Download the OCRd text for ALL the digitised journals in Trove!
This notebook helps you download all the OCRd text from all (or most of?) Trove's digitised periodicals, creating one text file for each issue. It also saves a CSV-formatted list of the issues in each periodical.
Preview
Using this notebook¶
To run this notebook using the ARDC Binder service you'll need to log in using an account from an Australian university or research organisation. If you don't have an account, try MyBinder instead.
The MyBinder service doesn't require any authentication, but it can be slow to start and will sometimes fail when busy. If you have a login at an Australian university, you'll probably get better results with ARDC Binder.
Binder is great for experimentation and quick tasks, but for some projects you might need a dedicated, persistent environment in which to work. There's information on other options in the run these notebooks section.
Related datasets¶
Additional documentation¶
Getting help¶
Cite as¶
Sherratt, Tim. (2024). GLAM-Workbench/trove-journals (version v2.2.0). Zenodo. https://doi.org/10.5281/zenodo.13744407