Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
NewsEye and READ project training data comprises 200 Finnish newspaper page images from the 19th century, provided by the National Library Finland. The dataset includes carefully annotated text in PAGE XML format, produced using the Transkribus platform. It serves as a training set for historical document analysis and optical character recognition tasks.
License is listed as 'Open Access (green)', but specific terms are not detailed. The dataset is specifically a training set.