Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
This repository by Chebart provides a dataset and implementation for Russian word optical character recognition (OCR) using Transformer architectures. Updated in November 2023, the project utilizes PyTorch and HuggingFace Transformers to process Cyrillic text. The dataset focuses specifically on the recognition of individual Russian words rather than full documents.
Users should verify the license status on the GitHub repository before use; the implementation requires PyTorch and HuggingFace Transformers libraries.