Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
TurkishLLaVA OCR Enhancement Dataset is a specialized collection of 100,000 books sourced entirely from Turkish materials. Created by ytu-ce-cosmos, it was last updated on December 17, 2024. The dataset is designed to improve the Turkish Optical Character Recognition capabilities of the Turkish-LLaVA-v0.1 model.
License is unknown, which may impose usage restrictions.