Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
CC-OCR is a benchmark dataset for evaluating large multimodal models in optical character recognition and literacy tasks. The dataset is hosted in a TSV format for use with the VLMEvalKit evaluation framework. It was created by author wulipc and last updated in December 2024.
Users must refer to the linked GitHub repository for full documentation, evaluation code, and the benchmark leaderboard. The dataset is intended for evaluation within the VLMEvalKit.