Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
OCRBench v2 is a benchmark dataset for evaluating large multimodal models, containing at least 10,000 test samples. It includes subsets for English and Chinese, with the English subset containing 7,400 samples. The dataset focuses on visual text localization and reasoning tasks.
The dataset is loaded via the Hugging Face `datasets` library; users must specify the split ('test') and optionally the subset ('EN', 'CN'). The full description is on the Hugging Face dataset page.