Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Ocrbench V2 is a multimodal dataset for evaluating optical character recognition systems, containing images paired with text. The dataset was created by author 'ling99' and uploaded to Hugging Face in February 2025. Platform tags indicate it contains at least 100,000 data points and includes both image and text modalities.
License is listed as MIT on the platform, but users should verify the specific terms. The dataset is multimodal (image+text), requiring tools capable of handling both data types.