Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
OCR VQA is a multimodal dataset for visual question answering tasks based on text extracted from images via optical character recognition. The dataset was created by qnguyen3 and uploaded to Hugging Face in October 2023. Specific details on the number of images, questions, or data volume are not provided in the available metadata.
License is unknown, which must be verified before use. The dataset's specific structure, file formats, and column schema are not described.