Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A curated collection of 1,000 image and caption pairs. Each sample pairs an image with a detailed natural language description, making it suitable for training and evaluating vision-language models. The dataset was created by prithivMLmods and was last updated on July 4, 2026.
License is unknown, which may restrict usage.