Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A dataset containing paired images and captions designed for fine-grained multimodal concept understanding. Each data sample contains two images and two corresponding captions that differ only in one object, the color of an object, or the size of an object. The dataset was created by author 'phiyodr' and was last updated on October 2, 2024.
License is unknown; terms of use must be verified before application.