Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
400 million diverse image-caption pairs collected from the open web support multimodal AI research. The dataset includes rich metadata for filtering and deduplication. It was created by Ajax102 and last updated on Hugging Face in January 2026.
License is unknown; terms of use must be verified before application.