Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,647 datasets
coco_synth_qwen3_next_0_1 is a synthetic image dataset containing 10010 images. The dataset was created by author testdummyvt using an automated pipeline to mirror the COCO80 object list, pairing each generated image with a short scene description and structured object entries. It was last updated on Hugging Face on 2025-09-14.
Exempt organization information is extracted monthly from the Internal Revenue Service’s Business Master File. The dataset provides records for Connecticut-based organizations, containing fields such as EIN, NAME, STATE, ASSET_AMT, INCOME_AMT, REVENUE_AMT, and NTEE_CD. It is updated nightly and maintained by the State of Connecticut.
This large-scale collection of multi-view video sequences focuses on the spatio-temporal localization of human-object interactions in retail settings. It provides synchronized camera perspectives to support the development of models for complex activity recognition and object manipulation tracking in real-world environments.
A combined dataset of approximately 25 million images from the ImageNet21K and CC12M collections, recaptioned by the author. The ImageNet21K portion contains about 13 million examples across roughly 19,000 classes, while the CC12M portion contains 12 million images created in 2021. The dataset was uploaded to Hugging Face by user 'gmongaras' in September 2025.
Pokrovsky City Council's executive committee education management organizational structure is documented. The dataset likely contains information about departments, roles, and staffing as defined by an organizational and administrative document. It was last updated on September 22, 2025, and originates from a Ukrainian government data portal.
Afri-Aya is a community-curated multilingual image dataset covering 13 major African languages with AI-powered categorization. It was created by the Cohere Labs Regional Africa community during the six-week Expedition Aya open-build challenge to include low-resource languages in vision-language tasks.
5,000 question-answer pairs demonstrating the Socratic method of teaching through guided questioning. The dataset was created by sanjaypantdsd and last updated on Hugging Face in September 2025. It has been cleaned to remove romantic and potentially inappropriate content.
A single photograph depicting the interior of an 'agadir', a traditional fortified granary or fortress in Morocco. This image is part of the 'Enhanced Tashelhiyt Dictionary' project, contributed by author John Telleman and hosted by the DANS Data Station Social Sciences and Humanities Collection. The record was last updated on October 24, 2025.
A photograph of an 'agadir', a masoned wall or fortress enclosing a town. This image is part of the 'Enhanced Tashelhiyt Dictionary: Argan' collection, contributed by author John Telleman and hosted by the DANS Data Station Social Sciences and Humanities Collection. The dataset record was last updated on October 24, 2025.
Grefcoco is a benchmark dataset for Generalized Referring Expression (GREx) tasks, including segmentation, comprehension, and generation, developed by researcher henghuiding and presented at CVPR 2023. It provides a standardized framework for evaluating models on their ability to process natural language queries that may refer to multiple objects or no objects at all within an image.
Tri3D is a unified interface for accessing and processing major 3D autonomous driving datasets, including KITTI, nuScenes, and the Waymo Open Dataset. Developed by CEA-LIST and last updated in November 2025, it provides a standardized framework for handling multi-modal sensor data and 3D geometry.
Federicogirella's Sketchy dataset, adapted from Fashionpedia, provides multi-conditioning data for image generation. It was published for the ICCV25 paper 'LOTS of Fashion! Multi-Conditioning for Image Generation via Sketch-Text Pairing'. The dataset was last updated on September 18, 2025.
Img2Dataset is a high-performance utility created by rom1504 for converting massive URL collections into structured image datasets. It can download, resize, and package 100 million URLs in 20 hours on a single machine, facilitating the creation of large-scale multimodal repositories.
Aggregating 87,749 images annotated with metadata from the Iconclass classification system. It serves as a test set for applying machine learning to cultural heritage collections.
WFS XPlanung BPL ‘Flat part B 6. Amendment and rearrangement’ is a Web Feature Service (WFS) dataset from the XPlanung 5.0 standard. It describes a development plan amendment and reorganization for the 'At the new town hall' area in the municipality of Meißenheim, Germany, with a priority use designated as GE. The dataset is provided by the Bundesamt für Kartographie und Geodäsie and was last updated on September 9, 2025.
The period 1975-1986 in Spain is represented in this dataset, which is oriented towards analyzing the digital representation of that era. The data was authored by EIROA, MATILDE and last updated on October 14, 2025.
Cristina García Yebra's dataset contains X-ray diffraction measurements related to a catalytic chemical reaction. The data, harvested from the e-cienciaDatos Dataverse, was last updated on October 14, 2025. It likely supports the study of a specific iridium-based catalyst for converting formic acid.
M Dolores Porto compiled 200 articles from Spanish digital newspapers in 2021. The dataset includes 100 articles each from El Mundo and El País, all containing the word 'polarización' (polarization). It was harvested via e-cienciaDatos and last updated on October 14, 2025.
Annual data from the Department of Commerce details businesses with no paid employees and $1,000+ in receipts. It provides counts and total receipts by industry, broken down by legal form of organization at state and national levels and by receipts-size class nationally. The dataset is updated annually, with the latest metadata from September 2025.
Nano-Banana contains 9,457 synthetic images generated by bitmind using the Google Gemini 2.5 Flash Image Preview model in August 2025. The data is organized into Parquet files with images stored in an optimized binary format for machine learning workflows.