Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,661 datasets
Holosiivskyi District State Administration in Kyiv provides its official organizational chart. The data is structured according to the administrative document 'The structure and staffing of' and includes subordinate legal entities. It was last updated on 2025-04-29 via the eu_open_data platform.
1,555 images of rebar for object detection and instance segmentation, featuring fine-labeled bounding boxes and pixel-wise masks. The dataset includes diverse rebar specifications, layouts, application scenarios, and environmental conditions. It was uploaded by author 'tsrobcvai' to Hugging Face and last updated on 2025-05-05.
LeX-10K is a curated collection of 10,000 high-resolution images created by X-ART and released on Hugging Face in April 2025. The dataset is designed for text-to-image generation tasks, with a focus on aesthetic quality and text fidelity. It contains 1024x1024 images described as visually diverse and stylistically rich.
worm_data_short.parquet aggregates neural activity from 12 source datasets. The data was processed from raw formats including MATLAB and is associated with a 2024 conference paper. Author qsimeon uploaded it to Hugging Face on 2025-05-08.
Encompassing 144,810 images for fashion item detection, with 19,968 images of male subjects and 124,842 of female subjects. Fashion items are annotated with rectangular bounding boxes and categorized into four seasons: spring, autumn, summer, and winter. The data was created by Nexdata.
PH2D contains egocentric human-humanoid interaction data for co-training manipulation policies for humanoid robots. The dataset, created by RogerQi, was released in 2025 in conjunction with the associated research paper. Data is organized into folders by task and primarily stored in HDF5 files.
18,054 images of grape leaves across four classes: ESCA, Leaf Blight, Black Rot, and Healthy. The dataset was created by author adamkatchee as a modification of the original PlantVillage dataset. It was last updated on 2025-05-08.
3,700 high-resolution female portrait images at 1024×1024 pixels, hosted on Hugging Face by user strangerguardhf. The dataset is intended for training machine learning models on computer vision tasks. It was last updated on 2025-04-28.
Delivering ultra high-resolution image pairs categorized for image matting tasks. It supports the MEMatte framework for memory-efficient matting using adaptive token routing to process high-fidelity visual data.
NVIDIA's ClimbLab is a 1.2-trillion-token corpus for language model pre-training. It was created by OptimalScale using a semantic clustering method called CLIMB to reorganize and filter data from the Nemotron-CC and SmolLM-Corpus sources into 20 distinct clusters. The dataset was last updated on Hugging Face in May 2025.
Aggregating 18,880 images of 466 people, annotated with 3D instance segmentation masks and 22 anatomical landmarks. It was created by Nexdata and includes diversity in scenes, lighting, ages, shooting angles, and poses.
5,000 cleaned question-and-answer pairs focused on the C programming language sourced from StackOverflow. Each entry includes the original question and its corresponding accepted answer, with all text segments restricted to under 500 characters.
Cherkasy, Ukraine, provides data on seasonal trade placements within the city. The dataset includes information on trade types, organizing entities, work schedules, and permit details. It was last updated on 2025-04-30 and is published via the eu_open_data platform.
5 JSON files containing OCR-extracted Arabic text from book pages. Each entry maps a pageId to a source image url, a pdfPageNumber, and the corresponding ocrOutput string.
A collection of images sourced from CivitAI, licensed under CC BY-NC 4.0 for non-commercial use. The dataset is intended for research purposes, specifically for training NSFW classifiers and community enhancement measures. It was uploaded by author latentcanon and last updated on 2025-05-09.
A sample of 10,000 people for human pose recognition. It includes indoor and outdoor scenes, covering males and females with an age distribution from teenager to elderly, where middle-aged and young people are the majority.
ESA WorldCover 2020 v100 dataset provides a global 10-meter resolution land cover map. This dataset is derived from it for the task of weakly supervised semantic segmentation. The dataset was created by author j-h-f and last updated on the Hugging Face platform in April 2025.
GraspNet-1Billion provides 1 billion grasp poses for robotic manipulation research, maintained by the GraspNet team. The dataset serves as a large-scale benchmark for grasping tasks and was last updated in June 2025.
A May 1, 2025 update added a bounding box version of Web-Hybrid data, containing 757,000 datapoints after content moderation. The dataset is provided by the organization osunlp and hosted on Hugging Face. No conversation template is applied to this version, and bounding box coordinates are normalized to [0,999].
Sample of 5,199 3D face images from 5,199 people, collected in an indoor scene. It includes males and females with an age distribution from juvenile to elderly, primarily young and middle-aged people, captured using iPhone X and iPhone XR devices.