Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,335 datasets
Vietnamese Handwriting Ocr is a dataset for optical character recognition tasks, hosted on the Hugging Face platform by user manhha2502. The dataset was last updated on December 16, 2025. Its specific contents, such as the number of samples or annotation format, are not detailed in the available metadata.
Intelligent Interaction Agent Dataset V0.1 is a large-scale, multi-modal dataset designed for building AI assistants. The dataset, created by deepgo and last updated on 2025-11-03, supports tasks like vehicle interaction recognition, multi-turn dialogue, and emotion-aware agent development.
Fantastic Beasts is a dataset collected for the NeurIPS 2023 paper 'AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation'. It was created by author chaofanma and last updated on Hugging Face in October 2025. The dataset is designed to address the lack of rare or obscure vocabulary in existing segmentation benchmarks.
753 substation images with polygon mask annotations across 18 object categories, plus three additional subsets for transmission line classification and detection. The collection features drone-captured imagery specifically targeting power system infrastructure for computer vision applications.
Infinity-Doc-55K contains 55,000 real-world and synthetic scanned documents for full-text parsing, published by infly in 2025. The collection features layout variations and structural annotations across financial, medical, and academic report domains.
GRAID BDD100K is a dataset of structured question-answer pairs generated from object detection annotations in driving scenes. The dataset was created by author kd7 using the GRAID framework and was last updated on Hugging Face in October 2025. It is designed to test aspects of object localization, visual reasoning, and spatial reasoning.
EgoExoBench is a benchmark designed to evaluate cross-perspective understanding in multimodal large models. It contains synchronized and asynchronous egocentric and exocentric video pairs with multiple-choice questions. The dataset was authored by Heleun and last updated on November 3, 2025.
A dataset of PDF documents annotated for OCR classification tasks, published by HuggingFaceFW. It contains binary labels (OCR/NOCR) and file size information for each PDF. The dataset was last updated on October 20, 2025.
A dataset of CT scans with dense segmentation annotations for 14 anatomical targets, including adrenal glands, colon, duodenum, esophagus, gallbladder, kidneys, liver, lungs, pancreas, small bowel, spleen, stomach, trachea, and bladder. The dataset is hosted on Hugging Face by author Angelou0516 and was last updated on October 30, 2025. Data is provided in the NIfTI (.nii.gz) format.
A protected microdata file from Statistics Netherlands (CBS) provides information on the relationship between individuals and the labor market. The survey covers persons aged 15 and older in the Netherlands, excluding the institutionalized population, and links personal characteristics to their current or future labor market position. This version contains revised data for 2003-2012, aligned with later survey years and includes new weighting variables to match official CBS tables.
A longitudinal study tracking two cohorts of children in the Netherlands to examine the effects of different forms of childcare and early childhood education. The pre-COOL study, conducted by the Kohnstamm Instituut and the University of Amsterdam, collects data on children's cognitive and socio-emotional development, family background, and the quality of preschool and kindergarten facilities. The dataset includes a four-year-old cohort started in 2009 and a two-year-old cohort started in 2010, with children followed until the end of primary school.
GRAID NuImages is a question-answer dataset generated by the GRAID framework for enhancing spatial reasoning in Vision-Language Models. The dataset was created by author kd7 and is associated with a research paper and project page. It was last updated on October 29, 2025.
1989 to 2016 panel survey of employer establishments in the Netherlands, conducted biennially. It contains around 3,000 observations per wave and 14 measurement waves, designed to provide insight into labor demand and personnel policy. The data is managed by the Netherlands Institute for Social Research (SCP) and originated from the Organization for Strategic Labor Market Research (OSA).
Pre-COOL is a longitudinal study tracking two cohorts of children to understand the effects of different forms of childcare and early childhood education. The study, conducted by the Kohnstamm Instituut and the University of Amsterdam, collects data on children's cognitive and socio-emotional development, their family backgrounds, and the quality of preschool and kindergarten facilities. The dataset includes a four-year-old cohort started in 2009 and a two-year-old cohort started in 2010, with data collection continuing through the end of primary school.
The TotalSegmentator Organs dataset contains CT scans with dense segmentation annotations for 14 anatomical structures. The dataset is provided by MedOtter and was last updated on October 30, 2025. The data format is NIfTI (.nii.gz).
Code-170k-shona is a dataset containing 176,999 programming conversations, originally sourced from glaiveai/glaive-code-assistant-v2 and translated into Shona. It was created by michsethowusu and last updated on October 30, 2025. The dataset aims to make coding education accessible to Shona speakers through multi-turn dialogues.
A subset of 100,000 anime character images from the Zerochan webdataset. The images are filtered to be non-monochrome and depict a single person, head, and face with one primary character. Annotator animetimm created this dataset, which was last updated on November 5, 2025.
OpenPecha's benchmark dataset evaluates Tibetan optical character recognition models. It includes diverse scripts, writing styles, and print methods to enable testing across multiple domains. The dataset was last updated on October 30, 2025.
GRAID NuImages was generated using the GRAID framework, which transforms object detection annotations into structured question-answer pairs. The dataset tests various aspects of object understanding and spatial reasoning. It was created by authors Charles Xu, Qiyang Li, Jianlan Luo, and Sergey Levine.
High-resolution images of handwritten mathematical notes in English, including problem statements, worked examples, formulas, and annotated derivations. The dataset was created by HumynLabs and was last updated on the platform on 2025-10-23.