Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,713 datasets
1,000 object classes containing 1.3 million event-based recordings derived from the original ImageNet dataset. The data consists of asynchronous event streams captured by moving a 640x480 resolution event camera in front of a monitor displaying static images to simulate real-world motion.
London Borough of Camden's annual analysis of pay by gender, ethnic origin, and disability for the 2022-23 period. The report was produced by the Government Digital Service and last updated on November 1, 2023. This data collection follows legislation requiring organizations to report gender pay information from April 2017.
A benchmark dataset proposed in a paper at Findings of EMNLP 2023, focused on the task of letting models learn from data that has inherent disagreement. The dataset was created by MichiganNLP and last updated on October 30, 2023. It provides two splits, including an annotation split where each annotator's data is divided into train and test sets.
19,590 high-quality portrait images of faces captured across large pose variations. LPFF is a dataset introduced for ICCV 2023 by Yiqian Wu, Jing Zhang, Hongbo Fu, and Xiaogang Jin. The dataset is designed to support the creation of realistic facial images and 3D face shapes using generative networks.
CLEAR-Global compiled a corpus of 10,281 randomized sentences from books written by Kanuri authors Dr. Baba Kura Alkali Gazali, Lawan Dalama, Kaka Gana Abba, and Lawan Hassan. The corpus contains 90,706 words and was released on Hugging Face in October 2023. It was compiled for the creation of open-source language technology.
A dataset of 152 images for object detection, created by DanielCerda and last updated on October 28, 2023. It contains labeled images of industrial process and instrumentation diagram (PID) objects. The data is split into 128 training, 12 validation, and 12 test images.
Madanbaduwal published this GitHub-hosted directory of computer vision datasets in late 2023. The repository organizes external data sources according to specific computer vision problem domains such as classification or detection. It serves as a curated index rather than a direct host for raw image files.
A text dataset uploaded by tylercross to Hugging Face on December 8, 2023. The title suggests it contains dialogues involving the philosopher Socrates, as written by Plato. The specific content, size, and structure are unknown from the available metadata.
Data.ct.gov provides aggregate data on abuse and neglect reports accepted for response by the Connecticut Department of Children and Families (DCF). The dataset includes counts and rates of various allegation types, such as physical abuse, sexual abuse, and educational neglect, aggregated by town, region, and state fiscal year. It reflects policy changes, including the introduction of a voluntary Family Assessment Response for low-risk reports starting in April 2012.
The Exclusively Dark (ExDARK) dataset contains images from 10 different low-light conditions, ranging from very low-light to twilight. It includes annotations for 12 object classes, such as Dog, People, and Car, at both the image and object bounding box levels. The dataset was uploaded to Hugging Face by SatwikKambham in October 2023.
MAFAND-MT is the largest machine translation benchmark for African languages in the news domain, covering 21 languages. The dataset was created by Masakhane and published in 2022. Train, validation, and test splits are available for 16 languages, with validation and test sets for an additional 5 languages.
Invoices And Receipts Ocr V1 is a dataset for optical character recognition tasks, uploaded to HuggingFace by author dajor85570. The dataset was last updated on October 28, 2023. The description and column-level metadata are limited, requiring further inspection of the actual data files.
This repository by Chebart provides a dataset and implementation for Russian word optical character recognition (OCR) using Transformer architectures. Updated in November 2023, the project utilizes PyTorch and HuggingFace Transformers to process Cyrillic text. The dataset focuses specifically on the recognition of individual Russian words rather than full documents.
200 images of the anime character Hoshino Ai, collected via an auto-crawling system from sites like Danbooru, Pixiv, and Zerochan. The dataset was created by the CyberHarem organization and includes multiple processed versions, such as cropped and aligned images. It was last updated on September 17, 2023.
Developed by fjxmlzn for the IMC 2020 conference, this framework generates synthetic networked time series data using Generative Adversarial Networks (GANs). It provides a methodology for creating privacy-preserving versions of sensitive network measurement data while maintaining temporal fidelity.
Datastorehouse is an open-source collaborative platform for dataset aggregation and sharing, developed by neokd and last updated in October 2023. The project functions as a centralized repository for discovering and contributing diverse data across multiple domains using CSV and JSON formats.
14 million scene text images organized into 11 subsets representing diverse real-world challenges like curved and artistic text. The collection serves as a massive-scale pre-training corpus for Scene Text Recognition (STR) models to improve performance on irregular text.
2015-2021 bike rack installations in Chicago, recorded by the City of Chicago. The dataset includes geographic and administrative location details for each rack. Data was last updated on the platform in July 2023.
Test Image Classification is a dataset published on the Hugging Face platform by the user 'searchfind'. The dataset was last updated on October 30, 2023. Its specific content, scale, and intended application require verification after download due to minimal descriptive metadata.
1,555 images featuring text across three primary orientations including horizontal, multi-oriented, and curved layouts. The collection provides specialized data for scene text detection and recognition tasks involving complex geometric arrangements.