Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,713 datasets
Trail map URLs for all locations operated by the Olympic Regional Development Authority (ORDA) in New York State. The data is organized by venue, county, activity, and includes latitude and longitude coordinates. It was published by data.ny.gov and last updated on November 10, 2022.
CV is a collection of receipt images for benchmarking key information extraction from documents. The dataset provides OCR outputs including bounding boxes, recognized text, and class labels for each receipt. It was created by nvm472001 and last updated in December 2022.
LAION-Face is a subset of the LAION-400M dataset containing 50 million image-text pairs identified as containing human faces. It was first used to train the FaRL model for face analysis tasks. A 20 million image subset is also provided for faster evaluation.
A dataset of high-quality face images, likely derived from the CelebA dataset. The dataset is associated with the 2017 research paper 'Progressive Growing of GANs for Improved Quality, Stability, and Variation' by Karras et al. It was uploaded to Hugging Face by user Chris1 and last updated on November 18, 2022.
Imagewoof is a dataset for image classification containing a subset of 10 challenging dog breed classes from ImageNet. The breeds include Australian terrier, Border terrier, Samoyed, Beagle, Shih-Tzu, English foxhound, Rhodesian ridgeback, Dingo, Golden retriever, and Old English sheepdog. It was created by author frgfm and last updated in December 2022.
2021 boundaries for county subdivisions in Michigan, derived from the U.S. Census Bureau's MAF/TIGER Database. This dataset includes legally-recognized minor civil divisions (MCDs), statistical census county divisions (CCDs), and unorganized territories, providing a seamless geographic framework for the state.
1 dataset preparation guide for the YOLOv5 object detection framework. The content details the use of LabelImg for manual image annotation and label formatting.
Imagenette is a subset of the ImageNet dataset containing 10 easily classified image categories, including tench, English springer, and cassette player. It was created by author frgfm and last updated in December 2022.
OpenFire is an image classification dataset for wildfire detection, collected from web searches. The dataset is categorized as containing between 1,000 and 10,000 samples and was last updated in December 2022.
Encompassing approximately 2.4 million image pairs for improving VQGAN prediction quality. Each pair consists of a 512x512 image crop from Open Images and a corresponding 256x256 VQGAN-encoded-and-decoded version.
Aggregating between 100,000 and 1,000,000 non-anonymized news articles and summaries sourced from CNN and Daily Mail. Curated by ccdv and last updated in 2022, it provides paired text for training and evaluating abstractive summarization models.
341,968 audio recordings across Logical Access and Physical Access categories for the Third Automatic Speaker Verification Spoofing and Countermeasures Challenge. It includes genuine human speech and various spoofing attacks, such as synthetic speech generation and physical replay attempts, to support detection algorithm development.
The organizational and managerial structure of KP "Svyatoshinsky LPG" is described in this dataset. It is published as open data by the States site of Ukraine under the Law of Ukraine "On Access to Public Information", allowing free use and distribution. The dataset was last updated on October 6, 2022.
A best-fit lookup table mapping 2011 Lower Layer Super Output Areas (LSOAs) to 2020 Electoral Wards and Local Authority Districts in England and Wales. The dataset is provided by the Government Digital Service and was last updated on 2022-07-23. It includes fields for area codes and names, with a file size of 7 MB.
United Kingdom lookup file mapping electoral wards and divisions to their corresponding local authority districts as of 31 December 2020. The dataset includes codes and names for both administrative levels and was published by the Government Digital Service. It was last updated on the platform in July 2022.
Comprising approximately 36,000 unique pairs of protein sequences and ligand SMILES strings, along with the 3D coordinates of their complexes from the Protein Data Bank (PDB). Ligands are filtered to have at least 3 atoms, a molecular weight of 100 Da or more, and exclude the 280 most common PDB ligands. It was created by author jglaser and last updated in October 2022.
Five categories of text areas including "key", "value", "header", "other", and "background" are annotated across document images in this revised version of the FUNSD dataset. The data focuses on correcting connectivity inconsistencies between text areas to better support key-value extraction tasks.
Aggregating remote sensing satellite imagery categorized for small-object detection using an end-to-end edge-enhanced GAN. It features high-resolution image pairs processed through an integrated object detector network to identify minute geographical features.
A directory of the municipal institution "Maximum Lyceum" in Kropyvnytskyi, Ukraine, last updated on October 12, 2022. It contains three resources listing the organization, its subdivisions, and its employees. The data originates from the State site of Ukraine.
A version of the COCO dataset prepared for image captioning tasks using the Karpathy split. The dataset was created by the author yerevann and was last updated on Hugging Face in October 2022.