Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,713 datasets
22,000 annotated scene-centric images partitioned into 20,000 training and 2,000 validation samples across 150 semantic categories. The data includes exhaustive annotations for both background elements like sky and road, as well as discrete objects like persons and beds.
Featuring 0.5 million synthetic Korean document images generated by the SynthDoG tool for training the Donut (OCR-Free Document Understanding Transformer) model. It was created by naver-clova-ix and last updated in January 2024.
32,203 images containing 393,703 labeled faces categorized into 61 distinct event classes. The dataset features significant variability in facial scale, pose, and occlusion levels across its training, validation, and testing splits.
CyberHarem's dataset contains 130 images of the character Catalina from Granblue Fantasy, each with associated tags. The images were auto-crawled from sites like Danbooru, Pixiv, and Zerochan by the DeepGHS Team. The dataset was last updated on January 21, -2024.
26,000 anime-style images curated for segmentation tasks, split evenly between foreground subjects and backgrounds. The dataset was created by Zarxrax through automated and manual inspection to improve upon a foundational source. It was last updated on January 28, 2024.
500 images of the character Shibuya Rin from the media franchise THE iDOLM@STER: Cinderella Girls, along with associated tags. The dataset was created by the CyberHarem organization and last updated on January 16, 2024. Images were automatically crawled from multiple online art platforms.
A dataset of pet images across 37 categories, with roughly 200 images per class. The images exhibit large variations in scale, pose, and lighting. This instance includes standard train/test splits and was last updated on the platform on 2024-01-07.
The dataset contains two resources with information about seasonal and agricultural fairs and their organizers. It is published by the State site of Ukraine as open data under the Law of Ukraine 'On Access to Public Information'. The data was last updated on January 5, 2024.
CyberHarem's dataset contains 28 images of the anime character minami_mother from the Love Live! series. The images were crawled from sites like Danbooru, Pixiv, and Zerochan using an auto-crawling system. The dataset was last updated on January 17, 2024.
500 images of the Pokémon character 'suiren_s_mother' were auto-crawled from sites like Danbooru, Pixiv, and Zerochan. The dataset was created by the DeepGHS Team under the CyberHarem organization and uploaded to Hugging Face on 2024-01-16. Core character tags, such as blue_hair and mature_female, were pruned from the final collection.
500 images of the anime character Shibuya Kanon from Love Live! Superstar!!, each with descriptive tags. The dataset was created by the DeepGHS Team within the CyberHarem organization on Hugging Face and was last updated on January 17, 2024. Images were automatically crawled from multiple fan art websites, including Danbooru, Pixiv, and Zerochan.
D Cube is a vision-language dataset for object detection and segmentation introduced in the NeurIPS 2023 paper 'Described Object Detection: Liberating Object Detection with Flexible Expressions'. Created by the shikras team, it provides labels characterized by intricate and flexible natural language expressions rather than fixed category names. The data supports multi-modal learning tasks where visual grounding is driven by complex descriptive text.
248 images of the anime character Hidaka Ai from THE iDOLM@STER franchise. The dataset includes tags for each image, with core character tags like brown_hair and short_hair pruned. Images were auto-crawled from sites including Danbooru, Pixiv, and ZeroChan by the DeepGHS Team and uploaded to Hugging Face by CyberHarem in January 2024.
Imagenet W21 Webp Wds is a copy of the full Winter21 release of ImageNet in WebDataset tar format with WEBP encoded images. The dataset consists of 19,167 classes, which is 2,674 fewer classes than the original Fall11 release. It was created by author 'timm' and last updated on Hugging Face on 2024-01-07.
An open-source dataset intended as an encyclopedia for anime and manga characters. The collection is hosted on Hugging Face by the user 'lowres' and was last updated on January 14, 2024. Its specific scale and file formats are not detailed in the available metadata.
CCPD is a large-scale image dataset for license plate detection and recognition, released by detectRecog for the ECCV 2018 conference. It provides annotations for plate bounding boxes and character sequences across diverse real-world parking environments.
1,044 scenes containing over 1 million image pairs labeled as either 'doppelgangers' (visually similar but distinct locations) or 'non-doppelgangers' (same physical location). The dataset provides 3D reconstructions and camera poses to serve as ground truth for disambiguating visually identical structures in computer vision tasks.
A subset of the ImageNet-21K (Winter21) dataset, filtered according to the ImageNet-21-P methodology. It contains 10,450 classes, split into training and validation sets. The dataset was processed by re-encoding images into the WEBP format.
Annotated screenshots of websites are useful for training models in Robotic Process Automation. The dataset is hosted on Hugging Face by author Zexanima and was last updated on December 31, 2023. According to the description, labeling such screenshots can be expensive, and this collection aims to provide a pre-labeled resource.
RefSegRS is a dataset for referring remote sensing image segmentation tasks, likely containing satellite imagery paired with descriptive text. It was created by JessicaYuan and associated researchers for the RRSIS project, as documented in a 2023 arXiv preprint. The dataset was last updated on HuggingFace in February 2024.