Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,713 datasets
Research results from experiments analyzing the effect of silk fibroin and yerba mate nanoparticles on recycled polylactic acid (PLA). The dataset, authored by Beltrán González, Freddys R. and last updated in May 2024, contains measurements from migration tests in ethanolic food simulants and disintegration tests under composting conditions. Data is organized into Excel files with sheets for techniques including migration, mass loss, DSC, TGA, hardness, FTIR, and intrinsic viscosity.
Scene UNderstanding 397 is a dataset for scene categorization in computer vision, created to address the limited scope of earlier databases. The dataset, authored by 1aurent, was last updated on the Hugging Face platform in May 2024. It likely contains a variety of scene images intended to expand research beyond the 15-class limit of previous resources.
100 randomly selected classes from the ImageNet-1k collection are provided in this dataset, with all images pre-resized to 160 pixels on the shorter side. Each instance consists of an RGB image and its corresponding integer class label, derived from the original ILSVRC hierarchy.
A collection of Persian poems by Iran's great poets across historical and contemporary categories, converted from the Ganjoor database into CSV format. The data covers nearly the entire corpus of Persian poetry available in the Ganjoor repository for linguistic analysis.
A dataset of license plate images intended for optical character recognition (OCR) training. It was published on HuggingFace by AnirudhLanka2002 and last updated on June 26, 2024. The specific content, size, and annotation details require verification after download.
Francisco-Cruz provides 1,003 images of invoices and receipts, each with transcriptions for key financial fields. The dataset includes annotations for seller details, tax IDs, dates, and amounts. It was last updated on Hugging Face in May 2024.
EuroSAT is a dataset of satellite images for land use and land cover classification. It contains images labeled into 10 classes, including annual crop land, forest, and industrial buildings. The dataset was uploaded to Hugging Face by user 'tanganke' and was last updated on May 16, 2024.
SynthTIGER is a synthetic text image generator developed by Clova AI for the ICDAR 2021 conference. It enables the creation of large-scale scene text datasets for training and evaluating deep learning-based OCR recognition models.
An interview with an informant in Esperança provides Spanish and Portuguese designations for livestock and discusses differences between the two countries. The recording, coordinated by Álvarez Pérez, Xosé Afonso and harvested by e-cienciaDatos, was last updated on May 5, 2024. It covers traditional practices, recent introductions like sheep, and perceptions of declining meat quality.
Álvarez Pérez, Xosé Afonso coordinated this biographical dataset from the e-cienciaDatos Harvested Dataverse, last updated on May 5, 2024. The text describes the life of an informant from Tourém, covering childhood adversity, migration to Spain, and diverse economic activities. The narrative details work in the timber industry, smuggling, trade, and lending, including profiteering from high-cost products after the Spanish Civil War.
An oral history interview with Rosa Torrado describing life along the Spanish-Portuguese border in the villages of La Alamedilla and Batocas. The description details subsistence smuggling, uranium mining, cattle herding, depopulation, traditional festivals, and the political policing by the PIDE. The dataset was coordinated by Álvarez Pérez, Xosé Afonso and last updated in May 2024.
Families recounted stories over the years about relatives' actions during the Spanish Civil War, focusing on repression and forced head-shaving. The collection, coordinated by Álvarez Pérez, Xosé Afonso, was last updated on May 5, 2024. It also describes the post-war period where children often left school early for livestock care, domestic service, or emigration.
Álvarez Pérez, Xosé Afonso coordinated this oral history dataset documenting rural life in Herrera de Alcántara. The description details pastoralism, labor contracts, livestock breeds like the Merino sheep exported to Australia, and the socioeconomic structure of small landholdings versus large estates. The dataset was last updated on May 5, 2024, via the e-cienciaDatos Harvested Dataverse platform.
SATIN aggregates 27 satellite and aerial image datasets covering 6 distinct tasks. The imagery spans 5 orders of magnitude in resolution and contains over 250 distinct class labels. This collection was presented at the ICCV '23 TNGCV Workshop.
3 core components of the COCO data format—images, annotations, and categories—are programmatically generated by this suite of Python helper functions. The utility transforms binary masks and image metadata into the standardized JSON structure required for training object detection and instance segmentation models.
Geospatial boundaries define DSNY collection frequencies for refuse, recycling, organics, and bulk items across city sections. The dataset includes columns for SCHEDULECODE, FREQUENCY, and specific FREQ_* fields for each waste type. It is provided by data.cityofnewyork.us and was last updated in April 2024.
ChineseOCRBench is a dataset created to evaluate the performance of large multimodal models on Chinese Optical Character Recognition tasks. The dataset was extracted from the work 'On the Hidden Mystery of OCR in Large Multimodal Models' and published by author SWHL on Hugging Face in April 2024. It serves as a dedicated benchmark for Chinese OCR evaluation.
Álvarez Pérez, Xosé Afonso coordinated interviews with three informants in La Codosera. The dataset likely contains oral history recordings or transcripts covering livestock management, traditional farming, and food preparation practices. The data was harvested into the e-cienciaDatos Dataverse and last updated on May 5, —2024.
EasyPR features image data for Chinese license plate recognition in unconstrained environments, created by liuruoze for the 2017 China Graduate Contest on Smart-city Technology and Creative Design. The dataset supports supervised learning for identifying Chinese characters and alphanumeric sequences using OpenCV-based pipelines.
Spanish university reform project from 1967, known as PBRU. The dataset contains a tree diagram of the Ministry of Education and Science reorganization from January 1968 and results from a consultation with Spanish university professors, visualized in bar and pie charts. It was authored by María José Torres Parra and last updated on May 5, 2024.