Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,713 datasets
1,500 images of litter across 60 waste categories captured in diverse outdoor environments. The collection includes manual annotations for trash found in 'the wild' such as beaches, streets, and parks to facilitate environmental computer vision research.
Manuscript notes by Ramón Menéndez Pidal from the archive of the History of the Spanish Language, organized by author Inés Fernández-Ordóñez. The dataset covers linguistic evolution from Punic and Phoenician elements through Romanization, Vulgar and Literary Latin, to the Germanic element and Visigothic period (414-711). It was last updated on the Dataverse platform in May 2024.
Spanish manuscript notes from the archive of Ramón Menéndez Pidal's History of the Spanish Language. The dataset provides access to handwritten slips from the section titled 'Protohistory - Romanization', organized into sections covering Mediterranean, Iberian, Pyrenean, Celtic, and other pre-Roman peoples and languages. It was authored by Inés Fernández-Ordóñez and last updated on May 5, 2024.
Francisco Sánchez, Marisa Domínguez y Miguel Ángel (As Ellas / Eljas). Los juegos is a collection of descriptions of diverse traditional children's games. The dataset, coordinated by Álvarez Pérez, Xosé Afonso and harvested by e-cienciaDatos, includes information on game denominations, places, seasons, and gender divisions, alongside reflections on modern technology and school memories. It was last updated on May 5, 2024.
An interview with an informant in Malpica do Tejo, Portugal, detailing historical smuggling operations across the Portugal-Spain border. The description suggests the data likely contains narratives about products traded, collective evasion tactics, and contemporary border movement. The dataset was coordinated by Álvarez Pérez, Xosé Afonso and last updated on May 5, 2024.
A multilingual oral history session recorded in a bar captures discussions on domestic animals, roe deer, cattle, dairy, and traditional slaughter practices in the village of La Alamedilla. The recording is in Spanish and Portuguese, with sporadic interventions from a bar patron, documenting a cultural practice now maintained by only a couple of families. Xosé Afonso Álvarez Pérez coordinated this session, which was last updated in the dataverse on May 5, 2024.
A collection of ethnographic notes by Amparo López from La Alamedilla, harvested by e-cienciaDatos on 2024-05-05. The description suggests the data covers traditional knowledge and practices related to various livestock, including cattle, goats, sheep, pigs, and working animals, as well as associated activities like slaughter and sausage preparation.
Ramón Menéndez Pidal's handwritten notes from the archive of the History of the Spanish Language, covering the period 1380-1474. The dataset, authored by Inés Fernández-Ordóñez and last updated in May 2024, provides access to the original manuscript slips organized into sections on three generational stages, language evolution, and other Hispanic varieties.
Ramón Menéndez Pidal's handwritten notes from the archive of the History of the Spanish Language, covering the period from 1474 to 1555. The dataset, curated by Inés Fernández-Ordóñez, provides access to the contents of archiver 1, drawer 6, titled 'El español áureo. Renacimiento humanístico'. Its organization differentiates sections on Humanism, the triumph of Italianism, language evolution, other Hispanic languages, and international relations.
An oral history interview transcript with Manuel Barreira from Herrera de Alcántara, Spain, describing traditional agricultural and livestock practices. The description mentions topics including small-scale gardening, cereal harvesting, the absence of machinery, and the production of flour, bread, sausages, hams, and dairy products. The dataset was coordinated by Álvarez Pérez, Xosé Afonso and last updated on May 5, 2024.
A questionnaire response from an informant in San Martín de Trevellu/Trevejo covering various semantic fields. The data likely contains vocabulary and verbal conjugations related to livestock, insects, terrain, household items, and kinship. The dataset was coordinated by Álvarez Pérez, Xosé Afonso and last updated on May 5, 2024.
Gloria is an oral history interview from the e-cienciaDatos Harvested Dataverse, coordinated by Álvarez Pérez, Xosé Afonso. The dataset contains a first-person narrative from an informant who experienced a civil war at a young age in Santo Domingo de Guzmán. The account focuses on episodes of the war and, more prominently, the subsequent economic repression involving a Tax Prosecutor's Office and strict controls on cattle possession and slaughter.
Archival materials from the 'Historia de la Lengua Española' project by Ramón Menéndez Pidal, curated by Inés Fernández-Ordóñez. The dataset likely contains textual and philological records organized into four historical periods from the 8th to the 13th century. It was last updated in May 2024 via the e-cienciaDatos Harvested Dataverse.
1730-1823 handwritten notes by Ramón Menéndez Pidal from the archive of the History of the Spanish Language, organized into sections on criticism, neoclassicism, and preromanticism. The dataset provides access to these manuscript slips from the 'El español moderno. Renovación neoclásica' folder. It was contributed by Fernández-Ordóñez, Inés and last updated in May 2024.
A dataset coordinated by Álvarez Pérez, Xosé Afonso, harvested from e-cienciaDatos, describes changes in livestock populations and their uses over the years. It includes denominations for animals based on sex and age. The record was last updated on May 5,我们发现2024.
Jacinta González y Marcelina Preciado (Herrera de Alcántara). La vida de antes. El ganado is an oral history dataset from the e-cienciaDatos Harvested Dataverse. It contains testimonies about harsh living conditions, day labor in fields and houses, and the role of livestock like cattle, sheep, goats, and pigs for milk, tilling, and transport. The dataset was coordinated by Álvarez Pérez, Xosé Afonso and last updated on May 5, 2024.
2 geographic categories of air pollution images covering India and Nepal. These images document varying levels of atmospheric haze and smog across these two regions.
FEMNIST is a dataset for image classification of handwritten digits, lowercase, and uppercase letters, providing 62 unique labels. Each sample is a 28x28 grayscale image and includes writer and character information. The dataset is part of the LEAF benchmark and was curated by flwrlabs, with a last recorded update in April 2024.
ParsynthOCR 200K is a synthetic dataset for Persian optical character recognition. It is a preview version of the original 4 million sample dataset, ParsynthOCR-4M. The dataset was uploaded by hezarai and last updated on May 7, 2024.
Local Hotel Occupancy Tax (HOT) data has been compiled by Texas municipalities since 2018 and counties since January 2021. The data is self-reported by local governments to the Texas Comptroller of Public Accounts and includes revenue, tax rates, and allocation percentages for various uses. Specific questions about the data should be directed to the reporting local government entity.