Loading...
Loading...
Image classification, object detection, segmentation, face recognition, OCR, image generation, video understanding
17,479 datasets
ResearchNed's Startmonitor project, conducted annually since 2008-2009, tracks the information use, study choice process, and first-year integration of new students in Dutch higher education. The survey for the 2022-2023 academic year includes two waves of questionnaires administered in September and June. It aims to identify determinants of study success and dropout in the first year.
More than 7,500 Dutch secondary school students aged 12 to 18 participated in this survey between October and December 2008. The Nibud-Scholierenonderzoek 2008-2009 collected data on student income, expenditures, borrowing, saving, and problems with money management. The dataset was made representative through weighting based on October 2008 education statistics from CFI, adjusting for province, gender, age, and school type.
Dutch survey data from 1993 to 1996, collected by the Central Bureau of Statistics (CBS). It aims to provide a complete overview of victimization of common crimes, feelings of insecurity, precautionary measures, and the use of legal aid and police assistance. The target population is individuals aged 15 and older in private households in the Netherlands.
The Startmonitor is a research project by the Dutch Ministry of Education, Culture and Science (OCW) conducted annually since the 2008-2009 academic year. This dataset contains the September 2018 survey results, aiming to map the study choice process and first-year integration of new higher education students. The project seeks to identify determinants of study success and dropout in the first year.
414 organizations, including 163 companies, 167 non-profit institutions, and 84 societal organizations, participated in this 2020 survey. The Social and Cultural Planning Office of the Netherlands conducts this biennial survey to measure the share of women in decision-making positions in the market and civil society. The dataset uses longitudinal variable naming to facilitate comparisons with previous measurements dating back to 2000.
Since 2000, the Netherlands Institute for Social Research has conducted a biennial survey measuring the share of women in decision-making positions in the market and civil society. The 2018 survey includes responses from 414 organizations: 145 companies, 183 non-profit institutions, and 86 societal organizations. This data is part of the Emancipation Monitor, which aims to provide a cross-section of the progress of the emancipation process.
A YOLO-format dataset for object detection tasks. The dataset contains 16 classes representing pool balls numbered 0 through 15. It was created by author nhantu2107 and was last updated on Hugging Face on October 23, 2025.
176,999 programming conversations translated into Oromo from the Glaive Code Assistant v2 dataset. These multi-turn dialogues cover technical topics such as algorithms and data structures to support coding education for Oromo speakers.
1968 onward, this dataset contains Oregon workers' compensation claims counts and insurer performance metrics. It is provided by the Oregon Department of Consumer and Business Services and includes annual data on accepted and denied claims, fatality rates, and processing timeliness. The data supports the official state report on workers' compensation.
Mexico City crime victim data extracted from the open data portal of the Mexico City Government. The dataset was adapted for use in Tableau with OpenRefine and is used for a course on data visualization principles. The data was last updated on the platform on October 14,我们发现2025.
3,000 images and fine-grained alpha mattes serve as the foundation for zero-shot image matting benchmarks. These samples enable the development of models that require higher precision than standard segmentation masks for complex edges like hair or transparency.
ILSVRC 2012 (ImageNet 1K) contains over 1.2 million images categorized into 1,000 distinct object classes based on the WordNet hierarchy. Created by the ILSVRC team and released in 2012, it serves as the foundational benchmark for large-scale visual recognition and object classification.
French Cold War documents collected by Marc Trachtenberg for historical research. The collection covers the period between 1945 and 1959 and is organized into three folders for five-year intervals. It was harvested by QDR and last updated on October 20, 2025.
Marc Trachtenberg's collection of British Cold War documents supports his historical analysis, notably in 'A Constructed Peace'. The data covers the pivotal period from the end of WWII to 1964, organized into folders for five-year intervals and individual years. This particular project encompasses documents from the United Kingdom.
A cleaned version of the Aozora Bunko corpus processed with fugashi morphological analysis and the unidic-lite dictionary. Each morpheme was converted into hiragana, and all non-hiragana symbols were removed. The dataset was created by milano0017 and last updated on November 10, 2025.
2018 data on members of Dutch Provincial and Executive Councils, used for a 2019 investigative article on conflicts of interest. The dataset likely contains demographic details, education, occupations, side functions, contacts with civil society, political career, affiliation, and policy portfolios. It was authored by D.J. Berkhout and archived by DANS Data Station Social Sciences and Humanities.
MTADataset is a large-scale dataset designed for image inpainting. It contains images processed with Grounded-SAM to extract labels, bounding boxes, and masks, and uses LLaVA to generate detailed descriptions for approximately 5 masks per image. The dataset was created by huangjun12 and was last updated on October 23, 2025.
A collection of scattering-type scanning near-field optical microscopy (s-SNOM) images used to train a machine-learning denoising model. The data was generated by Carlos Baiz and last updated on October 15, 2025. The dataset likely contains paired or unpaired sets of images captured at rapid and extended acquisition times.
This harmonized dataset contains pseudonymized data on antibody responses, infections, and vaccination status from multiple research groups within the Harmony Consortium, focusing on healthy volunteers and immunocompromised patients. It includes measurements of anti-S1, anti-N, and anti-RBD antibody titers, with limited T-cell response data from a small patient group. The dataset is accompanied by a data management plan, codebook, and harmonization templates.
The Marine Corps Organizational Culture Research (MCOCR) Project gathered Marine perspectives on culture, gender bias, leadership, and cohesion between 2017 and 2020. Data consist of 179 transcripts from semi-structured interviews and focus groups conducted at six locations in 2017. The project was authored by Kerry Fosher and harvested by QDR.