Loading...
Loading...
Medical imaging (X-ray, CT, MRI), electronic health records, clinical trials, ECG/EEG, pathology
14,812 datasets
Over 70,000 CT scan studies, including 20,000+ with medical professional protocols, are available for computer vision research. The dataset, created by UniDataPro, focuses on detecting and analyzing brain pathologies such as tumors, hemorrhages, and cancers. It was last updated on April 4, 2025.
Vietnamese student mental health counseling conversations translated into English. The dataset contains 10,000 randomly sampled rows from an original Vietnamese-language collection, processed using Google Translator. It was created by user 'arafatanam' and last updated on Hugging Face in April 2025.
20,000 rows of student mental health counseling conversations translated from Vietnamese to English using Google Translator. The dataset is derived from the original chillies/student-mental-health-counseling-vn dataset and was last updated on 2025-04-02 by arafatanam.
National estimates of adult oral health indicators for even years from 2012 through 2020, prepared from the Behavioral Risk Factor Surveillance System (BRFSS) public use data sets. The data includes median prevalence estimates across U.S. states and the District of Columbia, provided by the CDC's Division of Oral Health via data.cdc.gov. Estimates are not age-adjusted and may differ from other sources due to definitional or rounding differences.
Points representing hospital locations created for the DC Geographic Information System. The dataset is maintained by the D.C. Office of the Chief Technology Officer and participating agencies, with locations identified from public records and digitized from a snapbase. It was last updated on April 16, 2025.
Encompassing 100,000 records related to mental health factors and suicide tendency. It includes attributes such as Age and Depression Severity, designed for research in mental health and predictive modeling.
Mimic Cxr is a dataset of chest X-ray images hosted on HuggingFace. The dataset was uploaded by ayyuce and was last updated on April 28, 2025. The specific content, size, and structure of the dataset require verification after download.
8,069 3D CT scans and associated medical annotations curated from the BIMCV database for the MICCAI24 paper. The dataset includes approximately 8,000 image-text pairs, representing over 2 million 2D CT slices. It was uploaded by author cyd0806 to Hugging Face in March 2025.
Mimic Cxr likely contains chest X-ray images, a common resource for medical AI research. The dataset was authored by MLforHealthcare and last updated on Hugging Face in April 2025. Specific details on the number of images, patient demographics, and annotation types are not provided in the available metadata.
A list of medical equipment for communal healthcare institutions under the Magal Village Council, compiled according to fixed asset inventory descriptions. The dataset originates from the States site of Ukraine and was last updated on March 27, 2025. The specific number of items and data fields are unknown.
HuggingFace hosts the 'Medical R1 Gsm Style' dataset, authored by 'rishiraj' and last updated on 2025-04-18. The dataset's title suggests it contains medical text, likely formatted in a question-answering style. Its specific content, size, and structure require verification after download.
A text dataset related to medical consultations in Vietnamese, published on the Hugging Face platform by author HieuNguyen203. The dataset was last updated on April 15, 2025. Its specific content, scale, and structure require verification after download.
A dataset of laboratory investigations for 100,000 Indian subjects, containing over 6.8 million readings. The dataset, named NidaanKosha, is described as a treasury of diagnostic information. It was created by author 'ekacare' and last updated on the Hugging Face platform on 2025-03-16.
Cirrmri600Plus contains over 600 MRI scans and segmentation masks for cirrhotic liver analysis, developed by NUBagciLab and updated in May 2025. The collection features both T1-weighted and T2-weighted MRI sequences specifically for liver disease research.
Saarland outpatient care facilities are represented in this dataset from the Statistical Office of the State. The data describes institutions focused on healing, preserving, and promoting health, and includes geocoded address information. It is served via an OGC WFS interface and was last updated on March 10, 2025.
Composed of a collection of expert-crafted Q&A pairs structured for Chain of Thought reasoning across a spectrum of rare diseases and health conditions. The content spans diagnostic challenges, genetic origins, treatment options, and the societal impact on patients to facilitate complex medical reasoning in AI models.
Pangeanic Dictionary Of Medical Terms is a multilingual translation dataset. It comprises 30 bilingual TMX files, with 460 translation units each, covering multiple language combinations. The dataset was uploaded by FrancophonIA and last updated on March 30, 2025.
RoleMRC is a composite benchmark for evaluating large language models in role-playing and instruction-following scenarios. It focuses on maintaining role identity and ability limits while following diverse instructions.
CDC data.cdc.gov provides archived weekly counts of COVID-19 cases among healthcare personnel. The dataset includes weekly case counts, percentages of known healthcare worker status, and MMWR week identifiers. It was last updated in February 2025.
The Pathology Images of Scanners and Mobilephones (PLISM) dataset was created by Ochi et al. in 2024 to evaluate AI model robustness to inter-institutional domain shifts. All histopathological specimens were sourced from patients diagnosed and treated at the University of Tokyo Hospital between 1955 and 2018. PLISM-wsi consists of consecutive slides digitized under 7 different scanners.