Loading...
Loading...
DNA/RNA sequences, gene expression, protein structures, metagenomics, single-cell sequencing
27,749 datasets
Giving access to numerical accessibility scores (0-100) for crown-owned and lease-purchased buildings assessed against the Treasury Board Standard on Barrier-Free Access to Real Property. The data was collected by Public Services and Procurement Canada between 2019-2020 and 2023-2024 as part of a five-year assessment project. Row and column counts are unknown.
Ocean Acquisition System for Interdisciplinary Science (OASIS) moorings collected measurements at the mouth of Monterey Bay in 1995. The dataset includes parameters related to ocean chemistry, optics, temperature, and salinity. It is managed by the OB_DAAC organization and is available on multiple platforms.
Peruvian women aged 18-44 years participated in a study examining psychosocial mechanisms. The dataset contains results from 251 survey respondents, analyzed using structural equation modeling (SEM) by Velia Graciela Vera-Calmet. It was last updated on figshare in March 2026.
A 100-row sample of Harmonized System headings and subheadings from the U.S. Harmonized Tariff Schedule. The dataset likely contains standardized codes used to classify traded goods internationally. The sample's origin and update date are unknown.
A cleaned dataset of anime TV series titles includes genres, ratings, popularity, and other metadata. The dataset appears to be sourced from The Movie Database (TMDB) platform. The specific number of rows, file formats, and last update date are unknown.
gss1147's Melodyne God Producer Dataset contains 13,122 examples for training large language models. The dataset aims to teach models to master Celemony Melodyne software across its Assistant, Editor, Studio, and Essential editions. It was last updated on HuggingFace on 2026-04-23.
Global Affairs Canada's International Scholarships Program (ISP) funds, manages, and promotes scholarship opportunities for students and researchers. The dataset is published under the OGL-CA-2.0 license and was last updated on April 9, 2026. It likely contains records related to scholarship administration, funding, and participant information.
A dataset titled 'Small Yingshi' was published on the Hugging Face platform by the author 'quejing'. The dataset was last updated on June 2, 2026. Its specific content and scale are not described in the available metadata.
The companion data release for The Platonic Universe: Do Foundation Models See the Same Sky? (UniverseTBD et al. 2025, arXiv:2509.19453). This dataset contains image embeddings from foundation models applied to astronomical survey data, testing the Platonic Representation Hypothesis. It was authored by UniverseTBD and last updated on 2026-04-20.
Five current meters on a single mooring recorded oceanographic parameters near Heard Island in the Southern Ocean. Data collection spanned from May 1990 to January 1991. The processed data is archived by the CMR Data Centre in Hobart.
NASA OSDR studies OSD-969 and OSD-970 on brown and white adipose tissue. The data likely contains RNA expression profiles from these tissues exposed to spaceflight conditions. The dataset's specific scale, collection dates, and detailed methodology are not provided in the input.
CTD and STD instrument profiles provide temperature, salinity, density, and other oceanographic measurements from the Pacific Ocean. Data were collected from multiple vessels, including the GYRE, as part of the International Decade of Ocean Exploration North Pacific Experiment (IDOE/NORPAX) project. The dataset was submitted by Scripps Institution of Oceanography and processed by the National Oceanographic Data Center into the high-resolution F022 standard format.
Interpretations from U-Pb detrital zircon dating of offshore petroleum well cuttings provide new information on sediment origin and changes in provenance for the Roebuck Basin. The data includes detrital zircon age spectra and grain shape analyses, aiming to better understand reservoir quality in Triassic delta sequences. This dataset was published by Geoscience Australia Data and last updated on 2026-03-25.
October 10 to November 3, 1979 temperature-depth profiles collected via expendable bathythermograph (XBT) casts from the R/V ATLANTIS 2 in the Atlantic Ocean. Data were submitted by the Woods Hole Oceanographic Institution for the IDOE/POLYMODE project and processed by the National Oceanographic Data Center into the Universal Bathythermograph Output (UBT) format. Each profile consists of paired temperature and depth values recorded at inflection points to define the temperature curve.
Statutory land valuation data from Queensland, Australia, showing annual changes in property values. The dataset is provided by the Queensland Department of Natural Resources and Mines, Manufacturing and Regional and Rural Development and was last updated in March 2026. It includes notes on local government area de-amalgamations, such as Douglas from Cairns, occurring from 2014.
Additional file 1 from a study published on figshare provides statistics from genome-wide bisulfite sequencing data. The 11.7 KB XLSX file was authored by Jenny Riekötter and was last updated in April 2026. It contains processed data supporting research into the role of DNA methylation in regulating Chinese yam tuber shape.
A thematic journal issue introduces geological research on the outer North West Shelf of Australia. The issue is hosted by the Australian Ocean Data Network and was last updated in April 2026. The content is available in HTML and PDF formats.
Hurdeal et al. describe a new rozellid parasite, Rozellomyces, isolated from the Gulf of Alaska. The dataset includes phylogenomic analysis based on 238 genes and amplicon sequence variants (ASVs) from the metaPR2 database to map its geographic distribution. This research provides insights into the diversity and ecological role of Rozellomycota parasites infecting the diatom Thalassiosira.
A dataset likely containing the lengths of hit songs from 1959 to 2025. The data was gathered via a web crawling process, as suggested by the description. The specific author, organization, and exact number of records are unknown.
119 pairs of linguistic examples from Old Norse-Icelandic, each pair containing a causative construction and its corresponding anticausative. Each example consists of three lines: the example, glossing, and translation. The dataset was compiled by Jóhanna Barðdal and colleagues across three research projects funded by the Norwegian Research Council, European Research Council, and Ghent University.