Loading...
Loading...
DNA/RNA sequences, gene expression, protein structures, metagenomics, single-cell sequencing
27,286 datasets
A collection of supplementary tables from a genomics study, including a list of IMRGs (likely Immune-Related Genes) and results from Gene Ontology (GO), KEGG, and Gene Set Enrichment Analysis (GSEA). The dataset was authored by Liang Liu and last updated on April 15, 2026. It is a small archive, 43.8 KB in size, and shared under a CC-BY-4.0 license.
Australia's marine jurisdiction covers over 10 million square kilometres, with less than 25% of its seafloor mapped at high-resolution. The AusSeabed program, facilitated by Geoscience Australia, coordinates national seabed mapping efforts to reduce duplication and improve data consistency. It includes a government priority plan, survey register, and is developing cloud-based data sharing infrastructure and common mapping tools.
Polygon features defining parcels of land created on survey plans in New South Wales, Australia. The dataset visualizes parcel boundaries, identifiers, and basic topographic features, forming the foundation fabric of land ownership. Spatial Services continuously updates the data, sourced from subdivision, registration, gazettal activity, and multiple government agencies.
Draft genome sequences of Fungi isolated from the Mars 2020 Spacecraft assembly facility are reported. The fungal strains were isolated from samples collected from cleanroom surfaces of Kennedy Space Center-Payload Hazardous Servicing Facility and Jet Propulsion Laboratory-Spacecraft Assembly Facility. Whole genome sequencing (WGS) of these isolates was carried out by the National Aeronautics and Space Administration.
800 images of 20 mammal species, preprocessed to a uniform 128x128 pixel resolution. This micro version is derived from the Animals with Attributes 2 dataset, originally collected from public sources like Flickr in 2016 for transfer-learning benchmarks. The dataset was curated by Meta-Album for few-shot image classification tasks.
NASA GMAO Decadal Analysis & Prediction for CMIP5 is a climate modeling dataset contributed to the CMIP5 project. It contains a three-member ensemble of decadal predictions from the GEOS-5 AOGCM, initialized each December 1 from 1960 to 2010. The dataset is managed by NASA's NCCS and is available from the CMIP5 Archive.
Survey data from the ACT Government's Epidemiology Section, last updated in March 2026. The dataset is available in multiple machine-readable formats including CSV, JSON, and XML.
Sequences of primers and probes for a novel detection platform targeting Pseudomonas aeruginosa. The dataset, authored by Haotian Lin and last updated in March 2026, is stored in an XLS file of 5.5 KB. The described method achieved a sensitivity of 10 DNA copies per reaction and was validated on 20 water samples.
68 serum proteins showed statistically significant differences between Dupuytren Disease patients and a healthy control group. This 9.5 KB dataset contains proteomic biomarker profiles and collagen epitope markers from a pilot study comparing plasma samples. The data was authored by Blake Hummer and last updated on March 18, 2026.
3.2 GB of input files, trajectories, and analysis scripts supporting a study on ion transport in angstrom-scale graphene slits. The dataset, authored by Fan Feng and last updated in April 2026, likely contains simulation configurations, raw trajectory data, and processed results. It is shared under a CC-BY-4.0 license on figshare.
Information related to attendance of show cause hearings. The dataset is provided by the Family Responsibilities Commission under a CC-BY-4.0 license and was last updated in April 2026. The data is available in CSV and DOCX formats.
Experimental data from a study combining serial plating with long-read sequencing to detect and quantify live Shiga toxin-producing Escherichia coli (STEC) in ground beef. The method, developed by the Department of Agriculture, was able to quantify STEC down to 1 colony-forming unit per gram. The dataset was last updated on March 13, 2026.
Family Responsibilities Commission records related to community conferences and application hearings. The dataset is published by the Family Responsibilities Commission under a CC-BY-4.0 license and was last updated on 2026-04-15.
Australian data from the Family Responsibilities Commission detailing outcomes of community conferences and application hearings. The dataset is provided by the Family Responsibilities Commission under a CC-BY-4.0 license and was last updated on 2026-04-15.
Texas German is a set of varieties based on donor dialects brought to Texas from the mid-1800s through World War I. The Texas German Dialect Archive holds materials including this dataset of translation task lists used in interviews with speakers since 2001. These lists were created by the Texas German Dialect Project and harvested by the Texas Data Repository.
Immunoblots of Campylobacter jejuni BumR and BumR binding DNA by EMSA analysis. The dataset was authored by David Hendrixson and is hosted by the Texas Data Repository Harvested Dataverse. It was last updated on June 8, 2026.
NASA's MAST archive provides access to the PanSTARRS 1 Data Release 2 catalog via a Cone Search service. This catalog likely contains data on astronomical objects such as stars and galaxies. The service was last updated on March 13, 2026.
Supplementary material for a study on CONSTANS-LIKE genes in cultivated strawberry. The dataset, published by Lixia Sheng on figshare, is a 51.6 MB ZIP file last updated on 2026-05-05. Its specific contents likely relate to the genome-wide identification and functional analysis of genes FaCOL57 and FaCOL59.
3.1 GB of genomic sequencing data for the brown alga Dictyota coriacea, collected in La Jolla, CA, USA. The dataset includes unfiltered and filtered genome assemblies, a transcriptome, and genome annotations. It was authored by Hannah Bone and last updated on April 15, 2026.
Nada El Makhzen's 2026 study provides nanopore long-read sequencing data for the CFTR gene from 9 Moroccan individuals. The dataset identifies six specific genetic variants, including p.Phe508del and p.Arg1162*, with pathogenicity confirmed via cellular assays. It demonstrates a sequencing and analytical pipeline for variant detection and phasing in an understudied population.