Loading...
Loading...
DNA/RNA sequences, gene expression, protein structures, metagenomics, single-cell sequencing
27,749 datasets
9.5 KB of experimental data from a study optimizing a constructed wetland-microbial fuel cell (CW-MFC) for rural domestic wastewater treatment under low-temperature winter conditions. The dataset, authored by Tuodi Zhang and last updated in May 2026, contains results from a response surface methodology (RSM) experiment varying electrode plate projection coefficient, inter-electrode distance, and external resistance to maximize COD removal efficiency.
Lars Brummel from Leiden University produced a three-part deliverable synthesizing academic literature on government responses to COVID-19. The work includes a systematic literature review, a conceptual framework for analyzing crisis governance legitimacy, and a proposed mixed-methods research design. The dataset likely contains the text and findings from this review and framework development.
The Southern Fairway Basin on the Lord Howe Rise in the Tasman Sea contains thick packages of Cretaceous and Tertiary sediment. Thirteen piston cores, collected by the RV L'Atalante in 1999 for the ZoNiCo 5 survey, were taken from depths between 1250 and 2753 meters below sea level to assess gas and petroleum potential. The cores document sediment, pore water, and gas composition in the shallow sedimentary section.
besti11's AI and AGI Governance Proposal outlines steps for AI and AGI systems before they act. The dataset, last updated on July 21, 2026, is hosted on Hugging Face and tagged with topics like policy and ethics. It likely contains textual proposals or frameworks related to AI governance and safety.
Yan Wang from Wannan Medical College created a dataset for investigating RNA methylation-related genes in cervical cancer. The data includes mRNA expression profiles, clinical data, and genes related to m6A, m5C, and m1A methylation modifications. Differential analysis identified 106 methylation-related differential genes, and a prognostic model was built using 10 genes validated with TCGA and GSE39001 data.
Evaluation data artefacts for experiments conducted using the TCtracer tool. The data supports the ICSE 2020 paper 'Establishing Multilevel Test-to-Code Traceability Links'. Robert White from University College London authored the paper and provided the data.
1997 and/or 2008 data on Linear Kinematic Features (LKFs) detected and tracked in sea-ice deformation fields simulated by models participating in the Sea Ice Rheology Experiment (SIREx). The dataset was created by Nils Hutter of the Alfred-Wegener-Institut Helmholtz-Zentrum für Polar- und Meeresforschung and serves as the basis for feature-based evaluation in a 2022 Journal of Geophysical Research: Oceans paper.
Each file contains a Maintenance of Wakefulness Test (MWT) trial recording from a patient, collected after noon. The data includes occipital EEG and EOG signals bandpass filtered between 0.5-45 Hz, along with expert scoring labels for wakefulness and microsleep episodes. The dataset is authored by Anneke Hertig-Godeschalk from the University of Bern and is based on the BERN scoring criteria.
Multiannual hourly ground temperature measurements from sixteen high elevation sites (3493–4377 m a.s.l.) in the Bale Mountains, Ethiopia, collected by Alexander Raphael Groos of the University of Bern. The dataset includes corrected and interpolated time series, original logfiles, and documentation of data modifications, covering a period from at least February 2017 to January 2020.
DinoREF is a curated reference database for the 18S rRNA gene of dinoflagellates (Dinophyceae). The database includes sequences from GenBank and literature sources, annotated with multiple taxonomy classifications and clustered into OTUs. Solenn Mordret and colleagues submitted the associated paper for revision in 2018.
The BiLI-TAS research project from Uppsala University explores the Turkish and Swedish language development of 102 bilingual children aged 4-7 growing up in Sweden. Findings are reported for vocabulary, inflectional morphology, and subordination, relating them to studies of Turkish-speaking children with other language combinations. The dataset likely contains structured assessments of language skills from these children.
Daisuke Tsugama's repository contains processed datasets and code for analyzing codon-mediated gene expression regulation. The 2.2 GB collection includes RNA-seq and Ribo-seq data, predicted expression indices, and observed TPM, mRNA half-life, and protein abundance for Arabidopsis thaliana, Oryza sativa, Homo sapiens, and Mus musculus. It was last updated on May 27, 2026.
Modelled dispersion data for three guided modes in a lithium niobate nanowaveguide structure, including an avoided mode crossing. The dataset also contains predicted modulation instability gain, XFROG spectrograms for soliton solutions, and simulations of soliton propagation. It was created by Will Rowe of the University of Bath.
Datasets and analyses for the CHI 2020 paper 'Affect Recognition using Psychophysiological Correlates in High Intensity VR Exergaming' by Soumya C. Barathi of the University of Bath. The data comprises two experiments investigating sensor-based affect recognition during different VR exergaming scenarios, including conventional exercise, sedentary VR gaming, and optimal/overwhelming game conditions. The release includes CSV data sheets, JASP files with statistical tests, and R scripts for correlation and regression analyses.
Simulations and analysis performed on two protein structures from the SARS-CoV-2 virus, PDB entries 6Y2E and 6LU7. The study includes rigidity analysis, elastic network modelling to identify normal modes, and all-atom geometric simulations of flexible motion. The data was produced by Stephen A. Wells at the University of Bath using tools like FIRST, FRODA, and Elnemo.
Rob W. Holland from Radboud University Nijmegen authored this dataset. It contains experimental data from three studies investigating how self-construal activation influences interpersonal proximity behavior. The studies measured seating distances in waiting rooms and dyadic settings based on primed or chronic self-construal.
2022 mutational signature data from a study published in Science, authored by Andrea Degasperi of the University of Cambridge. The data originates from whole-genome-sequenced cancers within the UK population, supporting research into cancer etiology.
A 1.05 Å resolution crystal structure describes a DNA oligonucleotide that self-associates into a non-G-quadruplex fold-back structure. Betty Chu from the University of Maryland, College Park authored the research, which reveals a tetrameric assembly formed by two-fold-back dimers interacting through noncanonical and Watson-Crick base pairs. This structure provides new sequence and structural contexts for fold-back quadruplexes and may inform the design of DNA nanoarchitectures or cation sensors.
Plamen Nikolov of Harvard University Press uses a laboratory binary choice minimum-effort coordination game to study gradualism. The experiment randomly assigned participants to three treatments to test how slowly increasing contribution thresholds affects coordination success. Findings suggest a simple, voluntary mechanism can promote coordination when sanction capacity is limited.
Supplementary data for a paper presented at the 17th International Web for All Conference (W4A'20) proposes a novel autism detection approach. The dataset contains individual eye-movement paths used to evaluate a Scanpath Trend Analysis (STA) method. It also includes Python code to re-run the evaluation, supporting reproducibility.