Loading...
Loading...
DNA/RNA sequences, gene expression, protein structures, metagenomics, single-cell sequencing
27,749 datasets
Brazilian research analyzes the editorial characteristics of top-tier (Qualis A1) medical journals across three subfields. The dataset, created by Luiz Roberto Curtinaz Schifini, profiles journals based on data from Sucupira, Ulrichsweb, DOAJ, Scimago Journal Rank, and Journal Citation Reports. It reports metrics including a median unified impact factor of 5.365, a median age of 45 years, and that 13% are Open Access.
England's Environment Agency provides an assessment of flood risk from rivers and the sea for areas linked to properties. The dataset includes Ordnance Survey references for linking to AddressBase and OS MasterMap. RoFRS outputs estimate flood risk to an area of land and are generally not suitable for property-level assessment.
MOSTWAS models and summary statistics for transcriptome-wide association studies are contained in this dataset. The data includes models trained on TCGA breast cancer and ROS/MAP pre-frontal cortex multi-omic data, simulation results, and TWAS associations for breast cancer survival, Alzheimer's disease, and major depressive disorder. Arjun Bhattacharya from the University of North Carolina at Chapel Hill created this dataset to accompany the 2020 paper.
Sedimentological and geochemical properties of seabed and suspended sediments in north and central Torres Strait, Australia. The dataset was created to investigate links between sediment delivery and widespread seagrass dieback. It was published by the Australian Ocean Data Network and last updated on 2026-06-16.
171.99 Mb genome assembly of the parasitic wasp Chouioia cunea, containing 6 chromosomes and 12,258 annotated protein-coding genes. The assembly, contributed by Ziqi Wang, includes annotations for repeat sequences, noncoding RNAs, and gene family evolution. Results show 26.06 Mb of repeat sequences and 431 predicted noncoding RNAs, including micro-RNAs and transfer RNAs.
Nanchukmacdon genome data includes a chromosome-level assembly built from PacBio and short-read sequencing. The assembly comprises 1,942 high-quality contigs and is annotated with 20,588 protein-coding genes averaging 47.06 Kbp in length, plus non-coding RNAs. The data was produced by Bioinfo lab and sequencing reads are available under NCBI SRA project PRJNA967127.
Leontodon longirostris genome draft is a 418 Mb assembly provided as FASTA sequences. It includes 853 MAKER-based gene models and 472 GeneAssembler-based chimaeric gene models, both with functional annotations. The dataset was contributed by M. Gonzalo Claros and is available via paperswithcode.
Alegria, Rio de Janeiro, Brazil is the study location. This dataset describes the technical, chemical, and biological characterization of biosolids from a sewage treatment plant and their use as a substrate component for producing seedlings of Anadenanthera macrocarpa. The study tested four substrate formulations with varying proportions of commercial substrate and biosolids, measuring seedling growth after fifty days.
A cross-sectional study of 39 patients with a median age of 63 years evaluated the prevalence of atherosclerotic lesions in the Left Internal Thoracic Artery (LITA) using selective preoperative angiography. The research, authored by Hadrien Felipe Meira Balzan, identified a 7.7% prevalence of disorders that made the LITA unsuitable for use as a graft in Coronary Artery Bypass Graft (CABG) surgery. Statistical analysis was performed using SPSS® software version 20.
Course materials and an online book created for a three-day spring school on high-throughput 16S rRNA gene sequencing data analysis. The materials were developed by Shetty Sudarshan A in collaboration with Wageningen University & Research and the University of Turku. The tutorial focuses on downstream analysis from OTU tables and BIOM files using R-based tools like Phyloseq and ggplot2.
1,153 users in Pernambuco State, Brazil, were interviewed about their participation in physical activity programs. The study found 40% of barriers and 77.5% of facilitators belonged to the intrapersonal domain, with 'current health condition' as the top barrier and desire 'to be healthier' as the top facilitator. This cross-sectional research by Caroline Silva profiles user demographics, health status, and program engagement factors.
Scopus publication affiliation data from 1996-2018 was used to infer internal migration among researchers across Mexican states. The dataset, created by Andrea Miranda-González and colleagues, represents each movement as a row from a source to a target state in a specific year. It is provided under a CC BY-NC-SA 4.0 license and can be used to model migration flows or as an edge-list for network analysis.
Carnarvon Basin covers over 1,000 km of Western Australia's coast and contains up to 15,000 m of sedimentary infill. This dataset provides descriptive attributes for groundwater features, grouped into themes like hydrogeology, groundwater management, and land use. It is published by the Australian Ocean Data Network via data.gov.au.
Australian Ocean Data Network provides environmental DNA (eDNA) data collected during the RV Investigator voyage IN2022_V09. The voyage, titled 'Valuing Australia’s new Gascoyne Marine Park,' took place between November 19 and December 19, 2022, departing from and returning to Fremantle. The study compared three eDNA processing methods and validated fish detections against trawl survey data and the regional species pool.
An archaeological dataset from the University of Basel lists sites where brooch and belt buckle types found in Basel during the Early Middle Ages are attested. The CSV file includes site coordinates, artifact types, and chronological frames, while accompanying TIFF files contain high-resolution distribution maps generated via Kernel Density Estimation. This data supplements a published article in ZAM - Zeitschrift für Archäologie des Mittelalters.
The International Seabed Geomorphology Mapping Working Group's Seabed Geomorphology Classifier (ISGM-SGC) is a standardized tool for classifying seabed features. It was developed by geoscience agencies from the United Kingdom, Norway, Ireland, and Australia. The tool assigns five levels of geomorphology classification attributes and 16 additional attributes for describing geomorphic interpretations.
Baer's pochard (Aythya baeri) is a critically endangered duck species historically widespread in East Asia. Lei Zhang reports the first high-quality genome assembly for this species, with a total length of 1.14 Gb. The assembly is anchored to 35 chromosomes and includes annotation for 18,581 protein-coding genes.
A study by Dawei Liu from Sun Yat-sen University analyzed 54 osteosarcoma tissues and matched nontumor tissues. It investigated the expression of the Prospero homeobox 1 (PROX1) gene and its correlation with clinical characteristics and patient survival.
Supplementary materials for a systematic review and meta-analysis on SARS-CoV-2 RNA prevalence in blood products. The underlying data includes metadata for 28 citations, RT-PCR results from 212 serum samples from a UK clinical cohort and 142 convalescent donor samples, and raw microscope images of cell cultures. The report was authored by Monique Andersson.
Goiás State data from a cross-sectional study of 53 medical records of pregnant and postpartum women who died at a reference hospital. The study correlates maternal changes and pregnancy outcomes, reporting a maternal mortality ratio of 228.4. The author is Maíra Ribeiro Gomes De Lima.