Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,077 datasets
A 41.2 KB spreadsheet contains the complete raw data collected during in vitro and in vivo experiments, including all measurements and experimental conditions. The file also holds the detailed results of all statistical tests performed, including ANOVA and post-hoc tests. Author Zhonghua Lv published the dataset under a CC-BY-4.0 license on figshare, with a last update timestamp of 2026-05-14.
Jawad Ali's dataset details the development of a saponin and salt-activated nuclease method for depleting host DNA in blood cultures. The 11.5 KB Excel file contains experimental results for optimizing conditions to improve bacterial DNA recovery for sepsis diagnosis. The data was uploaded to figshare in April 2026.
A research paper details the development and optimization of a method for host DNA depletion in blood cultures to improve sepsis diagnosis via nanopore sequencing. Author Jawad Ali published the study in April 2026. The file is a 415.7 KB PDF describing experimental results with spiked and clinical blood cultures.
A 2026 study by Martina Colombo details the development and evaluation of a three-dimensional alginate-based vitrification protocol for domestic cat cumulus-oocyte complexes. The 1.5 MB document compares this 3D method against a standard 2D vitrification approach, reporting metrics on cryoprotectant permeation, oocyte viability, maturation rates, actin distribution, and embryo development. It provides a proof-of-concept for optimizing fertility preservation techniques in feline species.
S-NPP CrIS IMG data from GES DISC contain collocated Visible Infrared Imaging Radiometer Suite (VIIRS) statistics within each Cross-track Infrared Sounder (CrIS) footprint. The dataset provides Level 1B radiance measurements from CrIS across 1,317 spectral channels and summary statistics from VIIRS's 22 bands, including cloud mask data. Products are constructed on six-minute boundaries with a spatial sampling of 30 field-of-regards cross-track and 45 along-track.
Over a hundred parameters per second are recorded in this global altimeter dataset from the TOPEX/POSEIDON mission launched in August 1992. The data provides sea surface height measurements with a precision of 3 cm and accuracy of 13 cm, alongside significant wave height, ionospheric corrections, and tides. Files are arranged in 10-day cycles separated into 254 passes, each approximately 56 minutes long.
Daily products on a ¼ x ¼ degree grid covering the continental United States (CONUS). This dataset provides fused estimates of near-surface vapor pressure deficit, generated by the Spatial Statistical Data Fusion (SSDF) algorithm combining data from the AIRS instrument on Aqua and the CrIMSS suite on Suomi-NPP. The algorithm weights input data based on estimated variance to infer a value for each grid point.
Spatial Statistical Data Fusion (SSDF) combines infrared and microwave sounder data from NASA's Aqua and Suomi-NPP satellites to produce daily, bias-corrected surface air temperature estimates. The algorithm weights measurements from the AIRS and CrIMSS instruments based on their estimated variance to infer values on a consistent ¼-degree grid. This Level-3 product provides a spatially fused, daily record for climate and atmospheric research over the continental United States.
Germany's statistical units derived from the digital landscape model at a 1:250,000 scale. The service is published by the Bundesamt für Kartographie und Geodäsie and is mapped via EuroBoundaryMap to satisfy INSPIRE conformance. It is provided free of charge under the federal GeoNutzV ordinance.
SLM-Bench is a programmatic benchmark containing 3,000 questions designed to evaluate language models with fewer than 10 million parameters. It covers six core capability areas with 500 questions each, including arithmetic problems. The dataset was created by liodon-ai and was last updated on June 14, 2026.
Data for optimizing coordinated trajectories of dual-arm robotic systems. The dataset was contributed by Qi Wang and is available via the Papers with Code platform under an Open Access license. Specific details on the data's size, format, and collection date are not provided.
Phylogenetic trees generated from a concatenated alignment, including a RAxML best tree file with bootstrap values and a Bayesian consensus tree. The dataset was contributed by author Rebecca B. Dikow and is available via the paperswithcode platform under an Open Access (green) license. Specific details on the number of taxa, alignment length, and creation date are not provided in the input.
Joris Meurs authored a paper titled 'Flow rates in liquid chromatography, gas chromatography and supercritical fluid chromatography: A tool for optimization'. The dataset likely contains flow rate data used for optimizing chromatographic separation methods. The paper is available via Open Access (green).
A historical dataset of repeated leg measurements from 16 patients, resulting in 256 total observations, used to model arterial occlusive diseases. The data was analyzed in a study by Endris Assen Ebrahim, comparing Bayesian MCMC methods, with results last updated in April 2026. The dataset is provided in an XLS format under a CC-BY-4.0 license.
A historical dataset from a study proposing a Bayesian hierarchical modeling framework for mixed-type outcomes. It contains 256 repeated leg measurements from 16 patients with arterial occlusive diseases. The dataset was created by Endris Assen Ebrahim and last updated on 2026-04-15.
A table of Wilcoxon rank-sum test results comparing evolutionary conservation and protein complex node scores between high- and low-centrality gene groups. The dataset, authored by Takanori Sasaki, is a 10.8 KB XLSX file last updated on 2026-05 06. It contains results for tests on 32 genes per centrality group.
Data from a 2015 study by Matthew Inglis and Andrew Aberdein published in *Philosophia Mathematica*. The dataset likely contains mathematicians' evaluations or ratings of different mathematical proofs, as described in the paper 'Beauty is not simplicity: An analysis of mathematicians' proof appraisals.'
Statistical analyses of the likelihood of the extended doubleton motif MxxxxxK arising by chance in three short amino acid sequences. The dataset is associated with a project by author John E Hart and is available via the paperswithcode platform under an Open Access (green) license. The specific temporal coverage, data volume, and file formats are not detailed.
Jackson Penfield's dataset contains molecular dynamics simulation trajectories for Human β defensin type 3 (hBD-3) monomers and dimers. The data comprises 57.0 microseconds of all-atom simulations across wild-type and analog protein forms embedded in four types of model lipid membranes. The dataset was last updated on April 30, 2026.
Kinematic data derived from high-speed video recordings of rats reaching for pellets. The dataset includes raw input data, processed output data, and statistical values and tables, created by Andrew Spence and last updated in May 2026. It was collected to study descending brain circuit remodeling after injury to inform better treatments.