Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,069 datasets
DICOM data and FreeSurfer recon-all generated files for mathematical modeling of the human brain. The dataset was created by Kent‐André Mardal at the University of Oslo. The temporal coverage and size of the dataset are unknown.
Google AI Quantum and collaborators produced experimental quantum data from superconducting processor experiments. The data likely contains results from quantum approximate optimization runs on non-planar graph problems. The dataset is associated with an Open Access paper published on arXiv.
69.5 KB of summary statistics and statistical comparisons for all quantitative data from a specific manuscript. The data is provided by author Nicholas D. Christman in an XLSX file under a CC-BY-4.0 license and was last updated on June 4, 2026.
144.9 MB replication files analyze fragmented identity documentation among Syrian refugees in Lebanon. Latent class analysis identifies four prevalent patterns of document ownership, linking them to varied access to services and vulnerabilities. The dataset, authored by V.N. Tran and released under CC-BY-4.0, provides files for replicating this research.
A theoretical document proposing an alternative model to the electron cloud concept in atomic structure. The author, Richard Morefield, describes a layered tetrahedral model for electron arrangement and bonding points, with implications for proton acceleration in elements like lutetium and polonium. The document was last updated on May 8, 2026 and is shared under a CC-BY-4.0 license.
Ethan Davis published model performance scores, compute profiling metrics, and MCMC convergence diagnostics on 2026-05-05. The data originates from a large-scale benchmark comparing Bayesian and frequentist pipelines for motor imagery EEG classification across twenty publicly available MOABB datasets and six pipeline pairs. Performance was assessed using six metrics including AUROC and MCC, with compute data covering training times.
Data and code accompany a 2017 research paper by Tad Dallas et al. on competitive outcomes in experimental communities. The repository likely contains tabular data from controlled experiments examining how initial abundance and stochasticity influence species competition. The dataset is provided under an Open Access license to support reproducibility of the published analyses.
12.8 MB of data from a techno-economic optimization study comparing a hybrid CO2 capture process to standalone methods. The dataset, authored by Sunny Pawar and last updated in April 2026, explores flue gas CO2 compositions ranging from 3.5 to 30 mol %. It contains results from process models and surrogate artificial neural networks used to calculate CO2 avoided costs.
A simulation experiment using seabed mud content samples from the Geoscience Australian Marine Samples database to compare statistical and mathematical spatial interpolation techniques. The study assessed prediction accuracy using cross-validation and analyzed factors like region, sample density, and method. Outcomes can be applied to modeling physical properties for marine biodiversity prediction.
Supplementary Material 1 contains densitometric quantification data from Western blot analyses and cell viability assays for a study on the radiosensitizing effects of a dual-target inhibitor. The data includes measurements of protein levels like p-STAT3 and Ac-H3/H4, and cell viability results from CCK8 assays across A549, MDA-MB-231, and B16 cell lines. Qi Wang published this dataset on figshare in April 2026 under a CC-BY-4.0 license.
ProofWiki Math Problems and Solutions pairs theorem statements with their formal proofs extracted from the ProofWiki knowledge base. The dataset includes raw MediaWiki wikitext alongside resolved text for both problems and solutions. It was created by user 'avewright' and last updated on 2026-06-19.
A crossover randomized trial of 60 sedentary college students measured hemodynamic and perceptual responses to 5-minute walking sessions with varying limb occlusion pressures. Data includes blood pressure, heart rate, perceived exertion, discomfort, and step counts recorded before, immediately after, and 5 minutes post-intervention. The dataset was published by Yuke Zhu on figshare in April 2026.
A 932.5 KB research paper by Kazuki Tomioka, last updated on 2026-05-08, proposing an estimation framework for panel stochastic frontier models. The framework accommodates heterogeneity through latent group structures and is demonstrated with an empirical application to U.S. commercial banking cost efficiency. The paper includes simulation studies and is available in PDF and TXT formats under a CC-BY-4.0 license.
R code for analyzing long-term maternal mortality associated with placental abruption and retention. The analysis is based on a cohort of 638,911 vaginal deliveries, with mortality rates of 6.4, 9.8, and 12.0 per 1,000 for normal, abruption, and retention groups, respectively. The code, authored by Sona Jasani and last updated in April 2026, performs statistical modeling to evaluate hazard ratios and temporal mortality patterns.
A 2026 study by Samuel J. Pearl examines the relationship between math anxiety and metacognitive monitoring in U.S. adults performing fraction arithmetic. The dataset includes responses from 685 adults who completed a fraction task, pre- and post-task performance judgments, and reported their math anxiety and self-concept. Findings suggest adults with higher math anxiety had less accurate monitoring of their performance.
A study presents a greedy optimization algorithm for allocating individuals to crosses in autogamous crop breeding programs. The algorithm was tested with an experimental barley resistance dataset and 60 simulated datasets, using population sizes of 400 and 1,200 and heritabilities of 0.7 and 0.9. The research was authored by Uche Joshua Okoye and includes user-friendly R code for improving breeding program efficiency.
Rev. Daniel Elis Axelrod created a dataset documenting the geometric reconstruction of a crop circle. The project involved mapping the crop circle's size and shape, estimating its area, and reconstructing it using circles and straight lines to derive extrusion volume and surface area. The dataset includes multiple file formats and was last updated on 2026-05-12.
A dataset supporting the research paper 'Weight reduction optimization and extreme sea-state sensitivity analysis of large jacket structures under multiple load cases and code constraints'. It includes a deepwater jacket structural calculation model, implementation code for a feasibility-first discrete TR-GA algorithm, and code-compliant strength verification results under multiple wave-current coupled load cases. The dataset also contains global sensitivity analysis data based on Morris screening and Sobol methods, authored by Li, Yuhang and hosted on Harvard Dataverse.
Matlab scripts and empirical data for applying an integral projection model based on dynamic energy budget theory to two species. The model aims to identify which species is more sensitive to shifts in temporal autocorrelation structure. The dataset includes parameterization data for the terrestrial crustacean Orchestia gammarellus and is authored by Isabel M. Smallegange.
An educational and research implementation of a Simple Genetic Algorithm (SGA) authored by Rafael Lahoz-Beltrá. The algorithm is applied to optimize a specific mathematical function, f(x)=abs(x-5/2+sin(x)), within the range 0<=x<=15, where the maximum value occurs at x=11. The dataset's size, row count, and last update date are unknown.