Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,062 datasets
Linz municipal data from 2019 details credit changes and deviations from budget estimates. The dataset includes notes on credit transfers, overruns, and household remnants. It was published by Cooperation OGD Österreich and Wikimedia Österreich under a CC-BY-4.0 license.
A dataset from the Australian Ocean Data Network describing mathematical filters, or admittance functions, for modeling isostatic processes. The data likely contains spectral representations of the relationship between gravity and topography based on elastic and visco-elastic rheologies. It was last updated on 2026-06-27.
Xinglong Lu published a dataset of 246 locally advanced gastric cancer patients who received neoadjuvant immunochemotherapy at The First Hospital of Lanzhou University between 2021 and 2024. The data includes hematological parameters measured before and after treatment, used to predict major pathologic response. The dataset was last updated on 2026-05-25.
Jamie Davis created a high-performance angular normalization engine for motion control systems. The dataset, last updated on June 2, 2026, is a 1.7 KB text file describing a module designed to eliminate boundary discontinuities in continuous motion tracking. It is engineered for constant-time execution under 5 nanoseconds with zero heap allocation.
November 2014 statistical classification and delineation of settlements in Northern Ireland by the DOE Planning Service. Boundaries are defined for settlements exceeding thresholds of 20 or more households and 50 or more usual residents. The dataset is published by OpenDataNI under the OGL-UK-3.0 license.
A data-driven risk-aware model predictive control framework for discrete-time linear systems under process noise. The dataset, published under CC-BY-4.0 by Pouria Tooranjipour, includes files in PNG and EPS formats totaling 223.8 KB. It was last updated on June 2, 2026.
A case-study dataset from the HPP Belo Monte spillway, comparing physical and mathematical model results for free-flow and gate-controlled operations. The data was generated by author Marcus Fernandes Araujo Filho to evaluate the FLOW-3D® software's capability in modeling complex low-drop spillway flows. Comparisons include discharge capacity, flow lines, and pressure distribution.
Yuanlong Zhang from Tsinghua University created this dataset for a deep learning project on large depth-of-field ultra-compact microscopes. It includes blurry test images, sharp network output images, and pre-trained network weights. The dataset is hosted on Papers with Code and is licensed as Open Access.
A methodological paper by Mithat Gönen discusses criteria for comparing Bayes factors used in two-sample studies. The work proposes a new objective criterion based on classification theory to evaluate the performance of different Bayesian prior choices. It is published as an Open Access paper on the paperswithcode platform.
Mike Behrisch from TU Wien created this dataset containing formal verification proofs for a partial ternary Boolean conjunction. The verification approach translates the problem into Boolean satisfiability problems specified in SMT-LIB2.0 and solved using the Z3 solver from Microsoft Research. The dataset includes SMT-LIB2.0 implementation files, solver outputs, and formal proof documents.
Fernando Raphael Pinto Guedes Rogério authored a preliminary study on the reproducibility and agreement of different dynamic baropodometry protocols during gait. The dataset likely contains measurements of peak plantar pressure and pressure-time integral for eight foot masks, collected from fifteen volunteers across three time points. The study compared shortened one-step and three-step protocols to a standard midgait protocol.
Three pre-fitted Bayesian models for gene panel selection, each derived from a distinct spatial transcriptomics dataset. The models are fitted on data from the Zhang, Moffit, and Codeluppi datasets, as described in the associated paper. Author Yida Zhang from Harvard University made these models available via the paperswithcode platform.
An open-access raw source implementation for an embedded High-Level Data Link Control (HDLC) bit-stuffing framing engine. The module, authored by Jamie Davis, is licensed under CC BY 4.0 and was last updated on May 29, 2026. It is designed to evaluate high-velocity telemetry streams for transmission over synchronous serial links.
Experimental data and scripts from the IJCAR 2022 paper describing Evonne, an interactive proof visualization tool for description logics. The dataset was created by Christian Alrabbaa of Technische Universität Dresden. The data likely contains results and configurations used to evaluate the Evonne system.
Feng Tang's proof-of-concept lipidomics dataset analyzes cerebrospinal fluid from pediatric meningitis patients. It contains lipid metabolite profiles from 13 CSF samples across 10 patients, grouped by disease type and phase. The data was generated using UPLC-MS/MS on a Q Exactive mass spectrometer and processed with LipidSearch software.
Feng Tang's lipidomics dataset contains 13 cerebrospinal fluid samples from 10 pediatric patients analyzed by UPLC-MS/MS. The data compares lipid metabolite profiles across acute-phase purulent meningitis, recovery-phase purulent meningitis, acute-phase viral meningitis, and non-meningitic controls. The dataset was last updated on 2026-05-29.
38.2 MB of material property data compiled from multiple published studies, hosted on figshare by Wei-Ting Tang. The dataset includes hydrogen and nitrogen uptake in metal-organic frameworks, joint CO2 and CH2 uptake, water solubility, and LogD distribution coefficients. It was last updated on 2026-05-28.
Richard Goodman of IAOM presents machine-certified mathematical results on quantum measurement limits, including a proven finite Cramér-Rao bound and a concrete two-outcome readout saturating the quantum limit. The work, last updated in July 2026, derives eigenstructures for degenerate information matrices and introduces a trichotomy for handling division-by-zero failures. It includes a theorem demonstrating that thermal noise can restore parameter identifiability.
33 years of historical under-five mortality rates from 1990 to 2022 across the eight divisions of Bangladesh, based on 64,697 records from the Bangladesh Demographic and Health Survey 2022. The dataset was created by Md. Ismail Hossain and includes forecasts up to 2030 using a Bayesian Spatiotemporal model. It enables analysis of regional disparities and national progress in child health.
Aggregated data from scans of Pension Credit data is provided at the 1992 ward level. The dataset likely contains counts or statistics on benefit claimants for geographic analysis. It is available under the Open Government Licence for the United Kingdom (OGL-UK-3.0).