Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,063 datasets
Leicester City Council provides quarterly data on the types of buses operating in the city, categorised according to the EU Emissions Directive. The dataset is part of a public transport dashboard and links to detailed technical documentation from the UK Department for Environment, Food & Rural Affairs. It was last updated on 2026-06-17.
Farouk Mark Mukiibi's dataset documents verifiable citations and references to the Minimum Viable Relationships (MVR) Framework by major AI systems. It consolidates multi-platform proof of attribution across OpenAI ChatGPT, xAI Grok, Perplexity AI, Google Gemini, Meta AI, and Microsoft Copilot. The 1.1 MB dataset, last updated on 2026-05-24, includes machine-readable JSON evidence, public archives, and SHA-256 integrity hashes.
255 plant-based larvicidal compounds against the Zika vector Aedes aegypti were analyzed using QSAR models developed with CORAL software and Monte Carlo optimization. The dataset includes pLC50 bioactivity values and molecular docking results for selected compounds. S. Lotfi published the data on figshare in 2026.
Version 1.0.0 biometric data released by Birmingham City University on 1 August 2025. It is an Excel workbook with six worksheets containing participant demographics, task outcomes, time-series biometric measurements, and derived statistical summaries. The time-series worksheets contain raw biometric measurements recorded during two cycles of activity.
A dataset from 2026 describes a series of biaryl-substituted pyrazolopyrimidine inhibitors of the TgCDPK1 enzyme for treating Toxoplasma gondii infections. The data, shared by Michael P. Mannino on figshare, includes compound properties optimized for metabolic stability, plasma protein binding, efflux, and pharmacokinetics. It is a small dataset of 8.8 KB in CSV format.
5.5 KB of statistical analysis results for the AMGST traffic forecasting model. The dataset, authored by Pei Shi and last updated in June 2026, contains experimental results from evaluating the model on four public traffic datasets. It is hosted on figshare under a CC-BY-4.0 license.
Slow Ripening Grapevine Genotypes contains supplementary datasets from a project characterizing grapevine material that fails to accumulate sugar until late in the season. The repository includes 11 Excel tables with statistical comparisons and logistic regression model parameters for traits like total soluble solids, berry weight, and firmness across multiple years and experiments. Pietro Previtali authored the dataset, which was last updated on June 4, 2026.
Hang Li's dataset contains survey results from 272 orthopedic theatre nurses across eight tertiary hospitals in Shanxi Province, China, collected between September and December 2024. The data includes scores for Knowledge, Attitude, and Practice (KAP) regarding the use of orthopedic power tools and results from multivariable regression analysis. It was last updated on figshare in May 2026.
Geoscience Australia conducted a simulation experiment comparing statistical and mathematical techniques for predicting seabed mud content across the Australian continental margin. The study assessed factors including regions, sample densities, and interpolation methods using bathymetry, distance-to-coast, and slope as secondary variables. A novel combined method, random forest and ordinary kriging (RKrf), demonstrated a relative mean absolute error up to 17% less than a control method.
Event permit applications for Chicago's public parks include details on the applicant, event type, description, park location, and scheduled times. The dataset is maintained by the Chicago Park District and tracks permit statuses. It is available in multiple formats including CSV, JSON, and XML.
Statistical returns from central government monitored bodies in the UK, collected by the Ministry of Justice. The dataset is licensed under OGL-UK-3.0 and was last updated on 2026-07-08.
Supplementary file 1 from a hybrid study by Ye Liu assesses modifiable risk factors for atrial fibrillation and flutter in young adults aged 15-39 years. The dataset likely contains results from a global burden analysis spanning 1990 to 2021 and a local cohort study, including age-standardized rates and regression analyses. It was last updated on figshare in May 2026 under a CC-BY-4.0 license.
Sixty-five anonymized clinical radiotherapy plans were used to benchmark three dose calculation algorithms against a Monte Carlo standard. The study, authored by Yutong Zhao and last updated in May 2026, found AXB generally most accurate, while CCC performed comparably for lung cases. All deterministic algorithms exhibited systematic dose deviations in lung tissue and planning target volumes.
Analysis of combat fighting in Homer's Iliad contains data from a textual analysis of the epic poem. The analysis was performed independently by two reviewers with classical studies backgrounds, using a Modern Greek translation by Prof. Dimitrios N. Maronitis that received the 2011 State Award for Interlanguage Translation. Thematic analysis of transcripts was conducted by three investigators who established final themes by consensus.
MMLKG is a thesaurus for describing a repository of computer-verified mathematical papers written in the Mizar language. The dataset includes a GraphML file for loading into graph databases, CSV relationship files, and RDF data. It was prepared by Dominik Tomaszuk and is available under an Open Access license.
A dataset from a paper detailing the use of a statistical design of experiments (DoE) strategy to optimize a rubber compound formulation. The work by Pablo Ernesto Salvatori focuses on maximizing soybean oil content while meeting tire tread performance targets. It includes fitted models and response surfaces for properties like glass transition temperature and Mooney viscosity.
A research article by Christopher Wolf of KU Leuven analyzes equivalent keys in Multivariate Quadratic public key schemes. The work demonstrates reductions in private and public key space size by several orders of magnitude for schemes like Matsumoto-Imai, hidden field equations, and unbalanced oil and vinegar. These findings have applications in cryptanalysis and memory-efficient implementations of MQ-schemes.
Larissa R. Terra's research paper details an experiment to optimize solid-phase microextraction (SPME) parameters for analyzing volatile metabolites from whole papaya. The work evaluates two fiber types and uses central composite design to assess the effects of conditioning and exposure time on the number of detectable compounds. A conditioning time of 10 minutes and exposure time of 30 minutes was sufficient for detecting more than 100 compounds.
Trust statistical tables provide key taxation and select accounting information for all trust tax returns that have been assessed or reassessed. The data is published by the Canada Revenue Agency and covers tax years 2019 to 2023. The dataset was last updated on 2026-06-22.
A monograph chapter by William B. Hansen of Atrium Health Wake Forest Baptist, sourced from paperswithcode. The text discusses statistical power in prevention research, referencing a review of 46 longitudinally followed cohorts. It critiques an overemphasis on sample size as the primary method for increasing power.