Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,080 datasets
European pension fund statistical balance sheet items, broken down by geographical area and maturity, on a quarterly basis. The dataset originates from the Dutch Ministry of the Interior and Kingdom Relations and has been collected since 1986. It is published under a CC-BY-4.0 license via the EU Open Data platform.
Quarterly statistical cash flow statements for pension funds, starting from 2006. The data is provided by the Ministerie van Binnenlandse Zaken en Koninkrijksrelaties and is available in JSON format under a CC-BY-4.0 license.
Quarterly statistical balance sheet data for insurance corporations, broken down by sector and maturity. The dataset originates from the Dutch Ministry of the Interior and Kingdom Relations and covers a time series starting in 2002. It is provided in a fracture-free format under a CC-BY-4.0 license.
Dan Jiang's 5.5 KB Excel file provides a unified collection of mathematical formulas and descriptions for machine learning evaluation. The dataset, last updated in April 2026, consolidates performance measures, LIME interpretability measures, and statistical validation tests. Its small size suggests it is a reference sheet rather than a large observational dataset.
A structural comparison table of the proposed Felis Catus Optimization (FCO) algorithm against 17 competitor metaheuristic methods across seven design criteria. The 5.5 KB Excel file uses checkmarks and crosses to denote feature presence or absence. It was authored by Mohammad Salehi and last updated on April 15, 2026.
Geoscience Australia data from a 2010 study comparing methods for spatial interpolation of seabed sand content across the Australian Exclusive Economic Zone (AEEZ). The study evaluated 18 methods and 36 variable combinations, finding RFIDS and RFOK to be among the most accurate, reducing prediction error by up to 7%.
5.5 KB of Spearman correlation scores and p-values comparing predicted versus experimental gene activity changes from knockout experiments. The data, authored by Clelia Corridori and last updated in April 2026, is stored in an XLS file. It highlights statistically significant correlations (p < 0.05) in bold formatting.
369 paired statistical comparisons of explanation faithfulness metrics, specifically per-image AUCs, for two AI models. Md Mehedi Hasan Santo published this dataset on figshare in April 2026. It contains mean differences and p-values from paired t-tests and Wilcoxon signed-rank tests.
Health statistics for Year 7 students in the Australian Capital Territory, published by the ACT Government's Epidemiology Section. The dataset was last updated in March 2026 and is available in multiple formats including CSV and JSON.
2026 data from the ACT Government's Epidemiology Section provides public health statistics on selected cancers. The dataset is available in multiple machine-readable formats including CSV, JSON, and XML. Specific row and column counts are not provided in the source metadata.
Performance metrics for regression models analyzing muscle parameters. The dataset contains statistical measures including standard errors, standardized coefficients, Cohen's fΒ² effect sizes, and statistical power. Byungmun Kang published this data in March 2026.
57.3 KB of R code and a CSV data table accompany a specific academic paper. The dataset was authored by Kirill Korznikov and is available under a CC-BY-4.0 license. It was last updated on April 20, 2026.
OptiVerse is a benchmark dataset for evaluating large language models on complex optimization tasks. The dataset was created by Waicheng and was accepted for publication at ACL 2026 Findings. The description indicates it aims to address limitations in existing benchmarks that focus narrowly on Mathematical Programming and Combinatorial Optimization.
Monte Carloβderived uncertainty estimates for carbon stock, carbon emissions, and carbon budget. The dataset is a 9.5 KB Excel file authored by Ruijing Zhang and last updated on April 24,ζ们εη° 2026. It is shared under a CC-BY-4.0 license on the figshare platform.
15 indicators across preparation, administration, and post-election phases form a composite index for Indonesia's 38 provinces, yielding a national mean score of 63.43. The dataset, created by Ahmad Nur Hidayat and last updated in March 2026, is derived from verified administrative records for the 2024 general election. It reveals territorial disparities, with western provinces outperforming eastern ones and post-election engagement scoring highest at an average of 68.48.
The Indonesian Electoral Participation Index (IPP) is a multidimensional statistical model of electoral engagement. It is constructed from 15 indicators across preparation, administration, and post-election phases for the 2024 general election in 38 provinces. The dataset was created by Ahmad Nur Hidayat and published on figshare under a CC-BY-4.0 license.
25 high-level combat athletes participated in a cross-sectional study measuring energy and oxidative stress markers alongside six strength performance indicators. Data includes plasma ATP, antioxidant markers (T-AOC, SOD, GPX), oxidative damage marker MDA, and performance tests like standing long jump and 1-RM bench press. The research was conducted by Jinling Huang to examine associations between physiological indicators and athletic performance.
UKCCSRC Call 2 project data from a poster presented in London on 27 June 2016. The project, led by the British Geological Survey, aimed to develop a multi-modal sensing system for measuring CO2 mass flow in Carbon Capture and Storage pipelines. The work focused on creating a calibration platform and evaluating sensor performance under single-phase and two-phase flow conditions.
A secondary analysis of a randomized trial involving 23 tobacco-dependent adults who underwent a single rTMS session targeting either the left frontopolar cortex or vertex. The dataset includes response times and self-reported craving ratings from an image-based cue-reactivity paradigm conducted before and after stimulation. It examines the modulation of these implicit and explicit measures by stimulation site and their association with smoking outcomes.
Larissa Vieira's secondary analysis dataset from a randomized trial investigates response times to craving-item ratings in tobacco-dependent adults. The data originates from a cue-reactivity paradigm before and after a single 1 Hz rTMS session targeting either the left frontopolar cortex (n=12) or vertex (n=11). It examines the modulation of response time by stimulation site and its association with smoking outcomes at baseline and one-week follow-up.