Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,080 datasets
Supplementary material 3 from a Bayesian network meta-analysis evaluating therapies for granulomatous mastitis. The 25.0 KB XLS file, authored by Pin Wang and last updated in April 2026, provides operational definitions for clinical outcomes like Complete Response (CR), Overall Response Rate (ORR), and Recurrence Rate (RR) used in the included studies.
Supplementary data from a Bayesian network meta-analysis evaluating therapies for granulomatous mastitis. The dataset contains matrices of pairwise comparisons for overall response rates, reported as odds ratios and 95% credible intervals. It was authored by Pin Wang and published on figshare under a CC-BY-4.0 license.
NYC Department of City Planning's Housing Database tracks net changes in housing units for New York City Community Districts. It aggregates data from Department of Buildings-approved construction and demolition jobs filed or completed since January 1, 2010. The dataset includes census unit counts, net changes, and units pending completion.
A 9.5 KB Excel file containing results from a hyperparameter optimization study. The work by Mario Koddenbrock, last updated in May 2026, demonstrates that tuning on synthetic SynthMT images can improve the SAM3Text model to human-grade performance on unseen, real IRM data.
A 5.5 KB Excel file containing the results of a model performance comparison with statistical analysis. The dataset, authored by Shanyue Wang, was last updated on April 28, 2026. It calculates the delta, or difference, between the outputs of two models named MsgaBpred and EpiGraoh.
Statistical results from coal petrographic identification of the B-coal seams in the Xishanyao Formation. The dataset is a 9.5 KB XLS file authored by Bin Chen and last updated on May 5, 2026. It is shared under a CC-BY-4.0 license on the figshare platform.
Statistical results of coal petrographic identification for the B-coal seams in the Xishanyao Formation. The dataset is a 19.1 KB XLSX file authored by Bin Chen and last updated on 2026-05-05. Its license is CC-BY-4.0, facilitating open reuse.
Hongjing Chang published a dataset titled 'Samples’ Evaluations on 4 Translations in RRQ Based on Descriptive Statistical Analysis and One-Sample T Test' on figshare in May 2026. The 5.5 KB XLS file likely contains statistical evaluation data for four different translations, possibly related to a research questionnaire (RRQ). The dataset's specific row count and column details are not provided in the metadata.
28-day compressive strength data from cement plants was used to validate a hybrid Transformer-XGBoost prediction model. The model achieved an average R² of 0.94 in 25 Monte Carlo cross-validations, demonstrating high accuracy for small-sample scenarios. The dataset contains the results of this optimization study, authored by Dianyuan Ju and shared in 2026.
Statistical results from a study proposing a Transformer-XGBoost model for predicting 28-day cement compressive strength. The method was validated using real-world strength testing data from cement plants, achieving an average R² of 0.94 in Monte Carlo cross-validation.
892.7 KB of research materials authored by Rachid Belfadli, last updated on April 22, 2026. The content includes a paper proving the existence and uniqueness of solutions for two classes of doubly reflected backward stochastic differential equations driven by pure jump Markov and jump semi-Markov processes. The analysis is based on the Snell envelope technique and a penalization method.
Simulation, training, and optimization data supporting a 2026 research publication on Floating Production Storage and Offloading (FPSO) units. The dataset, 65.2 MB in size, was created by Jiaqi Zhang and colleagues to capture non-linear fluid-structure-mooring interactions under extreme environmental conditions. It is stored in DAT and XLSX file formats.
Opus 4.6 10000X is a dataset of 10,000 high-fidelity reasoning traces synthesized using the Claude Opus 4.6 model. It was created by user 'ansulev' and last updated on Hugging Face in May 2026. The dataset is designed to capture the model's internal 'Chain of Thought' and reasoning patterns.
981 globally distributed hosting providers form the basis of this large-scale quantitative study correlating server technology stacks with DNS resolution efficiency. The analysis isolates the impact of Cloudflare, LiteSpeed, Apache HTTPD, and Nginx using a 10% trimmed mean methodology to exclude latency outliers. Empirical findings indicate a statistically significant performance advantage for edge-native, decentralized architectures over legacy centralized setups.
A cleaned mathematics supervised fine-tuning dataset built by kaushik-harsh-99 and last updated on 2026-05-17. It contains instruction-solution pairs, mathematical proofs, derivations, and olympiad-style solutions for theorem reasoning and stepwise explanations. The dataset is designed specifically for mathematical supervised fine-tuning and removes explicit chain-of-thought tags.
949,100 Monte Carlo simulation replications accompany a methodological study on stabilizer variables for measurement invariance. The results are organized into six phases, including a core performance evaluation of 800,000 replications and sensitivity analyses. The dataset, created by Salim Yılmaz, was updated in March 2026.
A figshare-hosted dataset by Jihang Jia, last updated March 2026, containing materials for a statistical test of the mean matrix. The 3.4 MB resource includes PDF, ZIP, and TXT files detailing a projection-based method that incorporates structural information of matrix-valued data.
A 39.8 KB dataset from figshare, authored by Daniel Ernesto Rojas Ventura and last updated on 2026-04-24. It contains consolidated architectural correlation tables mapping physical constants like the Bekenstein Bound and Planck Area to signed integer limits. The dataset also includes a forensic validation of the 1986 Chernobyl disaster as a macro-scale arithmetic overflow event.
Results from a Bayesian nonparametric analysis using a Hierarchical Dirichlet Process (HDP) model, likely applied to ETF (Exchange-Traded Fund) data. The dataset was published by author P2SAMAPA on the Hugging Face platform. It was last updated on 2026-06-22 08:16:08.
CO-OPS water level stations across coastal U.S. states and territories provide annual exceedance probability levels for extreme high and low water events. The National Oceanic and Atmospheric Administration produced this dataset by analyzing historical data from stations with at least 30 years of records. The statistical analysis focuses on storm tides, excluding wave runup and tsunami peaks.