Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,069 datasets
ZhiPeng Wu's dataset provides a performance evaluation on extreme events and categorical metrics for Wind Speed (u10). The data is stored in a 5.5 KB XLS file and was last updated on June 4, 2026. The license is CC-BY-4.0, indicating open access for reuse.
STGLDWeather outperforms competing methods in a Root Mean Square Error (RMSE) comparison. The 13.5 KB Excel file contains results where statistical significance (p < 0.05) compared to GraphCast is indicated. Author ZhiPeng Wu uploaded this comparative analysis to figshare in June 2026.
Results of Monte Carlo simulations and probability density function (PDF)βbased error assessments authored by Jessica V. Eberle. The dataset is a 13.5 KB XLS file last updated on June 4, 2026. It is available under a CC-BY-4.0 license on the figshare platform.
A 9.5 KB Excel file provides a statistical summary of Quantitative Trait Loci (QTLs) significantly associated with dietary fiber-related traits. The QTLs were identified by Genome-Wide Association Study (GWAS) in populations of Vaccinium meridionale. The dataset was authored by Ginna Patricia Velasco Anacona and last updated on June 4, 2026.
A simulation experiment using samples from the Geoscience Australian Marine Samples database to compare statistical and mathematical techniques for predicting seabed mud content. The study assessed five factors affecting accuracy, including regions, methods, and sample densities, using ten-fold cross-validation and secondary variables like bathymetry. Outcomes aim to improve the modeling of physical properties for marine biodiversity prediction.
A comparative evaluation scores the performance of Google Gemini 2.5 Flash and DeepSeek-V3.2 LLMs against expert Committee on Publication Ethics responses. The study analyzes 12 authorship and contributorship cases using three prompting strategies, with responses rated across seven domains on a 5-point Likert scale. The dataset was authored by Kannan Sridharan and published on figshare in April 2026.
A PDF document presents a cross-sectional analysis comparing the performance of Google Gemini and DeepSeek LLMs against expert COPE forum responses. The study includes 12 authorship and contributorship cases, with responses scored across seven domains on a 5-point Likert scale by independent raters. Author Kannan Sridharan published this 89.1 KB file on figshare in April 2026.
12 authorship and contributorship dispute cases from the Committee on Publication Ethics forum were used to evaluate two large language models. Kannan Sridharan published this comparative analysis in April 2026. The dataset contains performance scores across seven evaluation domains for Google Gemini 2.5 Flash and DeepSeek-V3.2.
Twelve authorship dispute cases from the Committee on Publication Ethics forum were used to evaluate two large language models. Kannan Sridharan published this comparative analysis in April 2026. The dataset contains scored model responses across seven evaluation domains.
12 authorship and contributorship dispute cases from the Committee on Publication Ethics (COPE) forum were used to evaluate two large language models. The study, authored by Kannan Sridharan and shared in 2026, scored model responses across seven domains, including Actionability of Recommendations and Consistency with COPE Principles. Both Gemini and DeepSeek models achieved perfect scores in one domain but showed weaknesses in identifying specific ethical issues.
12 authorship and contributorship cases from the Committee on Publication Ethics (COPE) forum were used to evaluate two large language models. The dataset contains scores across seven domains, including Actionability of Recommendations and Identification of Ethical Issues, rated on a 5-point Likert scale. Kannan Sridharan published this comparative analysis in April 2026.
12 authorship and contributorship cases from the Committee on Publication Ethics forum were used to evaluate two large language models. Kannan Sridharan published comparative performance scores across seven domains, including Actionability of Recommendations and Consistency with COPE Principles, in April 2026. The dataset contains Likert scale scores and qualitative disagreement rates from this cross-sectional analysis.
Brian Liu's research on figshare, last updated April 28, 2026, proposes an estimator for extracting compact decision rules from tree ensembles. The 12.5 MB repository includes code and data supporting the development of exact and approximate algorithms for rule extraction. The work establishes non-asymptotic prediction error bounds and demonstrates performance against existing algorithms through experiments.
180 self-contained operations research tasks benchmarked from academic literature. Each task includes a natural-language problem description, mathematical formulation, reference Gurobi implementation, test instances, and an automated feasibility checker. Created by SmartOR, the dataset was last updated on June 14, 2026.
A 46.7 MB dataset in XLSX format, last updated on May 13, 2026. It was created by Jose Hernandez and shared under a CC-BY-4.0 license on figshare. The data examines coloring as a representational practice in primary mathematics textbooks from Chile and Japan.
A discovery chemistry campaign describes novel camptothecin-based linker-payloads for antibody-drug conjugates (ADCs). The dataset, created by Vlad Bacauanu and last updated in April 2026, likely contains results from the synthesis and evaluation of cytotoxic analogs and optimized linkers. It includes data on constructs enabling high drug-to-antibody ratio ADCs with good biophysical properties and in vivo efficacy.
Julijus Bogomolovas published a structured dataset for statistical analysis of cardiac function in PKN2 cardiomyocyte-specific knockout and control mice. Each row corresponds to one animal at one time point, including genotype, animal identifier, age, and measured parameters such as fractional shortening and left ventricular dimensions. The dataset is formatted for mixed-effects modeling with repeated measurements nested within animals.
19.9 KB of Excel data supports a quantitative risk assessment of the developmental neurotoxicity of F-53B, a PFOS replacement. The study by Longfei Feng integrates in vitro phenotypic profiling, transcriptomics, adverse outcome pathway modeling, and pharmacokinetic modeling. It estimates fetal brain concentrations of 0.09β14.66 ng/mL based on human biomonitoring data.
2.5 KB of data from a drug discovery project targeting the Werner syndrome helicase (WRN) protein. The dataset, authored by Justin A. Caravella and shared on figshare, describes the structure-based design of covalent inhibitors, including compound 26, which demonstrated in vivo efficacy in an MSI-H Xenograft tumor model. It was last updated on April 29, 2026.
Data and code supporting a study on joint optimization of land carbon uptake and albedo for climate cooling. The dataset, by Alexander Graf of Forschungszentrum JΓΌlich, accompanies research published in Communications Earth & Environment. It likely contains model outputs or parameters related to land-atmosphere interactions and climate feedbacks.