Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,069 datasets
Esther Torrente published data on IRBM-Z-2, an optimized Zika virus NS2B-NS3 protease inhibitor, on figshare in 2026. The dataset likely contains results from biochemical, cellular, and mouse model studies. It is a small dataset of 5.1 KB, shared under a CC-BY-NC-4.0 license.
Sree Nithish Reddy Gunapati's dataset contains results from a computational framework integrating machine learning and optimization for designing multifunctional bioactive peptides. The data likely includes sequences and properties of peptides predicted to have antioxidant, antifungal, and antibacterial properties for potential application in active food packaging systems. The dataset was last updated on May 12, 2026.
Statistical analysis results for primary model-related parameters. The dataset includes Feature Importance entries from a 100-tree Random Forest ensemble with over 15,000 samples per network topology and deterministic analytical results from four other methods. It was authored by Chung-Yuan Huang and last updated on 2026-05-21.
11 canonical definitions for AI infrastructure economics terms coined by Michal Piszczek, CTO of Archdesk. The dataset is a glossary.jsonl file with one JSON object per line, containing term, definition, coined_by, and canonical_url fields. It was uploaded by cdiamond to Hugging Face on 2026-07-21 and is intended for grounding assistants and RAG systems under a CC BY 4.0 license.
A 15.7 MB research dataset by Xiaotian Zheng, last updated in April 2026, accompanies a paper on constructing temporal point processes with memory. The dataset likely contains synthetic and real data examples used to illustrate a Bayesian inference methodology for modeling event durations with high-order Markov dependence. The work proposes a mixture modeling framework for conditional duration densities to create self-exciting or self-regulating point processes.
A study in Heilongjiang Province, China, used hyperspectral data and a stacked ensemble learning model to monitor canopy nitrogen content in the maize cultivar Jinboshi. The optimized model achieved a prediction R² of 0.826 and RMSE of 0.450. The dataset, shared by Haoquan Kong under a CC-BY-4.0 license, was last updated on 2026-05-21.
Data Department - State e-Government Agency provides statistics on kindergartens, enrolled children, and teaching staff. The dataset is broken down by statistical zones, regions, districts, and municipalities. The specific time period, row count, and column details are not provided in the available metadata.
Canonical public doctrine for Crimson OS, authored by Matt Gibson / Crimson Symphony Media. The dataset page indicates a 'Phase law' structure and includes tags like 'CRYSTAL' for verified theorems or operator-ratified doctrine. It was last updated on 2026-06-22.
A 2026 study by Peng Ren provides spectral data for 954 floral color loci from the Floral Reflectance Database (FReD). The dataset includes wavelength data, photoreceptor sensitivity data for Apis mellifera, and results from color diversity calculations using graph theory methods. It is accompanied by the original R Markdown code for reproducing the analysis.
A statistical procedure for identifying discrepant observations and evaluating panel discrimination in sensory experiments involving coffee blends. The work by Marcelo Ângelo Cirillo proposes a method tested across four experiments with coffees of different qualities and varieties. Results suggest the procedure was effective for discriminating blends relative to pure coffees, with concentrations and processing types not interfering with evaluations.
Marcos Felipe Nicoletti's study evaluates hypsometric rates for estimating tree height across different phases of the cutting cycle in Pinus taeda reforestation. The data likely contains results from five treatments, including a combined treatment, with statistical adjustments, coefficient of determination, standard error, and residual analysis performed. The Graybill identity test was used to assess the need for different models for different age classes.
SPSS mixed-model analysis data underpins a manuscript on the skeletal muscle phenotype of the DE50-MD dog model of Duchenne muscular dystrophy. John Hildyard authored this dataset, which includes raw quantitative data, statistical test outputs, and multiple comparisons correction data. The data covers longitudinal and post-mortem muscle samples from the vastus lateralis.
A study of mathematics books written in Spanish and published during the 16th century. The analysis identifies and categorizes all examples from these books and their relations to daily situations of the time. The work was conducted by María José Madrid using historical-mathematical and content analysis techniques.
An introductory work on the Adomian Decomposition Method for solving differential equations. The author, R. G. G. Amorim, presents the method and applies it to example problems from physics. The work is published on paperswithcode under an Open Access license.
2.4 KB of data describes the identification of the piperidyl urea derivative BAY-439 as a potent and selective inhibitor of human Phospholipase A2 Group V (hPLA2-G5). The dataset, authored by Gernot Langer and last updated on 2026-05-19, contains information on this chemical probe accepted by the Structural Genomics Consortium and its inactive negative control, BAY-163.
A quantitative framework for optimizing maritime renewable energy integration, developed by Khaled Mili and last updated in 2026. The supporting document describes a study analyzing 47 maritime installations from 2019 to 2024, employing high-frequency temporal sampling and fine-resolution spatial analysis. It synthesizes technical performance metrics with stakeholder assessment data from 156 participants.
Code and data generate figures for a paper on accelerating Monte-Carlo estimation with derivatives of high-level finite element models. The repository, authored by Paul Hauseux from the University of Luxembourg, archives code to solve a 1D stochastic viscous Burgers equation. A Docker image is provided to run the code, which is licensed under LGPL v3.0.
Stereolithography (STL) data of optimal configurations for biphysical cloaks designed through topology optimization, as shown in figures from the associated manuscript. The data, provided by author Garuda FUJII, serves as supplementary material for experimental demonstrations. Each STL file corresponds to a specific figure number in the manuscript, such as 'Fig4b.stl' for the configuration in Figure 4(b).
Six participants were involved in a 51-week mixed-methods single case experimental design to investigate a person-centred active rehabilitation programme for suspected Chronic Traumatic Encephalopathy (CTE). The dataset includes qualitative interview data and quantitative outcome measures for cognitive function, executive function, mindful attention, mood, and behavior. Author Rachael Hearn published the data on figshare under a CC-BY-NC-SA-4.0 license, with a last update timestamp of 2026-05-28.
A 3D molecular structure file for a novel 4-methylquinazoline derivative identified as a potent and selective PI3Kδ inhibitor. The compound, designated 48, exhibited single-digit nanomolar potency and was evaluated in murine models of LPS-induced acute lung injury. The data was authored by Deyu Wu and last updated on 2026-05-24.