Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,058 datasets
3.0 KB of text codifies an industrial-grade, zero-allocation bare-metal processing framework. The registry, authored by Jamie Davis and licensed under CC BY 4.0, integrates a unified architecture for deterministic execution and high-resolution neuroanatomical mapping. It was last updated on June 1, 2026.
250 mL of UHT cow's milk provides approximately 267% of the daily human need for phosphorus, according to in vitro simulated gastric digestion results. This dataset contains total and dialyzed levels of Zn, Ca, and P determined by ICP OES for raw sheep’s milk, soybean extract, raw cow’s milk, UHT goat’s milk, and UHT cow’s milk. Accuracy was verified using a CRM (NIST 8435) and results were statistically evaluated by Student’s t-test.
A research paper and associated experiments focus on optimizing interactive debugging cycles for rule-based entity matching. The work, by Fatemah Panahi of the University of Wisconsin–Madison, proposes techniques like 'early exit' and 'dynamic memoing' to reduce computation time. Experiments validating the approach were conducted on six real-world data sets.
An empirical-Bayes method for exploiting spatial structure in large multiple-testing problems, presented by Wesley Tansey of The University of Texas at Austin. The method, called false discovery rate smoothing, finds spatially localized regions of significant test statistics and adjusts significance thresholds to control the overall false-discovery rate. It is applied to an fMRI experiment on spatial working memory and its code is publicly available in Python and R.
A methodological paper and likely accompanying code or synthetic data introducing novel statistical models for analyzing classification performance in hierarchical datasets. The work by Kay H. Brodersen from the University of Zurich proposes Bayesian mixed-effects models that account for within-subject and across-subject variance, using MCMC for model inversion and selection. It demonstrates the approach on both synthetic and empirical data to improve inference sensitivity and validity.
Module 129 details a zero-heap, zero-copy, real-time First-Order Lead-Lag Phase Shifter Network for bare-metal embedded systems. The module provides a recursive single-pole, single-zero filter network for discrete phase adjustments, using 64-bit integer operations and Q16.16 fixed-point format. It was authored by Jamie Davis and published on figshare under a CC BY 4.0 license, with a last update timestamp of 2026-05-31.
Module 128 from the Davis Logic V2 project details a fixed-point Phase-Locked Loop (PLL) loop filter node designed for bare-metal embedded systems. The component, authored by Jamie Davis and released under a CC BY 4.0 license, implements a zero-heap, zero-copy Proportional-Integral (PI) filter for real-time frequency and phase tracking. The dataset, last updated on 2026-05-31, is a 2.3 KB text file describing the mathematical architecture and computational performance of the filter.
Module 118 delivers a zero-heap, zero-copy, real-time First-Order Leaky Integrator with variable loss gain optimized for telemetry accumulation inside bare-metal embedded execution grids. The component processes streams sample-by-sample using a structural first-order lossy difference equation to prevent boundless state drift from sensor biases. Authored by Jamie Davis and released under a CC BY 4.0 license, this 3.4 KB text file was last updated on May 31, 2026.
2026 simulation results and analysis code for a psychometric method. The dataset contains pre-computed results from 270,000 applications of the Stabilizer Variable Test (SVT) across 540 factorial conditions, generated by Salim Yılmaz. It also includes three empirical datasets from a 2023 survey of Turkish adults and nine R scripts for full analysis replication.
Module 109 from the Davis Logic V2 project provides a zero-heap, real-time First-Order DC Blocking Filter designed to eliminate low-frequency drift from streaming data. The filter uses a Q16.16 fractional layout and 64-bit integer arithmetic to prevent quantization errors. Jamie Davis authored this 3.4 KB text file, last updated on 2026-05 31 under a CC BY 4.0 license.
Geoscience Australia's dataset supports a study comparing methods for spatial interpolation of seabed sand content within the Australian Exclusive Economic Zone (AEEZ). The research evaluates 18 machine learning and geostatistical methods, including RFIDS and RFOK, using samples extracted in August 2010. Model averaging and specific parameter choices are shown to improve prediction accuracy by up to 7%.
Comments on credit changes and deviations to budget estimates for the Austrian city of Linz in 2019. The dataset is provided as plain text and includes links to supporting documentation for proof of credit changes and discrepancies between estimates and invoices. It is published under a CC-BY-4.0 license by a cooperation of Austrian open government data entities.
Planning decisions for London boroughs, categorized by development type, decision speed, and enforcement actions. Statistics date back to the 1999-00 period and are published by the Greater London Authority. Regional figures are rounded and percentages may not sum to 100 due to exclusions and rounding practices.
A research paper evaluates the Pulsar algorithm for detecting outbreaks in syndromic time series data. The analysis uses daily syndromic counts from emergency departments of four major hospitals in the Athens area during August 2002 to August 2003. The work was authored by Urania Dafni of the University of West Attica.
Supplementary file 1 contains data from a study evaluating a novel enteric-coated formulation of tylvalosin tartrate against Mycoplasma hyopneumoniae in pigs. The dataset includes in vitro minimum inhibitory concentration and biofilm inhibition results, as well as clinical trial outcomes from 60 pigs across six treatment groups. The data was authored by Jiajing Li and last updated on figshare in June 2026.
Historical Major League Baseball team statistics compiled from Stathead. The data includes seasonal team performance metrics such as games played, wins, losses, runs scored and allowed, and win percentages. It also contains calculated win expectation estimates using different power exponents for converting runs to wins, along with error metrics for these estimates. The dataset was authored by Ricardo de la Peña and is hosted on Harvard Dataverse.
Experimental data from an ultrasound-assisted pectinase extraction process for flavonoids from peppermint leaves. The dataset includes results from response surface methodology optimization, yielding 12.87 g/100g DW total flavonoids, and UHPLC-MS/MS analysis identifying 38 compounds with quantitative data for four. The dataset was published by Bei Liu on figshare in 2026.
An open-access raw source implementation for an embedded real-time Quality of Service telemetry lag monitor tracks microsecond arrival deltas across data links. The module, authored by Jamie Davis and licensed under CC BY 4.0, is designed to isolate transmission jitter and network timing drift on the fly. Its architecture operates without runtime heap allocations and uses flat static buffers for continuous jitter estimation.
UK Ministry of Justice annual statistical bulletin on safety in prison custody. The dataset likely contains figures on deaths, self-harm incidents, and assaults within the prison system. It is designated as Official Statistics and is published under the OGL-UK-3.0 license.
Scottish Government National Statistics detail cancer waiting-time performance across NHS boards. The statistical release covers urgently-referred patients and is broken down by tumour site. Last updated on July 8, 2026, the data is designated as official National Statistics.