Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,062 datasets
Bruno Gama Magalhães conducted a quantitative study analyzing user satisfaction with the quality of services at CEOs (Specialized Dental Centers) in Pernambuco, Brazil. The dataset likely contains survey responses from 156 users present in waiting rooms who had undergone at least one clinical procedure. Data analysis was performed using SPSS version 13.0, employing descriptive and analytical statistics including Pearson's χ2 test at a 5% significance level.
Taubaté, Brazil, hospitalization data for respiratory diseases from August 2011 to July 2012, linked with estimated daily concentrations of air pollutants (CO, PM2.5, O3, NOx) and meteorological variables. The dataset contains 352 admissions and was used in a Poisson regression analysis to quantify health risks, finding a 17% increased hospitalization risk for a 3µg/m³ rise in nitrogen oxides. The research was conducted by Vanessa Villalta Lima Roman using data from Datasus and pollutant estimates from the CATT-BRAMS model.
18 recorded interactions between hearing parents and their children or adolescents with hearing loss were analyzed using a 22-behavior checklist. The data was collected by Laura Mochiatti Guijo from nine children and nine adolescents with bilateral sensorineural pre-lingual hearing impairment, all involved in an aural rehabilitation program. Interactions were scored on a Likert scale by three judges, achieving a 97.8% inter-rater agreement.
Research data on optimizing soaking and drying processes for ripe canistel fruit flour. The study, authored by Sri Rejeki Retna Pertiwi, determined sensory, phytochemical, physical, and chemical characteristics. Results identified optimal conditions as soaking in 7.5% NaCl for 30 minutes followed by drying at 40 °C for 6 hours.
A study dataset measuring sound pressure levels inside ambulances during emergency trips and full driver shifts. The data was collected and statistically analyzed by Rafaella Cristina Oliveira to assess compliance with Brazilian regulatory standards. Results indicate average noise levels exceeding 85 dB(A) and noise doses ranging from 17.51% to 155.68% of the recommended limit.
An experimental dataset for the Workflow Satisfiability Problem (WSP) with class-independent constraints, generated by a random instance generator described in a 2015 paper. The dataset includes instances stored in .wsp files and corresponding pseudo-Boolean satisfiability formulations in .opb files, alongside solution outputs from the SAT4J solver and a pattern-backtracking FPT algorithm. The dataset was created by Andrei Gagarin and colleagues for algorithm benchmarking.
Submitters to the STI-ENID Conference 2023 were asked to include an open science statement. Leo Waaijers of Leiden University collected these statements to analyze their compliance with open science principles. The statements were published openly on the conference website but vary widely in explanation, vocabulary, and detail.
A 2026 proof-of-concept reproduction of the REWIRE method for recycling low-quality web documents. It demonstrates using an LLM to rewrite discarded web documents, which are then re-scored by the original filter to keep improved versions. The dataset was created by davanstrien and is hosted on Hugging Face.
Two cannulated cattle were used to incubate soybean meal and cotton cake samples for up to 72 hours. Dry matter degradability for soybean meal ranged from 86.35% to 65.50% depending on passage rate, while cotton cake ranged from 53.44% to 35.21%. The study by Ubiara Henrique Gomes Teixeira evaluated five non-linear mathematical models to characterize the in situ degradation parameters.
A paper presents a capillary electrophoresis method for ultra-fast simultaneous determination of naphazoline with diphenhydramine, pheniramine, or chlorpheniramine. The method uses a 10 cm capillary column, achieving one analysis every 35 seconds with detection limits of 25 µmol L-1 for three compounds and 13 µmol L-1 for chlorpheniramine. Results were validated against high-performance liquid chromatography with no statistically significant differences at a 95% confidence level.
A framework for estimating expected possession value (EPV) in basketball using optical player tracking data, proposed by Daniel Cervone of New York University. The model is a multiresolution stochastic process differentiating between continuous player movements and discrete events like shots. A data sample and R code are provided in the supplementary material for model exploration.
The Kyoto 2006+ dataset provides real network traffic data collected from various honeypots between November 2006 and August 2009. Jungsuk Song from the National Institute of Information and Communications Technology created this dataset to address the outdated KDD Cup 99' dataset. It is intended to help researchers evaluate Network Intrusion Detection Systems (NIDSs) with data reflecting contemporary attack trends.
Two different films were sampled 40 times each to characterize their acoustic parameters in an impedance tube with an inner diameter of 44.44mm. Measurements span the [4, 4520] Hz band and include configurations with melamine and glass wool backings, which were also characterized using 3 and 2 samples respectively. The dataset was created by researchers including Mathieu Gaborit from the Centre National de la Recherche Scientifique and is released under a Creative Commons Attribution license.
Jamie Davis authored a C++ implementation of a real-time slew-rate limiter with variable rise and fall constraints. The algorithm is designed for bare-metal embedded systems to protect actuators from destructive transients. The code was uploaded to figshare on 2026-05-31 under a CC-BY-4.0 license.
Module 113 from the Davis Logic V2 project provides a zero-heap, zero-copy, real-time Resettable Integrator with Anti-Windup clamping logic. The component, authored by Jamie Davis and licensed under CC BY 4.0, is designed for safety-critical automated control loops on bare-metal systems to mitigate accumulator overruns and windup. It operates using Q16.16 fixed-point arithmetic and is stored as a 2.3 KB text file.
Module 111 is a C++ source code file for a zero-heap, zero-copy linear interpolator and extrapolator designed for bare-metal embedded systems. It was authored by Jamie Davis and released under a CC BY 4.0 license, with a last update recorded on 2026-05-31. The 3.7 KB file implements a Q16.16 fixed-point algorithm for real-time data projection and coordinate mapping.
Module 108 from the Davis Logic V2 project provides a zero-heap, real-time first-order lag/lead compensator filter designed for deterministic control loop stabilization in bare-metal embedded targets. The module, authored by Jamie Davis and licensed under CC BY 4.0, uses a recursive digital infrastructure with Q16.16 fixed-point arithmetic to optimize transient responses and modify phase margins. The text file, last updated on 2026-05-31, is 3.8 KB in size.
Southern Africa is the focus of this dataset used to validate the MODIS active fire product against higher-resolution ASTER imagery. NASA researchers employed 18 ASTER scenes from August to October 2001, applying logistic regression models to analyze detection probabilities across two MODIS algorithm versions. The data underpins a published 2005 study on improving satellite-based fire monitoring accuracy.
A sample of 309 college students provided keystroke log data through the cloud-based writing platform Clourite. The data was used to extract writing process features via a Bayesian hierarchical mixture model and to identify patterns differentiating stronger and weaker writers. The dataset, authored by Tingxuan Li, was last updated on 2026-05-28.
Kuan-Hsun Wu's research proposes a statistical functional and cumulative distribution function structure for estimating treatment effects like ATE and QTE. The 796.0 KB dataset, last updated in 2026, includes PDF, ZIP, and TXT files describing the methodology and numerical study results. The approach incorporates variable selection to identify informative and network structures in confounders.