Loading...
Loading...
Mathematical datasets, statistical benchmarks, probability, optimization, operations research
3,080 datasets
A methodological paper and associated data propose a novel approach for summarizing posterior inference in nonparametric Bayesian mixture models. The work, authored by Khai Nguyen and available on figshare, was last updated in April 2026. It introduces two variants of sliced Wasserstein distance for Gaussian mixtures to estimate mixing measures and validate partitions.
An anonymized and cleaned dataset used for a statistical analysis study. The data includes all variables collected from survey responses and is provided as a 17.4 KB XLSX file. It was authored by Abdullrahman Mohammed Alshehri and last updated on May 21, 2026.
79.9 KB of tabular data from figshare, authored by Ilju Yang and last updated on 2026-04-13. The dataset contains measurements of male damselfly morphology, including tibia area and abdomen length, alongside daily pairing success records and environmental variables like temperature and UV level.
A Web Map Service (WMS) provides regional statistical data for the districts of Hamburg and the city-state as a whole. The data is published by the German Federal Agency for Cartography and Geodesy (Bundesamt für Kartographie und Geodäsie). The specific variables, update frequency, and temporal coverage are not detailed in the provided metadata.
Fit statistics for a final model predicting the effects of soil iron content, salinity, temperature, and moisture levels. The dataset includes indicators for statistical significance at α = 0.05. Authored by Kanokporn Chaianunporn and last updated on 2026-05-18.
9.5 KB of statistical model fit data for predicting the effects of soil carbon-to-nitrogen ratio and salinity under different temperature and moisture levels. The dataset was authored by Kanokporn Chaianunporn and last updated on May 18, 2026. Asterisks in the data indicate statistical significance at α = 0.05.
Fit of the elements in the final model predicting effects of pH and salinity under different temperatures and moisture levels. The dataset, authored by Kanokporn Chaianunporn, is a 5.5 KB XLS file last updated on 2026-05-18. Asterisks in the data indicate statistical significance at α = 0.05.
Pragya Singhal published a dataset on figshare containing psychological health and WHOQOL scores. The data includes mean and standard deviation scores for Physical, Psychological, Social, and Environment domains across three groups: Cases, caregivers, and college students. It also contains Kruskal Wallis statistical values for comparing these groups.
A 2026 dataset by Baljit Kaur on figshare details the structure-guided optimization of CHI3L1-binding small molecules for Alzheimer's disease. It includes a structure-activity map for 24 prioritized derivatives, with biophysical data such as binding affinity (Kd) ranging from 45 μM to 236 μM. The dataset identifies G721-0377 as a lead compound with improved affinity and functional efficacy in reversing astrocytic dysfunction.
Structure-activity relationship data for 24 small-molecule derivatives of the CHI3L1-binding compound G721-0282, optimized for Alzheimer's disease research. The dataset includes biophysical binding affinity measurements, with the lead compound G721-0377 showing a Kd of 45 μM. It was created by Baljit Kaur and published on figshare in April 2026.
24 prioritized derivatives of the CHI3L1-binding molecule G721-0282 were generated through virtual screening and structure-guided optimization. The dataset includes biophysical analysis identifying lead compound G721-0377, which has a binding affinity (Kd) of 45 μM. The data was authored by Baljit Kaur and last updated on 2026-04-28.
Trend data from the NSW Bureau of Crime Statistics and Research, last updated in May 2026. This table provides information on short (2-year), medium (5-year), and long (10-year) term trends in incidents of use/possess amphetamines and cocaine offences recorded by the NSW Police Force. The data is organized by region (Statistical Area).
Sediment sources to the Fitzroy River coastal zone have been identified and quantified using an integrated geochemical and modeling approach. The study, likely hosted by the Australian Ocean Data Network, found that the proportion of basaltic material deposited in the coastal zone has increased in recent time and is now the dominant catchment source. The data likely reflects changes in catchment sediment sources over time, influenced by rainfall events and land-use changes following European settlement.
Simulation results from a study investigating a prescribed-time trajectory tracking control strategy for unmanned surface vehicles (USVs) under stochastic environmental loads. The dataset, 401.1 KB in size, was authored by Ruye Cong and shared under a CC-BY-4.0 license on figshare in May 2026. It was used to analyze and demonstrate that the proposed control scheme can achieve trajectory tracking within a prescribed time.
Five distinct seabed sediment classes were identified in Keppel Bay, Central Queensland, using statistical techniques. The classification is based on sediment grainsize, chemical composition, and modelled seabed shear stress from waves and tidal currents. Data was collected by the Australian Ocean Data Network through sediment sampling and acoustic seabed mapping.
A 5.5 KB dataset by Lukas Waltenberger, last updated on 2026-05-05, contains performance measures for binary logistic regression and Bayesian models. The analysis employed post-hoc filtering, excluding subjects with a sex prediction probability below 75% to refine model evaluation. The dataset is shared under a CC-BY-4.0 license on figshare.
A methodological guide for converting common statistical summary measures like P-values, LSD, and CI into variance estimates suitable for meta-analysis. The resource, authored by David LeBauer and hosted on Papers with Code, focuses on agricultural and biological research contexts. It provides rules for prioritizing direct variance estimates and transforming range statistics.
Zahid Ullah from Changwon National University presents a method for processing brain MRI images. The methodology uses histogram equalization for contrast enhancement and mathematical morphology for skull stripping, implemented in MATLAB R2015a. Results were evaluated using Mean Square Error and Peak Signal to Noise Ratio metrics.
A 50,000-step cyber-physical simulation dataset for overhead transmission lines, created by Song and published in 2026. It includes baseline physical parameters, dynamic load fluctuations, environmental disturbances, and results from algorithm comparisons. The data supports research on monitoring algorithms and digital twin applications under stochastic conditions.
Karanvir Singh designed and synthesized a series of triazole-linked sesamol conjugates as potential antifungal agents. The dataset includes in vitro and in vivo efficacy data, such as MIC and MFC values, for these compounds against pathogenic Candida strains. It was last updated on April 28, 2026, and is shared via figshare under a CC-BY-NC-4.0 license.