Loading...
Loading...
General ML benchmarks, tabular data, AutoML, recommendation systems, anomaly detection, evaluation suites
192,488 datasets
Henry S. Rzepa from Imperial College London published this dataset on paperswithcode. The title suggests it contains data related to a transition state intrinsic reaction coordinate (TS IRC) calculation for a chemical reaction. The specific content and scale of the data require verification after download.
A dataset of quantum chemical calculations for the HCOH molecule, likely formaldehyde. The data includes energy values calculated using the wB97XD functional and Def2-TZVPP basis set, with reported Gibbs free energy and energy difference values. It was published by Henry Rzepa of Imperial College London on the Papers with Code platform under an Open Access license.
The Redcar and Cleveland WFS Service provides geographically maintained or monitored areas in the borough of Redcar and Cleveland. The service includes numerous layers representing the extents of environmental, historical, and regulated areas managed by Redcar and Cleveland Borough Council. It was last updated on 2026-07-08.
Triple-H with methanol presents computational chemistry data from a specific quantum chemical calculation. Published by Henry Rzepa of Imperial College London on Papers with Code, the dataset likely contains molecular energies and properties. The title indicates the use of the MN15L+G(d,p) functional and Lanl2dz basis set with a solvent model for methanol.
Henry S. Rzepa from Imperial College London published this dataset on paperswithcode. The title suggests it contains data related to a Hetero-Diels-Alder (HDA) reaction mechanism, specifically transition state (TS) calculations with a p=NH2 substituent, likely analyzed via Intrinsic Reaction Coordinate (IRC). The dataset is licensed for Open Access.
Thomas Heinis of Imperial College London authored a paper titled 'Neuromorphic Hardware As Database Co-Processors: Potential and Limitations'. The paper likely discusses the application of neuromorphic computing architectures to database operations. It is published on paperswithcode and is under an Open Access (green) license.
Replication package for the paper 'LLHD: A Multi-level Intermediate Representation for Hardware Description Languages' by Fabian Schuiki of ETH Zurich. The dataset likely contains artifacts related to the LLHD intermediate representation, which may include code, benchmarks, or specifications. The specific contents, size, and format are not detailed in the available metadata.
Gustavo Alonso from ETH Zurich authored this paper on data processing in modern hardware. The dataset likely contains performance metrics and architectural details relevant to computing systems. It is published on paperswithcode under an Open Access license.
A replication package associated with a machine learning article authored by Gagleen Singh from ETH Zurich. The package likely contains data, code, or models necessary to reproduce the study's results. Its specific contents, size, and structure are not detailed in the available metadata.
Artifact for the PLDI 2018 paper 'Bayonet', authored by Timon Gehr of ETH Zurich. The dataset is a research artifact published on paperswithcode, likely containing data related to the paper's experiments. Its specific content, such as row count or file format, is not detailed in the available metadata.
A replication package from the paperswithcode platform, associated with an article authored by Tobias Gysi of ETH Zurich. The package likely contains data and code to reproduce the results of a specific machine learning study. Its exact contents, scale, and structure are unspecified in the available metadata.
A nomenclature document defining terms for transport phenomena in electrolytic systems. The work was authored by N. Ibl and is associated with ETH Zurich. The dataset is hosted on the paperswithcode platform.
An investigation of equilibria authored by G. Anderegg from ETH Zurich. The dataset is published on the paperswithcode platform and is tagged with topics including Machine Learning, Equilibrium, and Physics. The specific data format, size, and content details are not provided in the metadata.
FinQA_TAT-QA_financial_finetuning_dataset provides a unified format for question answering over financial documents combining tabular and textual data. The dataset, created by hellotayssir, is built to support training and evaluating models on numerical and discrete reasoning tasks in finance. It draws on the structure and style of established finance-QA benchmarks such as TAT-QA and FinQA.
A 0.02° x 0.02° gridded map of sea surface temperature (SST) for the Southern Ocean region from 3°E to 158°W and 27°S to 78°S. The product, created by the Australian Ocean Data Network, provides one-month averages of SST derived from NOAA AVHRR satellite observations, with a reported 2014 bias of less than 0.03°C and standard deviation of 0.6°C. Production of this single-sensor product ceased in June 2025.
RoboVista is a benchmark of 474 multiple-choice questions for evaluating vision-language models on robot perception and reasoning skills. The dataset pairs images from real robot deployments with answer choices and ground-truth answers. It was created by sy-xie and last updated on July 6, 2026.
songyiren's dataset contains full generation and evaluation results from running the Visual Reasoning Benchmark Suite v3 on two image generation models. The benchmark includes 1,705 items across 10 tasks such as figure completion, spatial generation, mazes, Sudoku, board games, matchsticks, orthographic views, and visual math proofs. Results were judged by the gemini-3.1-pro-preview model and the dataset was last updated on July 8, 2026.
Leaf functional trait measurements describe leaf structure, chemistry, and metabolism. Data was collected from the Alice Mulga site in 2014. The dataset is provided by the Terrestrial Ecosystem Research Network's Data Discovery platform.
919,869 English chat and agentic conversations are formatted in ChatML (Anthropic-style) Messages. Every assistant turn begins with a short, dense chain-of-thought reasoning tag (<reasoning>) that leads to the final answer. The dataset was created by GasaiAI and is built on top of the curated answers from HuggingFaceTB/smoltalk2, with a last update timestamp of 2026-06-21.
Rewina Tilahun Gessese provides annual under-five stunting prevalence data for Ethiopia from 2000 to 2024, sourced from the WHO Global Health Observatory. The dataset includes forecasts for 2025-2030 generated using time series models, with the Exponential Smoothing (ETS) model selected as the best performer. The data was last updated on April 22, 2026.