Loading...
Loading...
Social network graphs, knowledge graphs, citation networks, molecular graphs for GNN, web link graphs
426 datasets
Per-logbook claim verdicts produced by the logbook-judge Space. The dataset was authored by ICML-2026-agent-repro and last updated on July 17, 2026. It is hosted on Hugging Face and the data is contained within a verdicts.json file.
Two California traffic datasets, PeMSD4 and PeMSD8, used to benchmark a Transformer-based Hypergraph Convolutional Network for traffic flow prediction. The 5.5 KB XLS files were published by SiWei Wei on figshare under a CC-BY-4.0 license and last updated in April 2026. Experiments described in the metadata involved 5 independent random seed runs to achieve statistically significant performance improvements on core metrics like MAE, RMSE, and MAPE.
RelBench hosts the Temporal Graph Benchmark datasets converted to its manifest format. The repository contains multiple sub-datasets, each with a self-describing manifest, parquet tables, and task-specific labels. The data was last updated on 2026-06-12.
A knowledge graph-Bayesian network-driven suspect screening (KGBS) strategy integrates text mining, probabilistic reasoning, and high-resolution mass spectrometry. The framework extracted 54,838 associations between phthalate esters and their metabolites from 3,167 publications to construct a Bayesian inference-embedded knowledge graph. Applied to paired indoor dust and human urine samples, the KGBS platform identified 68 PAEs and 49 metabolites, including 18 PAEs and 14 metabolites newly annotated in the study.
Beamtime ES-1145 contains all raw X-ray diffraction (XRD) data from an experiment conducted at the European Synchrotron Radiation Facility (ESRF). The dataset includes a logbook describing samples, experimental conditions, and scan numbers. It was authored by ZHAO BIN of the Centre National de la Recherche Scientifique and is available under an Open Access license.
A 2.8 MB research dataset from figshare, last updated in April 2026, accompanies a paper proposing the SpecMBA method for community detection. The data likely contains multi-layer network adjacency matrices used to validate the proposed spectral method with moment integration and bias adjustment. Author Xuefei Wang released the dataset under a CC-BY-4.0 license.
56.5 MB of research data from a study investigating the role of Acetyl Tributyl Citrate in idiopathic pulmonary fibrosis. The dataset includes results from computational toxicology, machine learning on transcriptome data, molecular docking simulations, and a pilot experiment in bleomycin-induced fibrotic mice. It was authored by Dajia Fu and last updated on 2026-05-15.
Supporting materials for the paper 'Deep Learning of Subsurface Flow via Theory-guided Neural Network' authored by Nanzhe Wang. The data likely contains simulation results or measurements for modeling fluid flow in porous media. Its specific size, format, and row count are not detailed in the provided metadata.
A machine-readable identity graph representing the professional, research, creative, and digital identity infrastructure of Hamed Behrouzi. The dataset contains structured identity nodes, semantic relationships, linked profiles, publications, professional credits, research outputs, and provenance metadata. It is maintained as a living research artifact by Hamed Behrouzi and was last updated on June 11, 2026.
MONDO-DOID Entity Alignment 12K is a biomedical benchmark for aligning disease entities between the Mondo Disease Ontology (MONDO) and the Human Disease Ontology (DOID). It contains 12,000 manually curated exact-match mappings distributed by MONDO in SSSOM format. The dataset was created by vaibhavalakshmiravideshik and was last updated on 2026-05-29.
83 metro fire incident case files and Python source code reconstruct a knowledge graph with 792 nodes and 2,061 edges. The dataset supports a paper on constructing a safety risk assessment indicator system using knowledge graphs and fuzzy AHP. It was authored by Biao Ma and last updated on April 30, 2026.
MONDO-DOID Entity Alignment 12K is a biomedical entity alignment benchmark for heterogeneous knowledge graphs. It aligns disease entities between the Mondo Disease Ontology (MONDO) and the Human Disease Ontology (DOID) using manually curated exact-match mappings distributed by MONDO in SSSOM format. The dataset was authored by vaibhavalakshmiravideshik and last updated on Hugging Face in May 2026.
A 12.3 KB PDF document from Ladoke Akintola University of Technology (LAUTECH) in Ogbomoso, Nigeria. The file likely contains information and application forms for Direct Entry, Postgraduate Diploma (PGD), and Diploma programs for the 2025/26 academic year. It was uploaded by 'School Mail' on figshare in May 2026.
Osogbo, Southwestern Nigeria is the geographic scope of this dataset. It likely contains survey responses on patient attitudes, beliefs, and knowledge regarding blood transfusion practices. The data was published by Iosr Journals on the paperswithcode platform.
Pure PyTorch implementations of the ALIGNN and ALIGNN-FF models, released under a CC-BY-4.0 license. The 384.0 MB ZIP file contains machine learning models for predicting material properties, authored by Kamal Choudhary. It was last updated on May 20, 2026.
Cross-Medicine Knowledge Graph (CMKG) data source distribution and contributions. The dataset is a 5.5 KB Excel file authored by Zekun Zhou and last updated on May 13, 2026. It is licensed under CC-BY-4.0 and hosted on figshare.
Viacheslav Dubovitskii's repository provides code and a Jupyter notebook for constructing discrete-time quantum walk circuits. The workflow covers graph-based quantum walk formulation, symmetry-aware circuit construction, and transpilation for superconducting quantum devices. The dataset was last updated on 2026-05-08.
1,324 metabolites from 12 classes were measured across eight developmental stages and fifteen tissues of the 'Hongyingzi' waxy sorghum landrace. The dataset provides a spatiotemporal map of metabolic changes, constructed by Jibin Wang and published in 2026. It integrates metabolomic and transcriptomic data to build a regulatory network.
A heterogeneous graph dataset integrating Chinese Materia Medica, compounds, protein targets, and adverse drug reactions for pharmacovigilance prediction. The dataset, authored by Bowen Shi and last updated in March 2026, contains 27,062 curated CMM-ADR associations used to train and evaluate the MSAT graph neural network model.
5.5 KB of metrics describing global network behavior over time. The dataset includes DTDG metrics on total edges, average node connections, network profile, and the percentage of nodes in crisis. It was authored by Marcus Araujo and last updated on April 22, 2026.