Loading...
Loading...
Social network graphs, knowledge graphs, citation networks, molecular graphs for GNN, web link graphs
428 datasets
Software Heritage is the largest existing public archive of software source code and development history. The dataset is a fully deduplicated Merkle DAG representation linking file content, directories, commits, and repository states from major forges, distributions, and package managers. Author and committer information is anonymized.
Justin Goldston published this digital humanities resource in March 2026 to map relationships between biblical passages and thematic classifications. It contains a machine-readable scripture corpus alongside a cross-reference network graph designed for computational theology and knowledge graph construction.
GNN Vect GIN Ver2 is a dataset likely related to Graph Neural Networks, specifically the Graph Isomorphism Network (GIN) architecture. It is hosted on the Kaggle platform, but detailed metadata about its contents, size, and origin are not provided. The dataset's specific purpose and collection methodology must be verified after download.
1970 radar survey data records ice depth measurements on the Amery Ice Shelf. The Australian Antarctic Division collected this data using radar equipment mounted in a towed van. Original logbooks detailing equipment settings are archived by the Australian Antarctic Division.
Lampung Selatan Road Network Dataset is a graph representation of the road network in the South Lampung region of Indonesia, derived from OpenStreetMap. The dataset is published on Kaggle and is intended for routing and pathfinding applications. The specific scale, update date, and author are not provided in the available metadata.
An image dataset likely containing pictures of cats and dogs. The dataset is hosted on Kaggle, but its specific size, source, and creation details are not provided. The author and organization are unknown.
An evaluation dataset probing 18 Knowledge Graph-style reasoning tasks on the Qwen/Qwen3.5-2B-Base model. It was created by chayma-rhaiem and last updated on March 8, 2026. The dataset tests the model in its raw base form across parametric memory, standard grounded reasoning, and advanced grounded reasoning tasks.
Graph-TCGA-BRCA is a graph-level classification dataset derived from the TCGA-BRCA histopathology dataset. Each 224x224 patch image is converted into a cell-graph where nodes represent detected cell nuclei and edges encode spatial proximity. The dataset was created by author ogutsevda and last updated on 2026-03-03.
Kaggle dataset titled 'news-gnn-v2-filtered', likely containing news articles structured for graph neural network applications. The dataset's specific content, size, and origin are not detailed in the provided metadata. Its availability on Kaggle suggests it is intended for machine learning experimentation.
Graph-PanNuke is a node-level classification dataset derived from the PanNuke pan-cancer histology dataset, using all slides at 40× magnification. Each tissue patch is converted into a cell-graph where nodes represent detected cell nuclei, with the task of predicting cell type across 5 classes. Node features describe cell morphology and texture.
466 calendar years of tree ring data from Douglas fir specimens in Colorado, USA, covering the period from 466 to -44 years before present. The dataset is archived by NOAA's National Centers for Environmental Information as part of the Paleoclimatology World Data Service. The data was last updated in 1994.
News-GNN is a dataset hosted on Kaggle, likely containing news articles structured for graph-based machine learning. The dataset's specific content, size, and origin are not detailed in the provided metadata. Its platform tags indicate it is intended for applications involving Graph Neural Networks, text, and graph data.
A dataset titled 'exp73_gnn_lstm_fusion' was published on Kaggle. Its specific content and scale are unknown from the provided metadata. The title suggests it relates to an experiment combining Graph Neural Networks and Long Short-Term Memory architectures.
1977 logbooks record details of ice core drilling at site BHQ on Law Dome. The collection consists of 3 books containing stratigraphy information for some core segments. The Australian Antarctic Division (AU_AADC) archived a hard copy of this document.
Logbooks document geophysical and meteorological measurements taken during a traverse across Law Dome and Wilkes Land in Antarctica. The records were created by the Australian Antarctic Division during the 1981 field season. They contain readings for snow accumulation, barometric pressure, gravity, temperature, wind, and oxygen isotopes.
Logbooks from field work in Enderby Land between 1972 and 1980. The records contain observations of borehole temperatures, ice movement, gravity, ice radar notes, and barometric pressure. Physical copies are stored by the Australian Antarctic Division.
Summer 1970 records detail the daily activities and challenges faced by an Australian Antarctic Division traverse team on the Amery Ice Shelf. The dataset consists of a handwritten or typed logbook from the expedition. It was created by the Australian Antarctic Division and archived in 1970.
GNNModule is a dataset published on Kaggle, likely related to Graph Neural Networks. The dataset's specific content, size, and origin are not detailed in the available metadata. Users must download the dataset to verify its exact structure and potential applications in network science.
1985 data from the ADBEX III Antarctic voyage contains analyzed water density results from sea ice and snow samples. The Australian Antarctic Division (AU_AADC) collected and archived the results in physical logbooks. The dataset includes related records on oxygen isotope samples.
An identity graph dataset published on Kaggle. The dataset's title suggests it contains network data linking identities, likely for Q1 of an unspecified year. Metadata is minimal; actual content requires verification after download.