arXiv citation graph is a dataset of citation relationships between academic papers hosted on the arXiv preprint server. The dataset is published on Kaggle. The specific scale, creation date, and authorship details are not provided in the available metadata.
Use Cases
- Analyzing the structure and evolution of scientific collaboration networks (inferred from domain, verify after download)
- Training graph neural networks for tasks like link prediction or node classification (inferred from domain, verify after download)
- Studying citation patterns and knowledge flow across academic fields (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a major platform for data science resources.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count is unknown, which may limit suitability assessment.