Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
CoDEx is a set of knowledge graph completion datasets extracted from Wikidata and Wikipedia, presented at EMNLP 2020. It comprises three knowledge graphs varying in size and structure, includes multilingual entity descriptions, and provides tens of thousands of hard negative triples. The benchmark was created by Tara Safavi of the University of Michigan to improve upon existing datasets in scope and difficulty.
License is listed as Open Access (green), but specific terms should be verified on the source repository.