WikiGraphs

Dataset Information
Modalities
Texts, Graphs
Languages
English
Introduced
2021
Homepage

Overview

WikiGraphs is a dataset of Wikipedia articles each paired with a knowledge graph, to facilitate the research in conditional text generation, graph generation and graph representation learning. Existing graph-text paired datasets typically contain small graphs and short text (1 or few sentences), thus limiting the capabilities of the models that can be learned on the data.

WikiGraphs is collected by pairing each Wikipedia article from the established WikiText-103 benchmark with a subgraph from the Freebase knowledge graph. This makes it easy to benchmark against other state-of-the-art text generative models that are capable of generating long paragraphs of coherent text. Both the graphs and the text data are of significantly larger scale compared to prior graph-text paired datasets.

Variants: WikiGraphs

Associated Benchmarks

This dataset is used in 1 benchmark:

  • KG-to-Text Generation -

Recent Benchmark Submissions

Task Model Paper Date
KG-to-Text Generation Unconditional WikiGraphs: A Wikipedia Text - … 2021-07-20
KG-to-Text Generation BoW WikiGraphs: A Wikipedia Text - … 2021-07-20
KG-to-Text Generation GNN WikiGraphs: A Wikipedia Text - … 2021-07-20
KG-to-Text Generation Nodes WikiGraphs: A Wikipedia Text - … 2021-07-20

Research Papers

Recent papers with results on this dataset: