Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Welbl, Johannes; Stenetorp, Pontus; Riedel, Sebastian

Computer Science > Computation and Language

arXiv:1710.06481 (cs)

[Submitted on 17 Oct 2017 (v1), last revised 11 Jun 2018 (this version, v2)]

Title:Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Authors:Johannes Welbl, Pontus Stenetorp, Sebastian Riedel

View PDF

Abstract:Most Reading Comprehension methods limit themselves to queries which can be answered using a single sentence, paragraph, or document. Enabling models to combine disjoint pieces of textual evidence would extend the scope of machine comprehension methods, but currently there exist no resources to train and test this capability. We propose a novel task to encourage the development of models for text understanding across multiple documents and to investigate the limits of existing methods. In our task, a model learns to seek and combine evidence - effectively performing multi-hop (alias multi-step) inference. We devise a methodology to produce datasets for this task, given a collection of query-answer pairs and thematically linked documents. Two datasets from different domains are induced, and we identify potential pitfalls and devise circumvention strategies. We evaluate two previously proposed competitive models and find that one can integrate information across documents. However, both models struggle to select relevant information, as providing documents guaranteed to be relevant greatly improves their performance. While the models outperform several strong baselines, their best accuracy reaches 42.9% compared to human performance at 74.0% - leaving ample room for improvement.

Comments:	This paper directly corresponds to the TACL version (this https URL) apart from minor changes in wording, additional footnotes, and appendices
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:1710.06481 [cs.CL]
	(or arXiv:1710.06481v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1710.06481
Journal reference:	Transactions of the Association for Computational Linguistics (TACL), Vol 6 (2018), pages 287-302

Submission history

From: Johannes Welbl [view email]
[v1] Tue, 17 Oct 2017 19:35:07 UTC (736 KB)
[v2] Mon, 11 Jun 2018 17:08:20 UTC (857 KB)

Computer Science > Computation and Language

Title:Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators