Fetching the paper…
Reading the bibliography…
We introduce HoVer (HOppy VERification), a dataset for many-hop evidence extraction and fact verification.
Hierarchical graph network for multi-hop question answering
Yuwei Fang, Siqi Sun, Zhe Gan, Rohit Pillai, Shuohang Wang, and Jingjing Liu. 2019 · 1911
Earlier work this paper cites.
Local textual inference: can it be defined or circumscribed?
Annie Zaenen, Lauri Karttunen, and Richard Crouch. 2005 · 2005
Earlier work this paper cites.
Recognizing textual entailment: Rational, evaluation and approaches–erratum
Ido Dagan, Bill Dolan, Bernardo Magnini, and Dan Roth. 2010 · 2010
Earlier work this paper cites.
Fact checking: Task definition and dataset construction
Andreas Vlachos and Sebastian Riedel. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Emergent: a novel data-set for stance classification
William Ferreira and Andreas Vlachos. 2016 · 2016
Earlier work this paper cites.
Reading wikipedia to answer open-domain questions
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes. 2017 · 2017
Earlier work this paper cites.
“liar, liar pants on fire”: A new benchmark dataset for fake news detection
William Yang Wang. 2017 · 2017
Earlier work this paper cites.
Multi-hop inference for sentence-level TextGraphs: How challenging is meaningfully combining information for science question answering?
Peter Jansen. 2018 · 2018
Earlier work this paper cites.
WorldTree: A corpus of explanation graphs for elementary science questions supporting multi-hop inference
Peter Jansen, Elizabeth Wainwright, Steven Marmorstein, and Clayton Morrison. 2018 · 2018
Earlier work this paper cites.
Looking beyond the surface: A challenge set for reading comprehension over multiple sentences
Daniel Khashabi, Snigdha Chaturvedi, Michael Roth, Shyam Upadhyay, and Dan Roth. 2018 · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal. 2018 · 2018
Earlier work this paper cites.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Earlier work this paper cites.
The web as a knowledge-base for answering complex questions
Alon Talmor and Jonathan Berant. 2018 · 2018
Cited alongside, same era.
Fever: a large-scale dataset for fact extraction and verification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Cited alongside, same era.
Constructing datasets for multi-hop reading comprehension across documents
Johannes Welbl, Pontus Stenetorp, and Sebastian Riedel. 2018 · 2018
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W. Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Cited alongside, same era.
Multifc: A real-world multi-domain dataset for evidence-based fact checking of claims
Inherent disagreements in human textual inferences
Ellie Pavlick and Tom Kwiatkowski. 2019 · 2019
Later among the works it cites.
Answering complex open-domain questions through iterative query generation
Peng Qi, Xiaowen Lin, Leo Mehr, Zijian Wang, and Christopher D Manning. 2019 · 2019
Later among the works it cites.
Understanding the impact of text highlighting in crowdsourcing tasks
Jorge Ramírez, Marcos Baez, Fabio Casati, and Boualem Benatallah. 2019 · 2019
Later among the works it cites.
Towards debiasing fact verification models
Tal Schuster, Darsh J. Shah, Yun Jie Serene Yeo, Daniel Filizzola, Enrico Santus, and Regina Barzilay. 2019 · 2019
Later among the works it cites.
The fever2. 0 shared task
James Thorne, Andreas Vlachos, Oana Cocarascu, Christos Christodoulopoulos, and Arpit Mittal. 2019 · 2019
Later among the works it cites.
Learning to retrieve reasoning paths over wikipedia graph for question answering
Akari Asai, Kazuma Hashimoto, Hannaneh Hajishirzi, Richard Socher, and Caiming Xiong. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Isabelle Augenstein, Christina Lioma, Dongsheng Wang, Lucas Chaves Lima, Casper Hansen, Christian Hansen, and Jakob Grue Simonsen. 2019 · 2019
Cited alongside, same era.
Understanding dataset design choices for multi-hop reasoning
Jifan Chen and Greg Durrett. 2019 · 2019
Cited alongside, same era.
Seeing things from a different angle: Discovering diverse perspectives about claims
Sihao Chen, Daniel Khashabi, Wenpeng Yin, Chris Callison-Burch, and Dan Roth. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Avoiding reasoning shortcuts: Adversarial evaluation, training, and model development for multi-hop qa
Yichen Jiang and Mohit Bansal. 2019 · 2019
Cited alongside, same era.
Compositional questions do not necessitate multi-hop reasoning
Sewon Min, Eric Wallace, Sameer Singh, Matt Gardner, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2019 · 2019
Cited alongside, same era.
Revealing the importance of semantic retrieval for machine reading at scale
Yixin Nie, Songhe Wang, and Mohit Bansal. 2019b · 2019
Cited alongside, same era.
Closest in time.
Transformers as soft reasoners over language
P. Clark, Oyvind Tafjord, and Kyle Richardson. 2020 · 2020
Closest in time.
DeSePtion: Dual sequence prediction and adversarial examples for improved fact-checking
Christopher Hidey, Tuhin Chakrabarty, Tariq Alhindi, Siddharth Varia, Kriste Krstovski, Mona Diab, and Smaranda Muresan. 2020 · 2020
Closest in time.
Unsupervised question decomposition for question answering
Ethan Perez, Patrick Lewis, Wen-tau Yih, Kyunghyun Cho, and Douwe Kiela. 2020 · 2020
Closest in time.
Winogrande: An adversarial winograd schema challenge at scale
Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2020 · 2020
Closest in time.
WorldTree v2: A corpus of science-domain structured explanations and inference patterns supporting multi-hop inference
Zhengnan Xie, Sebastian Thiem, Jaycie Martin, Elizabeth Wainwright, Steven Marmorstein, and Peter Jansen. 2020 · 2020
Closest in time.
Transformer-xh: Multi-evidence reasoning with extra hop attention
Chen Zhao, Chenyan Xiong, Corby Rosset, Xia Song, Paul Bennett, and Saurabh Tiwary. 2020 · 2020
Closest in time.