Fetching the paper…
Reading the bibliography…
Explainable multi-hop question answering (QA) not only predicts answers but also identifies rationales, i.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Multi-hop question answering via reasoning chains
Jifan Chen, Shih-ting Lin, and Greg Durrett. 2019 · 1910
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Wt5?! training text-to-text models to explain their predictions
Sharan Narang, Colin Raffel, Katherine Lee, Adam Roberts, Noah Fiedel, and Karishma Malkan. 2020 · 2004
Earlier work this paper cites.
A survey on explainability in machine reading comprehension
Mokanarangan Thayaparan, Marco Valentino, and André Freitas. 2020 · 2010
Earlier work this paper cites.
MCTest: A challenge dataset for the open-domain machine comprehension of text
Matthew Richardson, Christopher J.C. Burges, and Erin Renshaw. 2013 · 2013
Earlier work this paper cites.
Rationalizing neural predictions
Tao Lei, Regina Barzilay, and Tommi Jaakkola. 2016 · 2016
Earlier work this paper cites.
RACE: Large-scale ReAding comprehension dataset from examinations
Guokun Lai, Qizhe Xie, Hanxiao Liu, Yiming Yang, and Eduard Hovy. 2017 · 2017
Earlier work this paper cites.
Looking beyond the surface: A challenge set for reading comprehension over multiple sentences
Daniel Khashabi, Snigdha Chaturvedi, Michael Roth, Shyam Upadhyay, and Dan Roth. 2018 · 2018
Earlier work this paper cites.
FEVER: a large-scale dataset for fact extraction and VERification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
Constructing datasets for multi-hop reading comprehension across documents
Johannes Welbl, Pontus Stenetorp, and Sebastian Riedel. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Earlier work this paper cites.
Interpretable neural predictions with differentiable binary variables
Jasmijn Bastings, Wilker Aziz, and Ivan Titov. 2019 · 2019
Earlier work this paper cites.
Understanding dataset design choices for multi-hop reasoning
Jifan Chen and Greg Durrett. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Compositional questions do not necessitate multi-hop reasoning
Sewon Min, Eric Wallace, Sameer Singh, Matt Gardner, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2019 · 2019
Cited alongside, same era.
Answering complex open-domain questions through iterative query generation
Peng Qi, Xiaowen Lin, Leo Mehr, Zijian Wang, and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Multi-hop reading comprehension across multiple documents by reasoning over heterogeneous graphs
Ming Tu, Guangtao Wang, Jing Huang, Yun Tang, Xiaodong He, and Bowen Zhou. 2019 · 2019
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Later among the works it cites.
Unsupervised alignment-based iterative evidence retrieval for multi-hop question answering
Vikas Yadav, Steven Bethard, and Mihai Surdeanu. 2020 · 2020
Later among the works it cites.
WinoWhy: A deep diagnosis of essential commonsense knowledge for answering Winograd schema challenge
Hongming Zhang, Xinran Zhao, and Yangqiu Song. 2020 · 2020
Later among the works it cites.
Towards interpretable natural language understanding with explanations as latent variables
Wangchunshu Zhou, Jinyi Hu, Hanlin Zhang, Xiaodan Liang, Maosong Sun, Chenyan Xiong, and Jian Tang. 2020 · 2020
Later among the works it cites.
Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Quick and (not so) dirty: Unsupervised selection of justification sentences for multi-hop question answering
Vikas Yadav, Steven Bethard, and Mihai Surdeanu. 2019 · 2019
Cited alongside, same era.
Learning variational word masks to improve the interpretability of neural text classifiers
Hanjie Chen and Yangfeng Ji. 2020 · 2020
Cited alongside, same era.
ERASER: A benchmark to evaluate rationalized NLP models
Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric Lehman, Caiming Xiong, Richard Socher, and Byron C. Wallace. 2020 · 2020
Cited alongside, same era.
Hierarchical graph network for multi-hop question answering
Yuwei Fang, Siqi Sun, Zhe Gan, Rohit Pillai, Shuohang Wang, and Jingjing Liu. 2020 · 2020
Cited alongside, same era.
Why do you think that? exploring faithful sentence-level rationales without supervision
Max Glockner, Ivan Habernal, and Iryna Gurevych. 2020 · 2020
Cited alongside, same era.
A simple yet strong pipeline for HotpotQA
Dirk Groeneveld, Tushar Khot, Mausam, and Ashish Sabharwal. 2020 · 2020
Cited alongside, same era.
SpanBERT: Improving pre-training by representing and predicting spans
Mandar Joshi, Danqi Chen, Yinhan Liu, Daniel S. Weld, Luke Zettlemoyer, and Omer Levy. 2020 · 2020
Cited alongside, same era.
Mor Geva, Daniel Khashabi, Elad Segal, Tushar Khot, Dan Roth, and Jonathan Berant. 2021 · 2021
Later among the works it cites.
Do multi-hop question answering systems know how to answer the single-hop sub-questions?
Yixuan Tang, Hwee Tou Ng, and Anthony Tung. 2021 · 2021
Later among the works it cites.
Teach me to explain: A review of datasets for explainable natural language processing
Sarah Wiegreffe and Ana Marasovic. 2021 · 2021
Later among the works it cites.
Exploiting reasoning chains for multi-hop science question answering
Weiwen Xu, Yang Deng, Huihui Zhang, Deng Cai, and Wai Lam. 2021 · 2021
Later among the works it cites.
Distantly-supervised dense retrieval enables open-domain question answering without evidence annotation
Chen Zhao, Chenyan Xiong, Jordan Boyd-Graber, and Hal Daumé III. 2021 · 2021
Later among the works it cites.
Diagnostics-guided explanation generation
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma, and Isabelle Augenstein. 2022 · 2022
Later among the works it cites.
From easy to hard: Two-stage selector and reader for multi-hop question answering
Xin-Yi Li, Wei-Jun Lei, and Yu-Bin Yang. 2022 · 2022
Later among the works it cites.
MuSiQue: Multihop questions via single-hop question composition
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022 · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, brian ichter, Fei Xia, Ed H. Chi, Quoc V Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Crepe: Open-domain question answering with false presuppositions
Xinyan Velocity Yu, Sewon Min, Luke Zettlemoyer, and Hannaneh Hajishirzi. 2022 · 2022
Later among the works it cites.