Fetching the paper…
Reading the bibliography…
Our goal, in the context of open-domain textual question-answering (QA), is to explain answers by showing the line of reasoning from what is known to the answer, rather than simply showing a fragment of textual evidence (a "rationale'").
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Learning-to-rank with BERT in TF-ranking
Shuguang Han, Xuanhui Wang, Mike Bendersky, and Marc Najork. 2020 · 2004
Earlier work this paper cites.
Bleurt: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, and Ankur P Parikh. 2020 · 2004
Earlier work this paper cites.
Teaching machine comprehension with compositional explanations
Qinyuan Ye, Xiaozhen Huang, and Xiang Ren. 2020 · 2005
Earlier work this paper cites.
Generative language modeling for automated theorem proving
Stanislas Polu and Ilya Sutskever. 2020 · 2009
Earlier work this paper cites.
The seventh pascal recognizing textual entailment challenge
L. Bentivogli, Peter Clark, I. Dagan, and Danilo Giampiccolo. 2011 · 2011
Earlier work this paper cites.
Recognizing Textual Entailment: Models and Applications
Ido Dagan, Dan Roth, Mark Sammons, and Fabio Massimo Zanzotto. 2013 · 2013
Earlier work this paper cites.
Benchmarking applied semantic inference: The pascal recognising textual entailment challenges
Roy Bar-Haim, I. Dagan, and Idan Szpektor. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Bidirectional attention flow for machine comprehension
Minjoon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi. 2016 · 2016
Earlier work this paper cites.
Natural language inference from multiple premises
Alice Lai, Yonatan Bisk, and J. Hockenmaier. 2017 · 2017
Earlier work this paper cites.
Think you have solved question answering? Try ARC, the AI2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord. 2018 · 2018
Cited alongside, same era.
Transforming question answering datasets into natural language inference datasets
Dorottya Demszky, Kelvin Guu, and Percy Liang. 2018 · 2018
Cited alongside, same era.
Training classifiers with natural language explanations
Braden Hancock, Paroma Varma, Stephanie Wang, Martin Bringmann, Percy Liang, and Christopher Ré. 2018 · 2018
Cited alongside, same era.
WorldTree: A corpus of explanation graphs for elementary science questions supporting multi-hop inference
Peter A. Jansen, Elizabeth Wainwright, Steven Marmorstein, and Clayton T. Morrison. 2018 · 2018
Cited alongside, same era.
Fever: a large-scale dataset for fact extraction and verification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Proc. 3rd Workshop on Fact Extraction and Verification . ACL
Christos Christodoulopoulos, James Thorne, Andreas Vlachos, Oana Cocarascu, and Arpit Mittal, editors. 2020 · 2020
Later among the works it cites.
R4C: A benchmark for evaluating RC systems to get the right answer for the right reason
N. Inoue, Pontus Stenetorp, and Kentaro Inui. 2020 · 2020
Later among the works it cites.
Learning to explain: Datasets and models for identifying valid reasoning chains in multihop question-answering
Harsh Jhamtani and P. Clark. 2020 · 2020
Later among the works it cites.
UnifiedQA: Crossing format boundaries with a single QA system
Daniel Khashabi, Sewon Min, Tushar Khot, Ashish Sabharwal, Oyvind Tafjord, Peter Clark, and Hannaneh Hajishirzi. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, M. Matena, Yanqi Zhou, W. Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W Cohen, Ruslan Salakhutdinov, and Christopher D Manning. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
ERASER: A benchmark to evaluate rationalized NLP models
Jay DeYoung, Sarthak Jain, Nazneen Rajani, E. Lehman, Caiming Xiong, R. Socher, and Byron C. Wallace. 2019 · 2019
Cited alongside, same era.
Explanation in artificial intelligence: Insights from the social sciences
T. Miller. 2019 · 2019
Cited alongside, same era.
Finding generalizable evidence by learning to convince Q&A models
Ethan Perez, Siddharth Karamcheti, Rob Fergus, Jason Weston, Douwe Kiela, and Kyunghyun Cho. 2019 · 2019
Cited alongside, same era.
Explain yourself! Leveraging language models for commonsense reasoning
Nazneen Rajani, B. McCann, Caiming Xiong, and R. Socher. 2019 · 2019
Cited alongside, same era.
PRover: Proof generation for interpretable reasoning over rules
Swarnadeep Saha, Sayan Ghosh, Shashank Srivastava, and Mohit Bansal. 2020 · 2020
Later among the works it cites.
Learning to prove theorems by learning to generate theorems
Ming-Zhe Wang and Jun Deng. 2020 · 2020
Later among the works it cites.
WorldTree V2: A corpus of science-domain structured explanations and inference patterns supporting multi-hop inference
Zhengnan Xie, Sebastian Thiem, Jaycie Martin, Elizabeth Wainwright, Steven Marmorstein, and Peter Jansen. 2020 · 2020
Later among the works it cites.
Did Aristotle use a laptop? A question answering benchmark with implicit reasoning strategies
Mor Geva, Daniel Khashabi, Elad Segal, Tushar Khot, D. Roth, and Jonathan Berant. 2021 · 2021
Closest in time.
General-purpose question-answering with Macaw
Oyvind Tafjord and Peter Clark. 2021 · 2021
Closest in time.
ProofWriter: Generating implications, proofs, and abductive statements over natural language
Oyvind Tafjord, B. D. Mishra, and P. Clark. 2021 · 2021
Closest in time.