Fetching the paper…
Reading the bibliography…
Building compositional explanations requires models to combine two or more facts that, together, describe why the answer to a question is correct.
UNIFIEDQA: Crossing format boundaries with a single QA system
Daniel Khashabi, Sewon Min, Tushar Khot, Ashish Sabharwal, Oyvind Tafjord, Peter Clark, and Hannaneh Hajishirzi. 2020 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
The measurement of observer agreement for categorical data
J Richard Landis and Gary G Koch. 1977 · 1977
Earlier work this paper cites.
Transformers as soft reasoners over language
Peter Clark, Oyvind Tafjord, and Kyle Richardson. 2020 · 2002
Earlier work this paper cites.
Automatic evaluation of summaries using n-gram co-occurrence statistics
Chin-Yew Lin and Eduard Hovy. 2003 · 2003
Earlier work this paper cites.
Learning-to-rank with bert in tf-ranking
Shuguang Han, Xuanhui Wang, Mike Bendersky, and Marc Najork. 2020 · 2004
Earlier work this paper cites.
A survey on explainability in machine reading comprehension
Mokanarangan Thayaparan, Marco Valentino, and André Freitas. 2020 · 2010
Earlier work this paper cites.
Question answering via integer programming over semi-structured knowledge
Daniel Khashabi, Tushar Khot, Ashish Sabharwal, Peter Clark, Oren Etzioni, and Dan Roth. 2016 · 2016
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. 2016 · 2016
Earlier work this paper cites.
Race: Large-scale reading comprehension dataset from examinations
Guokun Lai, Qizhe Xie, Hanxiao Liu, Yiming Yang, and Eduard Hovy. 2017 · 2017
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord. 2018 · 2018
Earlier work this paper cites.
WorldTree: A corpus of explanation graphs for elementary science questions supporting multi-hop inference
Peter Jansen, Elizabeth Wainwright, Steven Marmorstein, and Clayton Morrison. 2018 · 2018
Earlier work this paper cites.
Looking beyond the surface: A challenge set for reading comprehension over multiple sentences
Daniel Khashabi, Snigdha Chaturvedi, Michael Roth, Shyam Upadhyay, and Dan Roth. 2018 · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal. 2018 · 2018
Cited alongside, same era.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Cited alongside, same era.
Chains-of-reasoning at textgraphs 2019 shared task: Reasoning over chains of facts for explainable multi-hop inference
Rajarshi Das, Ameya Godbole, Manzil Zaheer, Shehzaad Dhuliawala, and Andrew McCallum. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
TextGraphs 2019 shared task on multi-hop inference for explanation regeneration
Peter Jansen and Dmitry Ustalov. 2019 · 2019
Cited alongside, same era.
Qasc: A dataset for question answering via sentence composition
Tushar Khot, Peter Clark, Michal Guerquin, Peter Jansen, and Ashish Sabharwal. 2020 · 2020
Later among the works it cites.
Explanatory completeness
Joanna Korman and Sangeet Khemlani. 2020 · 2020
Later among the works it cites.
Pgl at textgraphs 2020 shared task: Explanation regeneration using language and graph learning methods
Weibin Li, Yuxiang Lu, Zhengjie Huang, Weiyue Su, Jiaxiang Liu, Shikun Feng, and Yu Sun. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
Zero: Memory optimizations toward training trillion parameter models
Samyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, and Yuxiong He. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
KagNet: Knowledge-aware graph networks for commonsense reasoning
Bill Yuchen Lin, Xinyue Chen, Jamin Chen, and Xiang Ren. 2019 · 2019
Cited alongside, same era.
NLProlog: Reasoning with weak unification for question answering in natural language
Leon Weber, Pasquale Minervini, Jannes Münchmeyer, Ulf Leser, and Tim Rocktäschel. 2019 · 2019
Cited alongside, same era.
Autoregressive reasoning over chains of facts with transformers
Ruben Cartuyvels, Graham Spinks, and Marie-Francine Moens. 2020 · 2020
Cited alongside, same era.
R4C: A benchmark for evaluating RC systems to get the right answer for the right reason
Naoya Inoue, Pontus Stenetorp, and Kentaro Inui. 2020 · 2020
Cited alongside, same era.
CoSaTa: A constraint satisfaction solver and interpreted language for semi-structured tables of sentences
Peter Jansen. 2020 · 2020
Cited alongside, same era.
TextGraphs 2020 shared task on multi-hop inference for explanation regeneration
Peter Jansen and Dmitry Ustalov. 2020 · 2020
Cited alongside, same era.
Learning to explain: Datasets and models for identifying valid reasoning chains in multihop question-answering
Harsh Jhamtani and Peter Clark. 2020 · 2020
Cited alongside, same era.
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Later among the works it cites.
WorldTree v2: A corpus of science-domain structured explanations and inference patterns supporting multi-hop inference
Zhengnan Xie, Sebastian Thiem, Jaycie Martin, Elizabeth Wainwright, Steven Marmorstein, and Peter Jansen. 2020 · 2020
Later among the works it cites.
Experts, errors, and context: A large-scale study of human evaluation for machine translation
Markus Freitag, George Foster, David Grangier, Viresh Ratnakar, Qijun Tan, and Wolfgang Macherey. 2021 · 2021
Closest in time.
Genie: A leaderboard for human-in-the-loop evaluation of text generation
Daniel Khashabi, Gabriel Stanovsky, Jonathan Bragg, Nicholas Lourie, Jungo Kasai, Yejin Choi, Noah A Smith, and Daniel S Weld. 2021 · 2021
Closest in time.
Deepblueai at textgraphs 2021 shared task: Treating multi-hop inference explanation regeneration as a ranking problem
Chunguang Pan, Bingyan Song, and Zhipeng Luo. 2021 · 2021
Closest in time.
TextGraphs 2021 shared task on multi-hop inference for explanation regeneration
Mokanarangan Thayaparan, Marco Valentino, Peter Jansen, and Dmitry Ustalov. 2021 · 2021
Closest in time.
Teach me to explain: A review of datasets for explainable nlp
Sarah Wiegreffe and Ana Marasović. 2021 · 2021
Closest in time.