Fetching the paper…
Reading the bibliography…
Our goal is a question-answering (QA) system that can show how its answers are implied by its own internal beliefs via a systematic chain of reasoning.
QuaRTz: An open-domain dataset of qualitative relationship questions
Oyvind Tafjord, Matt Gardner, Kevin Lin, and Peter Clark. 2019 · 1909
Earlier work this paper cites.
Natural language inference
Christopher D. Manning and Bill MacCartney. 2009 · 2009
Earlier work this paper cites.
ProofWriter: Generating implications, proofs, and abductive statements over natural language
Oyvind Tafjord, B. D. Mishra, and P. Clark. 2020 · 2012
Earlier work this paper cites.
Recognizing Textual Entailment: Models and Applications
Ido Dagan, Dan Roth, Mark Sammons, and Fabio Massimo Zanzotto. 2013 · 2013
Earlier work this paper cites.
Transforming question answering datasets into natural language inference datasets
Dorottya Demszky, Kelvin Guu, and Percy Liang. 2018 · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? A new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal. 2018 · 2018
Earlier work this paper cites.
COMET: Commonsense transformers for automatic knowledge graph construction
Antoine Bosselut, Hannah Rashkin, Maarten Sap, Chaitanya Malaviya, Asli Celikyilmaz, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
A logic-driven framework for consistency of neural models
Tao Li, Vivek Gupta, Maitrey Mehta, and Vivek Srikumar. 2019 · 2019
Earlier work this paper cites.
Language models as knowledge bases?
F. Petroni, Tim Rocktäschel, Patrick Lewis, A. Bakhtin, Y. Wu, Alexander H. Miller, and S. Riedel. 2019 · 2019
Earlier work this paper cites.
Are red roses red? Evaluating consistency of question-answering models
Marco Tulio Ribeiro, Carlos Guestrin, and Sameer Singh. 2019 · 2019
Earlier work this paper cites.
Explanatory interactive machine learning
Stefano Teso and Kristian Kersting. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown et al. 2020 · 2020
Earlier work this paper cites.
Transformers as soft reasoners over language
Peter Clark, Oyvind Tafjord, and Kyle Richardson. 2020 · 2020
Cited alongside, same era.
Braid: Weaving symbolic and neural knowledge into coherent logical explanations
Aditya Kalyanpur, Tom Breloff, and David A. Ferrucci. 2020 · 2020
Cited alongside, same era.
Negated and misprimed probes for pretrained language models: Birds can talk, but cannot fly
Nora Kassner and H. Schütze. 2020 · 2020
Cited alongside, same era.
Unsupervised commonsense question answering with self-talk
Vered Shwartz, Peter West, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
LeapOfThought: Teaching pre-trained models to systematically reason over implicit knowledge
Alon Talmor, Oyvind Tafjord, P. Clark, Y. Goldberg, and Jonathan Berant. 2020 · 2020
Cited alongside, same era.
WorldTree V2: A corpus of science-domain structured explanations and inference patterns supporting multi-hop inference
Show your work: Scratchpads for intermediate computation with language models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, Charles Sutton, and Augustus Odena. 2021 · 2021
Later among the works it cites.
General-purpose question-answering with Macaw
Oyvind Tafjord and Peter Clark. 2021 · 2021
Later among the works it cites.
Teach me to explain: A review of datasets for explainable NLP
Sarah Wiegreffe and Ana Marasović. 2021 · 2021
Later among the works it cites.
Natural language deduction through search over statement compositions
Kaj Bostrom, Zayne Sprague, Swarat Chaudhuri, and Greg Durrett. 2022 · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhengnan Xie, Sebastian Thiem, Jaycie Martin, Elizabeth Wainwright, Steven Marmorstein, and Peter Jansen. 2020 · 2020
Cited alongside, same era.
Flexible generation of natural language deductions
Kaj Bostrom, Xinyu Zhao, Swarat Chaudhuri, and Greg Durrett. 2021 · 2021
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Cited alongside, same era.
Explaining answers with entailment trees
Bhavana Dalvi, Peter A. Jansen, Oyvind Tafjord, Zhengnan Xie, Hannah Smith, Leighanna Pipatanangkura, and Peter Clark. 2021 · 2021
Cited alongside, same era.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, E. Hovy, H. Schutze, and Yoav Goldberg. 2021 · 2021
Cited alongside, same era.
Paragraph-level commonsense transformers with recurrent memory
Saadia Gabriel, Chandra Bhagavatula, Vered Shwartz, Ronan Le Bras, Maxwell Forbes, and Yejin Choi. 2021 · 2021
Cited alongside, same era.
BeliefBank: Adding memory to a pre-trained language model for a systematic notion of belief
Nora Kassner, Oyvind Tafjord, Hinrich Schutze, and Peter Clark. 2021 · 2021
Cited alongside, same era.
Antonia Creswell, Murray Shanahan, and Irina Higgins. 2022 · 2022
Closest in time.
Metgen: A module-based entailment tree generation framework for answer explanation
Ruixin Hong, Hongming Zhang, Xintong Yu, and Changshui Zhang. 2022 · 2022
Closest in time.
Competition-level code generation with alphacode
Yujia Li, David H. Choi, et al. 2022 · 2022
Closest in time.
Towards teachable reasoning systems: Using a dynamic memory of user feedback for continual system improvement
Bhavana Dalvi Mishra, Oyvind Tafjord, and Peter Clark. 2022 · 2022
Closest in time.
Entailment tree explanations via iterative retrieval-generation reasoner
Danilo Neves Ribeiro, Shen Wang, Xiaofei Ma, Rui Dong, Xiaokai Wei, Henry Zhu, Xinchi Chen, Zhiheng Huang, Peng Xu, Andrew O. Arnold, and Dan Roth. 2022 · 2022
Closest in time.
Memory-assisted prompt editing to improve GPT-3 after deployment
Niket Tandon, Aman Madaan, Peter Clark, and Yiming Yang. 2022 · 2022
Closest in time.
Rationale-augmented ensembles in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, and Denny Zhou. 2022 · 2022
Closest in time.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Closest in time.
Generating natural language proofs with verifier-guided search
Kaiyu Yang, Jia Deng, and Danqi Chen. 2022 · 2022
Closest in time.