Fetching the paper…
Reading the bibliography…
Our goal is a modern approach to answering questions via systematic reasoning where answers are supported by human interpretable proof trees grounded in an NL corpus of authoritative facts.
Programs with common sense, 1959
John McCarthy · 1959
Earlier work this paper cites.
Baseball: An automatic question-answerer
Bert F. Green, Alice K. Wolf, Carol Chomsky, and Kenneth Laughery · 1961
Earlier work this paper cites.
Dynamic memory: A theory of reminding and learning in computers and people
Roger C Schank · 1983
Earlier work this paper cites.
Introduction to expert systems
P Jackson · 1986
Earlier work this paper cites.
Of brittleness and bottlenecks: Challenges in the creation of pattern-recognition and expert-system models
Mark A Musen and Johan Van der Lei · 1988
Earlier work this paper cites.
Planning as satisfiability
Henry A Kautz, Bart Selman, et al · 1992
Earlier work this paper cites.
Unifying sat-based and graph-based planning
Henry Kautz and Bart Selman · 1999
Earlier work this paper cites.
A probabilistic model of information retrieval: development and comparative experiments - part 1
Karen Sparck Jones, Steve Walker, and Stephen E. Robertson · 2000
Earlier work this paper cites.
Expert systems in production planning and scheduling: A state-of-the-art survey
Kostas S Metaxiotis, Dimitris Askounis, and John Psarras · 2002
Earlier work this paper cites.
The design and implementation of vampire
Alexandre Riazanov and Andrei Voronkov · 2002
Earlier work this paper cites.
Approximate reasoning by similarity-based sld resolution
Maria I Sessa · 2002
Earlier work this paper cites.
Recognising textual entailment with logical inference
Johan Bos and Katja Markert · 2005
Earlier work this paper cites.
Artificial intelligence: a modern approach
Stuart Russell and Peter Norvig · 2010
Earlier work this paper cites.
Using lookaheads with optimal best-first search
Roni Stern, Tamar Kulberis, Ariel Felner, and Robert Holte · 2010
Earlier work this paper cites.
Semantic representation
Lenhart Schubert · 2015
Earlier work this paper cites.
Neural module networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng · 2016
Earlier work this paper cites.
Answering complex questions using open information extraction
Tushar Khot, Ashish Sabharwal, and Peter Clark · 2017
Earlier work this paper cites.
End-to-end differentiable proving
Tim Rocktäschel and Sebastian Riedel · 2017
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord · 2018
Earlier work this paper cites.
Transforming question answering datasets into natural language inference datasets
Dorottya Demszky, Kelvin Guu, and Percy Liang · 2018
Earlier work this paper cites.
Scitail: A textual entailment dataset from science question answering
Tushar Khot, Ashish Sabharwal, and Peter Clark · 2018
Earlier work this paper cites.
Deepproblog: Neural probabilistic logic programming
Robin Manhaeve, Sebastijan Dumancic, Angelika Kimmig, Thomas Demeester, and Luc De Raedt · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal · 2018
Earlier work this paper cites.
Optuna: A next-generation hyperparameter optimization framework
Takuya Akiba, Shotaro Sano, Toshihiko Yanase, Takeru Ohta, and Masanori Koyama · 2019
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi · 2019
Earlier work this paper cites.
TextGraphs 2019 shared task on multi-hop inference for explanation regeneration
Peter Jansen and Dmitry Ustalov · 2019
Earlier work this paper cites.
Billion-scale similarity search with GPUs
Jeff Johnson, Matthijs Douze, and Hervé Jégou · 2019
Earlier work this paper cites.
On the possibilities and limitations of multi-hop reasoning under linguistic imperfections
Daniel Khashabi, Erfan Sadeqi Azer, Tushar Khot, Ashish Sabharwal, and Dan Roth · 2019
Cited alongside, same era.
Exploiting explicit paths for multi-hop reading comprehension
Souvik Kundu, Tushar Khot, Ashish Sabharwal, and Peter Clark · 2019
Cited alongside, same era.
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer · 2019
Cited alongside, same era.
Improving open-domain dialogue systems via multi-turn incomplete utterance restoration
Zhufeng Pan, Kun Bai, Yan Wang, Lianqiang Zhou, and Xiaojiang Liu · 2019
Cited alongside, same era.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych · 2019
Natural language deduction through search over statement compositions
Kaj Bostrom, Zayne Sprague, Swarat Chaudhuri, and Greg Durrett · 2022
Closest in time.
Faithful reasoning using large language models
Antonia Creswell and Murray Shanahan · 2022
Closest in time.
Towards teachable reasoning systems: Using a dynamic memory of user feedback for continual system improvement
Bhavana Dalvi Mishra, Oyvind Tafjord, and Peter Clark · 2022
Closest in time.
Rethinking with retrieval: Faithful large language model inference
Hangfeng He, Hongming Zhang, and Dan Roth · 2022
Closest in time.
METGEN: A module-based entailment tree generation framework for answer explanation
Ruixin Hong, Hongming Zhang, Xintong Yu, and Changshui Zhang · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
NLProlog: Reasoning with weak unification for question answering in natural language
Leon Weber, Pasquale Minervini, Jannes Münchmeyer, Ulf Leser, and Tim Rocktäschel · 2019
Cited alongside, same era.
Quick and (not so) dirty: Unsupervised selection of justification sentences for multi-hop question answering
Vikas Yadav, Steven Bethard, and Mihai Surdeanu · 2019
Cited alongside, same era.
Neural module networks for reasoning over text
Nitish Gupta, Kevin Lin, Dan Roth, Sameer Singh, and Matt Gardner · 2020
Cited alongside, same era.
Learning to explain: Datasets and models for identifying valid reasoning chains in multihop question-answering
Harsh Jhamtani and Peter Clark · 2020
Cited alongside, same era.
Qasc: A dataset for question answering via sentence composition
Tushar Khot, Peter Clark, Michal Guerquin, Peter Jansen, and Ashish Sabharwal · 2020
Cited alongside, same era.
Generative language modeling for automated theorem proving
Stanislas Polu and Ilya Sutskever · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Cited alongside, same era.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Yejin Bang, Andrea Madotto, and Pascale Fung · 2022
Closest in time.
Maieutic prompting: Logically consistent reasoning with recursive explanations
Jaehun Jung, Lianhui Qin, Sean Welleck, Faeze Brahman, Chandra Bhagavatula, Ronan Le Bras, and Yejin Choi · 2022
Closest in time.
Decomposed prompting: A modular approach for solving complex tasks
Tushar Khot, Harsh Trivedi, Matthew Finlayson, Yao Fu, Kyle Richardson, Peter Clark, and Ashish Sabharwal · 2022
Closest in time.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Closest in time.
Entailment tree explanations via iterative retrieval-generation reasoner
Danilo Neves Ribeiro, Shen Wang, Xiaofei Ma, Rui Dong, Xiaokai Wei, Henry Zhu, Xinchi Chen, Zhiheng Huang, Peng Xu, Andrew O. Arnold, and Dan Roth · 2022
Closest in time.
Natural language deduction with incomplete information
Zayne Sprague, Kaj Bostrom, Swarat Chaudhuri, and Greg Durrett · 2022
Closest in time.
Entailer: Answering questions with faithful and truthful chains of reasoning
Oyvind Tafjord, Bhavana Dalvi Mishra, and Peter Clark · 2022
Closest in time.
Diff-explainer: Differentiable convex optimization for explainable multi-hop inference
Mokanarangan Thayaparan, Marco Valentino, Deborah Ferreira, Julia Rozanova, and André Freitas · 2022
Closest in time.
Case-based abductive natural language inference
Marco Valentino, Mokanarangan Thayaparan, and André Freitas · 2022
Closest in time.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou · 2022
Closest in time.
Naturalprover: Grounded mathematical proof generation with language models
Sean Welleck, Jiacheng Liu, Ximing Lu, Hannaneh Hajishirzi, and Yejin Choi · 2022
Closest in time.
Generating natural language proofs with verifier-guided search
Kaiyu Yang, Jia Deng, and Danqi Chen · 2022
Closest in time.
LAMBADA: Backward chaining for automated reasoning in natural language
Mehran Kazemi, Najoung Kim, Deepti Bhatia, Xin Xu, and Deepak Ramachandran · 2023
Closest in time.
Faithful chain-of-thought reasoning
Qing Lyu, Shreya Havaldar, Adam Stein, Li Zhang, Delip Rao, Eric Wong, Marianna Apidianaki, and Chris Callison-Burch · 2023
Closest in time.
LINC: A neurosymbolic approach for logical reasoning by combining language models with first-order logic provers
Theo Olausson, Alex Gu, Ben Lipkin, Cedegao Zhang, Armando Solar-Lezama, Joshua Tenenbaum, and Roger Levy · 2023
Closest in time.
Logic-LM: Empowering large language models with symbolic solvers for faithful logical reasoning
Liangming Pan, Alon Albalak, Xinyi Wang, and William Wang · 2023
Closest in time.
Branch-solve-merge improves large language model evaluation and generation, 2023
Swarnadeep Saha, Omer Levy, Asli Celikyilmaz, Mohit Bansal, Jason Weston, and Xian Li · 2023
Closest in time.
MURMUR: Modular multi-step reasoning for semi-structured data-to-text generation
Swarnadeep Saha, Xinyan Yu, Mohit Bansal, Ramakanth Pasunuru, and Asli Celikyilmaz · 2023
Closest in time.
From word models to world models: Translating from natural language to the probabilistic language of thought, 2023
Lionel Wong, Gabriel Grand, Alexander K. Lew, Noah D. Goodman, Vikash K. Mansinghka, Jacob Andreas, and Joshua B. Tenenbaum · 2023
Closest in time.
SatLM: Satisfiability-aided language models using declarative prompting
Xi Ye, Qiaochu Chen, Isil Dillig, and Greg Durrett · 2023
Closest in time.
Towards Faithful Model Explanation in NLP: A Survey
Qing Lyu, Marianna Apidianaki, and Chris Callison-Burch · 2024
Closest in time.
Beyond chain-of-thought: A survey of chain-of-x paradigms for llms, 2024
Yu Xia, Rui Wang, Xu Liu, Mingyan Li, Tong Yu, Xiang Chen, Julian McAuley, and Shuai Li · 2024
Closest in time.