Fetching the paper…
Reading the bibliography…
Language models (LMs) can perform complex reasoning either end-to-end, with hidden latent state, or compositionally, with transparent intermediate state.
Concrete Problems in AI Safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mane · 2016
Earlier work this paper cites.
Approval-directed agents
Paul Christiano · 2016
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul Christiano, Jan Leike, Tom B. Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Earlier work this paper cites.
Effect of Vitamin C Infusion on Organ Failure and Biomarkers of Inflammation and Vascular Injury in Patients With Sepsis and Severe Acute Respiratory Failure: The CITRIS-ALI Randomized Clinical Trial
Alpha A. Fowler, Jonathon D. Truwit, R. Duncan Hite, Peter E. Morris, Christine DeWilde, Anna Priday, Bernard Fisher, Leroy R. Thacker, Ramesh Natarajan, Donald F. Brophy, Robin Sculthorpe, Rahul Nanchal, Aamer Syed, Jamie Sturgill, Greg S. Martin, Jonathan Sevransky, Markos Kashiouris, Stella Hamman, Katherine F. Egan, Andrei Hastings, Wendy Spencer, Shawnda Tench, Omar Mehkri, James Bindas, Abhijit Duggal, Jeanette Graf, Stephanie Zellner, Lynda Yanny, Catherine McPolin, Tonya Hollrith, David Kramer, Charles Ojielo, Tessa Damm, Evan Cassity, Aleksandra Wieliczko, and Matthew Halquist · 2019
Earlier work this paper cites.
Fine-Tuning Language Models from Human Preferences, January 2020
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving · 2020
Earlier work this paper cites.
Learning to summarize from human feedback
Nisan Stiennon, Long Ouyang, Jeff Wu, Daniel M. Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano · 2020
Earlier work this paper cites.
NILE : Natural Language Inference with Faithful Natural Language Explanations
Sawan Kumar and Partha Talukdar · 2020
Earlier work this paper cites.
Recursively Summarizing Books with Human Feedback, September 2021
Jeff Wu, Long Ouyang, Daniel M. Ziegler, Nisan Stiennon, Ryan Lowe, Jan Leike, and Paul Christiano · 2021
Earlier work this paper cites.
WebGPT: Browser-assisted question-answering with human feedback
Reiichiro Nakano, Jacob Hilton, Suchir Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, Xu Jiang, Karl Cobbe, Tyna Eloundou, Gretchen Krueger, Kevin Button, and Matthew Knight · 2021
Earlier work this paper cites.
Decomposing Complex Questions Makes Multi-Hop QA Easier and More Interpretable, October 2021
Ruiliu Fu, Han Wang, Xuejun Zhang, Jun Zhou, and Yonghong Yan · 2021
Earlier work this paper cites.
Pretrained Transformers for Text Ranking: BERT and Beyond, August 2021
Jimmy Lin, Rodrigo Nogueira, and Andrew Yates · 2021
Earlier work this paper cites.
Decontextualization: Making Sentences Stand-Alone
Eunsol Choi, Jennimaria Palomaki, Matthew Lamm, Tom Kwiatkowski, Dipanjan Das, and Michael Collins · 2021
Earlier work this paper cites.
A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers
Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A. Smith, and Matt Gardner · 2021
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe · 2022
Earlier work this paper cites.
Self-critiquing models for assisting human evaluators
William Saunders, Catherine Yeh, Jeff Wu, Steven Bills, Long Ouyang, Jonathan Ward, and Jan Leike · 2022
Earlier work this paper cites.
Supervise Process, Not Outcomes, 2022
Andreas Stuhlmüller and Jungwon Byun · 2022
Earlier work this paper cites.
Without specific countermeasures, the easiest path to transformative AI likely leads to AI takeover
Ajeya Cotra · 2022
Cited alongside, same era.
The alignment problem from a deep learning perspective
Richard Ngo, Lawrence Chan, and Sören Mindermann · 2022
Cited alongside, same era.
Solving math word problems with process- and outcome-based feedback, November 2022
Jonathan Uesato, Nate Kushman, Ramana Kumar, Francis Song, Noah Siegel, Lisa Wang, Antonia Creswell, Geoffrey Irving, and Irina Higgins · 2022
Cited alongside, same era.
Summarization Programs: Interpretable Abstractive Summarization with Neural Modular Trees, September 2022
Swarnadeep Saha, Shiyue Zhang, Peter Hase, and Mohit Bansal · 2022
Cited alongside, same era.
ReAct: Synergizing Reasoning and Acting in Language Models, October 2022
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2022
Cited alongside, same era.
Measuring and narrowing the compositionality gap in language models
Ofir Press, Muru Zhang, Sewon Min, Ludwig Schmidt, Noah A Smith, and Mike Lewis · 2022
Later among the works it cites.
Decomposed Prompting: A Modular Approach for Solving Complex Tasks, October 2022
Tushar Khot, Harsh Trivedi, Matthew Finlayson, Yao Fu, Kyle Richardson, Peter Clark, and Ashish Sabharwal · 2022
Later among the works it cites.
ThinkSum: Probabilistic reasoning over sets using large language models, October 2022
Batu Ozturkler, Nikolay Malkin, Zhen Wang, and Nebojsa Jojic · 2022
Later among the works it cites.
Re3: Generating Longer Stories With Recursive Reprompting and Revision, October 2022
Kevin Yang, Yuandong Tian, Nanyun Peng, and Dan Klein · 2022
Later among the works it cites.
PAL: Program-aided Language Models, November 2022
Luyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon, Pengfei Liu, Yiming Yang, Jamie Callan, and Graham Neubig · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Selection-Inference: Exploiting Large Language Models for Interpretable Logical Reasoning, May 2022
Antonia Creswell, Murray Shanahan, and Irina Higgins · 2022
Cited alongside, same era.
Maieutic Prompting: Logically Consistent Reasoning with Recursive Explanations
Jaehun Jung, Lianhui Qin, Sean Welleck, Faeze Brahman, Chandra Bhagavatula, Ronan Le Bras, and Yejin Choi · 2022
Cited alongside, same era.
FaiRR: Faithful and Robust Deductive Reasoning over Natural Language, March 2022
Soumya Sanyal, Harman Singh, and Xiang Ren · 2022
Cited alongside, same era.
Faithful Reasoning Using Large Language Models, August 2022
Antonia Creswell and Murray Shanahan · 2022
Cited alongside, same era.
Natural Language Deduction through Search over Statement Compositions, October 2022
Kaj Bostrom, Zayne Sprague, Swarat Chaudhuri, and Greg Durrett · 2022
Cited alongside, same era.
Locate Then Ask: Interpretable Stepwise Reasoning for Multi-hop Question Answering, August 2022
Siyuan Wang, Zhongyu Wei, Zhihao Fan, Qi Zhang, and Xuanjing Huang · 2022
Cited alongside, same era.
Successive Prompting for Decomposing Complex Questions, December 2022
Dheeru Dua, Shivanshu Gupta, Sameer Singh, and Matt Gardner · 2022
Cited alongside, same era.
Compositional Semantic Parsing with Large Language Models, September 2022
Andrew Drozdov, Nathanael Schärli, Ekin Akyürek, Nathan Scales, Xinying Song, Xinyun Chen, Olivier Bousquet, and Denny Zhou · 2022
Later among the works it cites.
Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions, December 2022
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal · 2022
Later among the works it cites.
LAMBADA: Backward Chaining for Automated Reasoning in Natural Language, December 2022
Seyed Mehran Kazemi, Najoung Kim, Deepti Bhatia, Xin Xu, and Deepak Ramachandran · 2022
Later among the works it cites.
STaR: Bootstrapping Reasoning With Reasoning, May 2022
Eric Zelikman, Yuhuai Wu, Jesse Mu, and Noah D. Goodman · 2022
Later among the works it cites.
Calibrating Trust of Multi-Hop Question Answering Systems with Decompositional Probes, October 2022
Kaige Xie, Sarah Wiegreffe, and Mark Riedl · 2022
Later among the works it cites.
Language Model Cascades, July 2022
David Dohan, Winnie Xu, Aitor Lewkowycz, Jacob Austin, David Bieber, Raphael Gontijo Lopes, Yuhuai Wu, Henryk Michalewski, Rif A. Saurous, Jascha Sohl-dickstein, Kevin Murphy, and Charles Sutton · 2022
Later among the works it cites.
Factored Cognition Primer, 2022
Andreas Stuhlmüller, Justin Reppert, and Luke Stebbing · 2022
Later among the works it cites.
Langchain/langchain at master ⋅ \cdot hwchase17/langchain
Harrison Chase · 2022
Later among the works it cites.
Sub-Task Decomposition Enables Learning in Sequence to Sequence Tasks, October 2022
Noam Wies, Yoav Levine, and Amnon Shashua · 2022
Later among the works it cites.
Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning, August 2022
Haokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin Raffel · 2022
Later among the works it cites.
Honest Students from Untrusted Teachers: Learning an Interpretable Question-Answering Pipeline from a Pretrained Language Model, October 2022
Jacob Eisenstein, Daniel Andor, Bernd Bohnet, Michael Collins, and David Mimno · 2022
Later among the works it cites.