Fetching the paper…
Reading the bibliography…
Chain-of-thought responses from language models improve performance across most benchmarks.
Designing and Interpreting Probes with Control Tasks
John Hewitt and Percy Liang · 2019
Earlier work this paper cites.
Challenging big-bench tasks and whether chain-of-thought can solve them
Mirac Suzgun, Nathan Scales, Nathanael Schärli, Sebastian Gehrmann, Yi Tay, Hyung Won Chung, Aakanksha Chowdhery, Quoc V. Le, Ed H. Chi, Denny Zhou, and Jason Wei · 2022
Earlier work this paper cites.
Towards revealing the mystery behind chain of thought: A theoretical perspective
Guhao Feng, Bohang Zhang, Yuntian Gu, Haotian Ye, Di He, and Liwei Wang · 2023
Earlier work this paper cites.
How LLMs are and are not myopic
Janus · 2023
Earlier work this paper cites.
Measuring faithfulness in chain-of-thought reasoning
Tamera Lanham, Anna Chen, Ansh Radhakrishnan, Benoit Steiner, Carson Denison, Danny Hernandez, Dustin Li, Esin Durmus, Evan Hubinger, Jackson Kernion, et al · 2023
Earlier work this paper cites.
The parallelism tradeoff: Limitations of log-precision transformers
William Merrill and Ashish Sabharwal · 2023
Earlier work this paper cites.
Future Lens: Anticipating Subsequent Tokens from a Single Hidden State
Koyena Pal, Jiuding Sun, Andrew Yuan, Byron C. Wallace, and David Bau · 2023
Cited alongside, same era.
Llms are (mostly) not helped by filler tokens, 2023
Kshitij Sachan · 2023
Cited alongside, same era.
Transformers as recognizers of formal languages: A survey on expressivity, 2023
Lena Strobl, William Merrill, Gail Weiss, David Chiang, and Dana Angluin · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models, 2023
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample · 2023
Cited alongside, same era.
Language models don't always say what they think: Unfaithful explanations in chain-of-thought prompting
Think before you speak: Training language models with pause tokens
Sachin Goyal, Ziwei Ji, Ankit Singh Rawat, Aditya Krishna Menon, Sanjiv Kumar, and Vaishnavh Nagarajan · 2024
Closest in time.
Why are sensitive functions hard for transformers?
Michael Hahn and Mark Rofin · 2024
Closest in time.
Representational strengths and limitations of transformers
Clayton Sanford, Daniel J Hsu, and Matus Telgarsky · 2024
Closest in time.
Do language models plan ahead for future tokens?, March 2024
Wilson Wu, John X. Morris, and Lionel Levine · 2024
Closest in time.
Quiet-star: Language models can teach themselves to think before speaking
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Miles Turpin, Julian Michael, Ethan Perez, and Samuel Bowman · 2023
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou · 2023
Cited alongside, same era.
A logic for expressing log-precision transformers
William Merrill and Ashish Sabharwal
Cited in the paper.
The expressive power of transformers with chain of thought, 2023c
William Merrill and Ashish Sabharwal
Cited in the paper.
Eric Zelikman, Georges Harik, Yijia Shao, Varuna Jayasiri, Nick Haber, and Noah D. Goodman · 2024
Closest in time.