Fetching the paper…
Reading the bibliography…
Pretrained language models have shown superior performance on many natural language processing tasks, yet they still struggle at multi-step formal reasoning tasks like grade school math problems.
From language to programs: Bridging reinforcement learning and maximum marginal likelihood
Kelvin Guu, Panupong Pasupat, Evan Zheran Liu, and Percy Liang · 2017
Earlier work this paper cites.
Neural semantic parsing with type constraints for semi-structured tables
Jayant Krishnamurthy, Pradeep Dasigi, and Matt Gardner · 2017
Earlier work this paper cites.
Neural symbolic machines: Learning semantic parsers on freebase with weak supervision
Chen Liang, Jonathan Berant, Quoc Le, Kenneth Forbus, and Ni Lao · 2017
Earlier work this paper cites.
Leveraging grammar and reinforcement learning for neural program synthesis
Rudy Bunel, Matthew Hausknecht, Jacob Devlin, Rishabh Singh, and Pushmeet Kohli · 2018
Earlier work this paper cites.
Execution-guided neural program synthesis
Xinyun Chen, Chang Liu, and Dawn Song · 2018
Earlier work this paper cites.
Improving text-to-sql evaluation methodology
Catherine Finegan-Dollak, Jonathan K Kummerfeld, Li Zhang, Karthik Ramanathan, Sesh Sadasivam, Rui Zhang, and Dragomir R Radev · 2018
Earlier work this paper cites.
Memory augmented policy optimization for program synthesis and semantic parsing
Chen Liang, Mohammad Norouzi, Jonathan Berant, Quoc V Le, and Ni Lao · 2018
Earlier work this paper cites.
Program synthesis through reinforcement learning guided tree search
Riley Simmons-Edler, Anders Miltner, and Sebastian Seung · 2018
Earlier work this paper cites.
Robust text-to-sql generation with execution-guided decoding
Chenglong Wang, Kedar Tatwawadi, Marc Brockschmidt, Po-Sen Huang, Yi Mao, Oleksandr Polozov, and Rishabh Singh · 2018
Earlier work this paper cites.
Learning to generalize from sparse and underspecified rewards
Rishabh Agarwal, Chen Liang, Dale Schuurmans, and Mohammad Norouzi · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Write, execute, assess: Program synthesis with a repl
Kevin Ellis, Maxwell Nye, Yewen Pu, Felix Sosa, Josh Tenenbaum, and Armando Solar-Lezama · 2019
Cited alongside, same era.
A discrete hard em approach for weakly supervised question answering
Sewon Min, Danqi Chen, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2019
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Merging weak and active supervision for semantic parsing
Program synthesis with large language models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, et al · 2021
Later among the works it cites.
GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow, March 2021
Sid Black, Leo Gao, Phil Wang, Connor Leahy, and Stella Biderman · 2021
Later among the works it cites.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman · 2021
Later among the works it cites.
Show your work: Scratchpads for intermediate computation with language models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, et al · 2021
Later among the works it cites.
Programming puzzles
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ansong Ni, Pengcheng Yin, and Graham Neubig · 2020
Cited alongside, same era.
Representing partial programs with blended abstract semantics
Maxwell Nye, Yewen Pu, Matthew Bowers, Jacob Andreas, Joshua B Tenenbaum, and Armando Solar-Lezama · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, Peter J Liu, et al · 2020
Cited alongside, same era.
Semantic evaluation for text-to-sql with distilled test suites
Ruiqi Zhong, Tao Yu, and Dan Klein · 2020
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al
Cited in the paper.
Latent execution for neural program synthesis beyond domain-specific languages
Xinyun Chen, Dawn Song, and Yuandong Tian
Cited in the paper.
Tal Schuster, Ashwin Kalyan, Alex Polozov, and Adam Tauman Kalai · 2021
Later among the works it cites.
Codet: Code generation with generated tests
Bei Chen, Fengji Zhang, Anh Nguyen, Daoguang Zan, Zeqi Lin, Jian-Guang Lou, and Weizhu Chen · 2022
Closest in time.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Closest in time.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Closest in time.