Fetching the paper…
Reading the bibliography…
In this paper, we explore how to leverage large language models (LLMs) to solve mathematical problems efficiently and accurately.
Measuring Mathematical Problem Solving With the MATH Dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt · 2021
Earlier work this paper cites.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc V Le, and Denny Zhou · 2022
Earlier work this paper cites.
The Impact of Large Language Models on Scientific Discovery: a Preliminary Study using GPT-4
Microsoft Research AI4Science and Microsoft Azure Quantum · 2023
Earlier work this paper cites.
Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks
Wenhu Chen, Xueguang Ma, Xinyi Wang, and William W. Cohen · 2023
Earlier work this paper cites.
ChatGPT’s Math Accuracy Decline and What That Has to Do With Content Creation
Jenn Greenleaf · 2023
Earlier work this paper cites.
Solving Math Word Problems by Combining Language Models With Symbolic Solvers
Joy He-Yueya, Gabriel Poesia, Rose E. Wang, and Noah D. Goodman · 2023
Earlier work this paper cites.
Give me a hint: Can LLMs take a hint to solve math problems?
Vansh Agrawal, Pratham Singla, Amitoj Singh Miglani, Shivank Garg, and Ayush Mangal · 2024
Earlier work this paper cites.
Large language models for mathematical reasoning: Progresses and challenges
Janice Ahn, Rishu Verma, Renze Lou, Di Liu, Rui Zhang, and Wenpeng Yin · 2024
Cited alongside, same era.
Cutting Through the Noise: Boosting LLM Performance on Math Word Problems
Ujjwala Anantheswaran, Himanshu Gupta, Kevin Scaria, Shreyas Verma, Chitta Baral, and Swaroop Mishra · 2024
Cited alongside, same era.
Saullm-7b: A pioneering large language model for law, 2024
Pierre Colombo, Telmo Pessoa Pires, Malik Boudiaf, Dominic Culver, Rui Melo, Caio Corro, Andre F. T. Martins, Fabrizio Esposito, Vera Lúcia Raposo, Sofia Morgado, and Michael Desa · 2024
Cited alongside, same era.
A survey on in-context learning, 2024
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Jingyuan Ma, Rui Li, Heming Xia, Jingjing Xu, Zhiyong Wu, Tianyu Liu, Baobao Chang, Xu Sun, Lei Li, and Zhifang Sui · 2024
Cited alongside, same era.
The Impact of Large Language Models in Academia: from Writing to Speaking
Mingmeng Geng, Caixi Chen, Yanru Wu, Dongping Chen, Yao Wan, and Pan Zhou · 2024
A survey of large language models for financial applications: Progress, prospects and challenges, 2024
Yuqi Nie, Yaxuan Kong, Xiaowen Dong, John M. Mulvey, H. Vincent Poor, Qingsong Wen, and Stefan Zohren · 2024
Closest in time.
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, Haowei Zhang, Mingchuan Zhang, Y. K. Li, Y. Wu, and Daya Guo · 2024
Closest in time.
What Makes Math Word Problems Challenging for LLMs?
Aditya Srivatsa and Ekaterina Kochmar · 2024
Closest in time.
AI achieves silver-medal standard solving international mathematical olympiad problems
Google DeepMind AlphaProof/AlphaGeometry teams · 2024
Closest in time.
Solving olympiad geometry without human demonstrations
Trieu Trinh, Yuhuai Tony Wu, Quoc Le, He He, and Thang Luong · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Kuei-Chun Kao, Ruochen Wang, and Cho-Jui Hsieh · 2024
Cited alongside, same era.
CHAMP: A competition-level dataset for fine-grained analyses of llms’ mathematical reasoning capabilities
Yujun Mao, Yoon Kim, and Yilun Zhou · 2024
Cited alongside, same era.
MathChat: Converse to Tackle Challenging Math Problems with LLM Agents
Yiran Wu, Feiran Jia, Shaokun Zhang, Hangyu Li, Erkang Zhu, Yue Wang, Yin Tat Lee, Richard Peng, Qingyun Wu, and Chi Wang · 2024
Closest in time.
Can LLMs Solve longer Math Word Problems Better?
Xin Xu, Tong Xiao, Zitong Chao, Zhenya Huang, Can Yang, and Yang Wang · 2024
Closest in time.