Fetching the paper…
Reading the bibliography…
Accurate mathematical reasoning with Large Language Models (LLMs) is crucial in revolutionizing domains that heavily rely on such reasoning.
Crafting papers on machine learning
Langley, P · 2000
Earlier work this paper cites.
The curious case of neural text degeneration
Holtzman, Ari, Buys, Jan, Du, Li, Forbes, Maxwell, and Choi, Yejin · 2019
Earlier work this paper cites.
Training verifiers to solve math word problems
Cobbe, Karl, Kosaraju, Vineet, Bavarian, Mohammad, Chen, Mark, Jun, Heewoo, Kaiser, Lukasz, Plappert, Matthias, Tworek, Jerry, Hilton, Jacob, Nakano, Reiichiro, et al · 2021
Earlier work this paper cites.
Measuring mathematical problem solving with the math dataset
Hendrycks, Dan, Burns, Collin, Kadavath, Saurav, Arora, Akul, Basart, Steven, Tang, Eric, Song, Dawn, and Steinhardt, Jacob · 2021
Earlier work this paper cites.
Zero-infinity: Breaking the gpu memory wall for extreme scale deep learning
Rajbhandari, Samyam, Ruwase, Olatunji, Rasley, Jeff, Smith, Shaden, and He, Yuxiong · 2021
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, Jason, Wang, Xuezhi, Schuurmans, Dale, Bosma, Maarten, Xia, Fei, Chi, Ed, Le, Quoc V, Zhou, Denny, et al · 2022
Earlier work this paper cites.
Learning from mistakes makes llm better reasoner
An, Shengnan, Ma, Zexiong, Lin, Zeqi, Zheng, Nanning, Lou, Jian-Guang, and Chen, Weizhu · 2023
Earlier work this paper cites.
Anil, Rohan, Dai, Andrew M, Firat, Orhan, Johnson, Melvin, Lepikhin, Dmitry, Passos, Alexandre, Shakeri, Siamak, Taropa, Emanuel, Bailey, Paige, Chen, Zhifeng, et al · 2023
Earlier work this paper cites.
https://www.anthropic.com/news/claude-3-family/
Anthropic · 2023
Earlier work this paper cites.
Llemma: An open language model for mathematics
Azerbayev, Zhangir, Schoelkopf, Hailey, Paster, Keiran, Santos, Marco Dos, McAleer, Stephen, Jiang, Albert Q, Deng, Jia, Biderman, Stella, and Welleck, Sean · 2023
Earlier work this paper cites.
Palm: Scaling language modeling with pathways
Chowdhery, Aakanksha, Narang, Sharan, Devlin, Jacob, Bosma, Maarten, Mishra, Gaurav, Roberts, Adam, Barham, Paul, Chung, Hyung Won, Sutton, Charles, Gehrmann, Sebastian, et al · 2023
Earlier work this paper cites.
Flashattention-2: Faster attention with better parallelism and work partitioning
Dao, Tri · 2023
Cited alongside, same era.
Fu, Jiayi, Lin, Lei, Gao, Xiaoyang, Liu, Pengli, Chen, Zhengzong, Yang, Zhirui, Zhang, Shengnan, Zheng, Xue, Li, Yan, Liu, Yuliang, et al · 2023
Cited alongside, same era.
Pal: Program-aided language models
Gao, Luyu, Madaan, Aman, Zhou, Shuyan, Alon, Uri, Liu, Pengfei, Yang, Yiming, Callan, Jamie, and Neubig, Graham · 2023
Cited alongside, same era.
Platypus: Quick, cheap, and powerful refinement of llms
Lee, Ariel N, Hunter, Cole J, and Ruiz, Nataniel · 2023
Cited alongside, same era.
Lightman, Hunter, Kosaraju, Vineet, Burda, Yura, Edwards, Harri, Baker, Bowen, Lee, Teddy, Leike, Jan, Schulman, John, Sutskever, Ilya, and Cobbe, Karl · 2023
Scaling relationship on learning mathematical reasoning with large language models
Yuan, Zheng, Yuan, Hongyi, Li, Chengpeng, Dong, Guanting, Tan, Chuanqi, and Zhou, Chang · 2023
Later among the works it cites.
Mammoth: Building math generalist models through hybrid instruction tuning
Yue, Xiang, Qu, Xingwei, Zhang, Ge, Fu, Yao, Huang, Wenhao, Sun, Huan, Su, Yu, and Chen, Wenhu · 2023
Later among the works it cites.
Orca-math: Unlocking the potential of slms in grade school math
Mitra, Arindam, Khanpour, Hamed, Rosset, Corby, and Awadallah, Ahmed · 2024
Closest in time.
https://openai.com/index/hello-gpt-4o/
OpenAI · 2024
Closest in time.
Deepseekmath: Pushing the limits of mathematical reasoning in open language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Wizardmath: Empowering mathematical reasoning for large language models via reinforced evol-instruct
Luo, Haipeng, Sun, Qingfeng, Xu, Can, Zhao, Pu, Lou, Jianguang, Tao, Chongyang, Geng, Xiubo, Lin, Qingwei, Chen, Shifeng, and Zhang, Dongmei · 2023
Cited alongside, same era.
Penedo, Guilherme, Malartic, Quentin, Hesslow, Daniel, Cojocaru, Ruxandra, Cappelli, Alessandro, Alobeidli, Hamza, Pannier, Baptiste, Almazrouei, Ebtesam, and Launay, Julien · 2023
Cited alongside, same era.
Code llama: Open foundation models for code
Roziere, Baptiste, Gehring, Jonas, Gloeckle, Fabian, Sootla, Sten, Gat, Itai, Tan, Xiaoqing Ellen, Adi, Yossi, Liu, Jingyu, Remez, Tal, Rapin, Jérémy, et al · 2023
Cited alongside, same era.
Gemini: a family of highly capable multimodal models
Team, Gemini, Anil, Rohan, Borgeaud, Sebastian, Wu, Yonghui, Alayrac, Jean-Baptiste, Yu, Jiahui, Soricut, Radu, Schalkwyk, Johan, Dai, Andrew M, Hauth, Anja, et al · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Touvron, Hugo, Lavril, Thibaut, Izacard, Gautier, Martinet, Xavier, Lachaux, Marie-Anne, Lacroix, Timothée, Rozière, Baptiste, Goyal, Naman, Hambro, Eric, Azhar, Faisal, et al · 2023
Cited alongside, same era.
Metamath: Bootstrap your own mathematical questions for large language models
Yu, Longhui, Jiang, Weisen, Shi, Han, Yu, Jincheng, Liu, Zhengying, Zhang, Yu, Kwok, James T, Li, Zhenguo, Weller, Adrian, and Liu, Weiyang · 2023
Cited alongside, same era.
Chen, Changyu, Wang, Xiting, Lin, Ting-En, Lv, Ang, Wu, Yuchuan, Gao, Xin, Wen, Ji-Rong, Yan, Rui, and Li, Yongbin
Cited in the paper.
Shao, Zhihong, Wang, Peiyi, Zhu, Qihao, Xu, Runxin, Song, Junxiao, Zhang, Mingchuan, Li, YK, Wu, Y, and Guo, Daya · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Team, Gemini, Georgiev, Petko, Lei, Ving Ian, Burnell, Ryan, Bai, Libin, Gulati, Anmol, Tanzer, Garrett, Vincent, Damien, Pan, Zhufeng, Wang, Shibo, et al · 2024
Closest in time.
Openmathinstruct-1: A 1.8 million math instruction tuning dataset
Toshniwal, Shubham, Moshkov, Ivan, Narenthiran, Sean, Gitman, Daria, Jia, Fei, and Gitman, Igor · 2024
Closest in time.
Yang, An, Yang, Baosong, Hui, Binyuan, Zheng, Bo, Yu, Bowen, Zhou, Chang, Li, Chengpeng, Li, Chengyuan, Liu, Dayiheng, Huang, Fei, et al · 2024
Closest in time.
Teaching language models to self-improve through interactive demonstrations
Yu, Xiao, Peng, Baolin, Galley, Michel, Gao, Jianfeng, and Yu, Zhou · 2024
Closest in time.
Mario eval: Evaluate your math llm with your math llm–a mathematical dataset evaluation toolkit
Zhang, Boning, Li, Chengxi, and Fan, Kai · 2024
Closest in time.