Fetching the paper…
Reading the bibliography…
Evaluating statement autoformalization, translating natural language mathematics into formal languages like Lean 4, remains a significant challenge, with few metrics, datasets, and standards to robustly measure progress.
Exploration of Neural Machine Translation in Autoformalization of Mathematics in Mizar
Qingxiang Wang, Chad Brown, Cezary Kaliszyk, and Josef Urban. 2020 · 1912
Earlier work this paper cites.
Isabelle/Hol a Proof Assistant for Higher-Order Logic
Tobias Nipkow, Lawrence C. Paulson, and Markus Wenzel. 2002 · 2002
Earlier work this paper cites.
Interactive theorem proving and program development. Coq’Art: The Calculus of inductive constructions
Pierre Castéran and Yves Bertot. 2004 · 2010
Earlier work this paper cites.
plastex/plastex
plasTeX Development Team. 2024 · 2014
Earlier work this paper cites.
The lean theorem prover (system description)
Leonardo de Moura, Soonho Kong, Jeremy Avigad, Floris van Doorn, and Jakob von Raumer. 2015 · 2015
Earlier work this paper cites.
Texygen: A benchmarking platform for text generation models
Yaoming Zhu, Sidi Lu, Lei Zheng, Jiaxian Guo, Weinan Zhang, Jun Wang, and Yong Yu. 2018 · 2018
Earlier work this paper cites.
The lean mathematical library
The mathlib Community. 2020 · 2020
Earlier work this paper cites.
A Promising Path Towards Autoformalization and General Artificial Intelligence
Christian Szegedy. 2020 · 2020
Earlier work this paper cites.
Measuring Mathematical Problem Solving With the MATH Dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
The lean 4 theorem prover and programming language
Leonardo de Moura and Sebastian Ullrich. 2021 · 2021
Earlier work this paper cites.
Towards a Mathematics Formalisation Assistant using Large Language Models
Ayush Agrawal, Siddhartha Gadgil, Navin Goyal, Ashvni Narayanan, and Anand Tadipatri. 2022 · 2022
Earlier work this paper cites.
Solving Quantitative Reasoning Problems with Language Models
Aitor Lewkowycz, Anders Johan Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay Venkatesh Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, Yuhuai Wu, Behnam Neyshabur, Guy Gur-Ari, and Vedant Misra. 2022 · 2022
Cited alongside, same era.
Mathematical Components
Assia Mahboubi and Enrico Tassi. 2022 · 2022
Cited alongside, same era.
leanprover-community/con-nf
Sky Wilshaw. 2025 · 2022
Cited alongside, same era.
Autoformalization with Large Language Models
Yuhuai Wu, Albert Q. Jiang, Wenda Li, Markus N. Rabe, Charles Staats, Mateja Jamnik, and Christian Szegedy. 2022 · 2022
Cited alongside, same era.
MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics
Kunhao Zheng, Jesse Michael Han, and Stanislas Polu. 2022 · 2022
Cited alongside, same era.
LeanDojo: Theorem Proving with Retrieval-Augmented Language Models
Kaiyu Yang, Aidan M. Swope, Alex Gu, Rahul Chalamala, Peiyang Song, Shixing Yu, Saad Godil, Ryan Prenger, and Anima Anandkumar. 2023 · 2023
Later among the works it cites.
Automated reasoning for mathematics
Jeremy Avigad. 2024 · 2024
Closest in time.
Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Alex Vaughan, Amy Yang, Angela Fan, Anirudh Goyal, Anthony Hartshorn, Aobo Yang, Archi Mitra, Archie Sravankumar, Artem Korenev, Arthur Hinsvark, and 542 others. 2024 · 2024
Closest in time.
V-STaR: Training Verifiers for Self-Taught Reasoners
Arian Hosseini, Xingdi Yuan, Nikolay Malkin, Aaron Courville, Alessandro Sordoni, and Rishabh Agarwal. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Albert Q. Jiang, Wenda Li, and Mateja Jamnik. 2023 · 2023
Cited alongside, same era.
Arithmo-mistral-7b: Mathematical reasoning model
Ashvini Jindal. 2023 · 2023
Cited alongside, same era.
Efficient memory management for large language model serving with pagedattention
Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph E. Gonzalez, Hao Zhang, and Ion Stoica. 2023 · 2023
Cited alongside, same era.
leanprover-community/repl
Kim Morrison. 2023 · 2023
Cited alongside, same era.
Large-scale formal proof for the working mathematician - lessons learnt from the ALEXANDRIA project
Lawrence C. Paulson. 2023 · 2023
Cited alongside, same era.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. 2023 · 2023
Cited alongside, same era.
ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Zhangir Azerbayev, Bartosz Piotrowski, Hailey Schoelkopf, Edward W. Ayers, Dragomir Radev, and Jeremy Avigad. 2023a
Cited in the paper.
Jiewen Hu, Thomas Zhu, and Sean Welleck. 2024 · 2024
Closest in time.
Rethinking and Improving Autoformalization: Towards a Faithful Metric and a Dependency Retrieval-based Approach
Qi Liu, Xinhao Zheng, Xudong Lu, Qinxiang Cao, and Junchi Yan. 2024 · 2024
Closest in time.
Gflean: An autoformalisation framework for lean via gf
Shashank Pathak. 2024 · 2024
Closest in time.
Code llama: Open foundation models for code
Baptiste Rozière, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Romain Sauvestre, Tal Remez, Jérémy Rapin, Artyom Kozhevnikov, Ivan Evtimov, Joanna Bitton, Manish Bhatt, Cristian Canton Ferrer, Aaron Grattafiori, Wenhan Xiong, Alexandre Défossez, and 7 others. 2024 · 2024
Closest in time.
rahul3613/ProofNet-lean4
Rahul Vishwakarma. 2024 · 2024
Closest in time.
Huajian Xin, Z. Z. Ren, Junxiao Song, Zhihong Shao, Wanjia Zhao, Haocheng Wang, Bo Liu, Liyue Zhang, Xuan Lu, Qiushi Du, Wenjun Gao, Qihao Zhu, Dejian Yang, Zhibin Gou, Z. F. Wu, Fuli Luo, and Chong Ruan. 2024 · 2024
Closest in time.
leanblueprint
Patrick Massot. 2025 · 2025
Closest in time.