Fetching the paper…
Reading the bibliography…
The research in AI-based formal mathematical reasoning has shown an unstoppable growth trend.
A survey of information requirements analysis techniques
William Taggart Jr and Marvin O Tharp. 1977 · 1977
Earlier work this paper cites.
The calculus of constructions
Gérard Huet. 1986 · 1986
Earlier work this paper cites.
Software requirements: analysis and specification
Alan M Davis. 1990 · 1990
Earlier work this paper cites.
Goal-based requirements analysis
Annie I Anton. 1996 · 1996
Earlier work this paper cites.
Model checking
Edmund M Clarke. 1997 · 1997
Earlier work this paper cites.
Model checking tla+ specifications
Yuan Yu, Panagiotis Manolios, and Leslie Lamport. 1999 · 1999
Earlier work this paper cites.
Specifying systems: the tla+ language and tools for hardware and software engineers
Leslie Lamport. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Natural language premise selection: Finding supporting statements for mathematical text
Deborah Ferreira and André Freitas. 2020 · 2004
Earlier work this paper cites.
An automated tool for generating uml models from natural language requirements
Deva Kumar Deeptimahanti and Muhammad Ali Babar. 2009 · 2009
Earlier work this paper cites.
Software model checking
Ranjit Jhala and Rupak Majumdar. 2009 · 2009
Earlier work this paper cites.
sel4: formal verification of an os kernel
Gerwin Klein, Kevin Elphinstone, Gernot Heiser, June Andronick, David Cock, Philip Derrin, Dhammika Elkaduwe, Kai Engelhardt, Rafal Kolanski, Michael Norrish, Thomas Sewell, Harvey Tuch, and Simon Winwood. 2009 · 2009
Earlier work this paper cites.
System requirements analysis
Jeffrey O Grady. 2010 · 2010
Earlier work this paper cites.
Dafny: An automatic program verifier for functional correctness
K Rustan M Leino. 2010 · 2010
Earlier work this paper cites.
Verified software toolchain
Andrew W. Appel. 2011 · 2011
Earlier work this paper cites.
Frama-c: A software analysis perspective
Pascal Cuoq, Florent Kirchner, Nikolai Kosmatov, Virgile Prevosto, Julien Signoles, and Boris Yakobowski. 2012 · 2012
Earlier work this paper cites.
Translating software requirements from natural language to formal specification
Agung Fatwanto. 2012 · 2012
Earlier work this paper cites.
Behavioral interface specification languages
John Hatcliff, Gary T. Leavens, K. Rustan M. Leino, Peter Müller, and Matthew Parkinson. 2012 · 2012
Earlier work this paper cites.
Feature model extraction from large collections of informal product descriptions
Jean-Marc Davril, Edouard Delfosse, Negar Hariri, Mathieu Acher, Jane Cleland-Huang, and Patrick Heymans. 2013 · 2013
Earlier work this paper cites.
Cgmurphi: Automatic synthesis of numerical controllers for nonlinear hybrid systems
Giuseppe Della Penna, Benedetto Intrigila, Daniele Magazzeni, Igor Melatti, and Enrico Tronci. 2013 · 2013
Earlier work this paper cites.
Formal verification of klibc with the wp frama-c plug-in
Nuno Carvalho, Cristiano da Silva Sousa, Jorge Sousa Pinto, and Aaron Tomb. 2014 · 2014
Earlier work this paper cites.
Ironclad apps: { \{ End-to-End } \} security via automated { \{ Full-System } \} verification
Chris Hawblitzel, Jon Howell, Jacob R Lorch, Arjun Narayan, Bryan Parno, Danfeng Zhang, and Brian Zill. 2014 · 2014
Earlier work this paper cites.
Code completion with statistical language models
Veselin Raychev, Martin Vechev, and Eran Yahav. 2014 · 2014
Earlier work this paper cites.
Compcert-a formally verified optimizing compiler
Xavier Leroy, Sandrine Blazy, Daniel Kästner, Bernhard Schommer, Markus Pister, and Christian Ferdinand. 2016 · 2016
Earlier work this paper cites.
Environment modeling-based requirements engineering for software intensive systems
Zhi Jin. 2017 · 2017
Earlier work this paper cites.
Pythia: Ai-assisted code completion system
Alexey Svyatkovskiy, Ying Zhao, Shengyu Fu, and Neel Sundaresan. 2019 · 2019
Earlier work this paper cites.
Learning to prove theorems via interacting with proof assistants
Kaiyu Yang and Jia Deng. 2019 · 2019
Earlier work this paper cites.
Mathematics in lean
Jeremy Avigad, Kevin Buzzard, Robert Y Lewis, and Patrick Massot. 2020 · 2020
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Earlier work this paper cites.
The lean mathematical library
The mathlib Community. 2020 · 2020
Cited alongside, same era.
Acsl: Ansi/iso c specification
Patrick Baudin, Jean-Christophe Filliâtre, Claude Marché, Benjamin Monate, Yannick Moy, and Virgile Prevosto. 2021 · 2021
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba. 2021 · 2021
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Camels in a changing climate: Enhancing lm adaptation with tulu 2
Hamish Ivison, Yizhong Wang, Valentina Pyatkin, Nathan Lambert, Matthew Peters, Pradeep Dasigi, Joel Jang, David Wadden, Noah A Smith, Iz Beltagy, et al. 2023 · 2023
Later among the works it cites.
Swe-bench: Can language models resolve real-world github issues?
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan. 2023 · 2023
Later among the works it cites.
Efficient memory management for large language model serving with pagedattention
Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph E. Gonzalez, Hao Zhang, and Ion Stoica. 2023 · 2023
Later among the works it cites.
Bert is not the count: Learning to match mathematical statements with proofs
Weixian Waylon Li, Yftah Ziser, Maximin Coavoux, and Shay B Cohen. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Measuring mathematical problem solving with the math dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Cited alongside, same era.
imer: Iterative process of entity relationship and business process model extraction from the requirements
Muhammad Javed and Yuqing Lin. 2021 · 2021
Cited alongside, same era.
The lean 4 theorem prover and programming language
Leonardo de Moura and Sebastian Ullrich. 2021 · 2021
Cited alongside, same era.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, Stephen H Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Teven Le Scao, Arun Raja, et al. 2021 · 2021
Cited alongside, same era.
Intelligent requirements elicitation and modeling: A literature review
Ye Wang, JW Chen, Xin Xia, and B Jiang. 2021 · 2021
Cited alongside, same era.
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Y Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M Dai, and Quoc V Le. 2021 · 2021
Cited alongside, same era.
Naturalproofs: Mathematical theorem proving in natural language
Sean Welleck, Jiacheng Liu, Ronan Le Bras, Hannaneh Hajishirzi, Yejin Choi, and Kyunghyun Cho. 2021 · 2021
Cited alongside, same era.
Minif2f: a cross-system benchmark for formal olympiad-level mathematics
Kunhao Zheng, Jesse Michael Han, and Stanislas Polu. 2021 · 2021
Cited alongside, same era.
Chengwu Liu, Jianhao Shen, Huajian Xin, Zhengying Liu, Ye Yuan, Haiming Wang, Wei Ju, Chuanyang Zheng, Yichun Yin, Lin Li, Ming Zhang, and Qun Liu. 2023 · 2023
Later among the works it cites.
Formalizing the proof of pfr in lean4 using blueprint: a short tour
Terence Tao, Yael Dillies, and Bhavik Mehta. 2023 · 2023
Later among the works it cites.
Norbert Tihanyi, Ridhi Jain, Yiannis Charalambous, Mohamed Amine Ferrag, Youcheng Sun, and Lucas C Cordeiro. 2023 · 2023
Later among the works it cites.
How far can camels go? exploring the state of instruction tuning on open resources
Yizhong Wang, Hamish Ivison, Pradeep Dasigi, Jack Hessel, Tushar Khot, Khyathi Chandu, David Wadden, Kelsey MacMillan, Noah A Smith, Iz Beltagy, et al. 2023 · 2023
Later among the works it cites.
Leandojo: Theorem proving with retrieval-augmented language models
Kaiyu Yang, Aidan M Swope, Alex Gu, Rahul Chalamala, Peiyang Song, Shixing Yu, Saad Godil, Ryan Prenger, and Anima Anandkumar. 2023 · 2023
Later among the works it cites.
Ai achieves silver-medal standard solving international mathematical olympiad problems.’25 july 2024
DeepMind AlphaProof and AlphaGeometry Teams. 2024 · 2024
Later among the works it cites.
Deepseek-coder: When the large language model meets programming–the rise of code intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Yu Wu, YK Li, et al. 2024 · 2024
Later among the works it cites.
Qwen2. 5-coder technical report
Binyuan Hui, Jian Yang, Zeyu Cui, Jiaxi Yang, Dayiheng Liu, Lei Zhang, Tianyu Liu, Jiajun Zhang, Bowen Yu, Keming Lu, et al. 2024 · 2024
Later among the works it cites.
Large language models for code completion: A systematic literature review
Rasha Ahmad Husein, Hala Aburajouh, and Cagatay Catal. 2024 · 2024
Later among the works it cites.
An evaluation of requirements modeling for cyber-physical systems via llms
Dongming Jin, Shengxin Zhao, Zhi Jin, Xiaohong Chen, Chunhui Wang, Zheng Fang, and Hongbin Xiao. 2024 · 2024
Later among the works it cites.
T \ \backslash " ulu 3: Pushing frontiers in open language model post-training
Nathan Lambert, Jacob Morrison, Valentina Pyatkin, Shengyi Huang, Hamish Ivison, Faeze Brahman, Lester James V Miranda, Alisa Liu, Nouha Dziri, Shane Lyu, et al. 2024 · 2024
Later among the works it cites.
A survey on deep learning for theorem proving
Zhaoyu Li, Jialiang Sun, Logan Murphy, Qidong Su, Zenan Li, Xian Zhang, Kaiyu Yang, and Xujie Si. 2024 · 2024
Later among the works it cites.
Starcoder 2 and the stack v2: The next generation
Anton Lozhkov, Raymond Li, Loubna Ben Allal, Federico Cassano, Joel Lamy-Poirier, Nouamane Tazi, Ao Tang, Dmytro Pykhtar, Jiawei Liu, Yuxiang Wei, et al. 2024 · 2024
Later among the works it cites.
Specgen: Automated generation of formal program specifications via large language models
Lezhi Ma, Shangqing Liu, Yi Li, Xiaofei Xie, and Lei Bu. 2024 · 2024
Later among the works it cites.
Introducing meta llama 3: The most capable openly available llm to date
AI Meta. 2024 · 2024
Later among the works it cites.
Laurel: Generating dafny assertions using large language models
Eric Mugnier, Emmanuel Anaya Gonzalez, Ranjit Jhala, Nadia Polikarpova, and Yuanyuan Zhou. 2024 · 2024
Later among the works it cites.
Towards large language models as copilots for theorem proving in lean
Peiyang Song, Kaiyu Yang, and Anima Anandkumar. 2024 · 2024
Later among the works it cites.
Calibration and correctness of language models for code
Claudio Spiess, David Gros, Kunal Suresh Pai, Michael Pradel, Md Rafiqul Islam Rabin, Amin Alipour, Susmit Jha, Prem Devanbu, and Toufique Ahmed. 2024 · 2024
Later among the works it cites.
Clover: Closed-loop verifiable code generation
Chuyue Sun, Ying Sheng, Oded Padon, and Clark Barrett. 2024 · 2024
Later among the works it cites.
Ai will become mathematicians’ “co-pilot”’
Terence Tao. 2024 · 2024
Later among the works it cites.
Solving olympiad geometry without human demonstrations
Trieu H Trinh, Yuhuai Wu, Quoc V Le, He He, and Thang Luong. 2024 · 2024
Later among the works it cites.
Theoremllama: Transforming general-purpose llms into lean4 experts
Ruida Wang, Jipeng Zhang, Yizhen Jia, Rui Pan, Shizhe Diao, Renjie Pi, and Tong Zhang. 2024 · 2024
Later among the works it cites.
Enchanting program specification synthesis by large language models using static analysis and program verification
Cheng Wen, Jialun Cao, Jie Su, Zhiwu Xu, Shengchao Qin, Mengda He, Haokun Li, Shing-Chi Cheung, and Cong Tian. 2024 · 2024
Later among the works it cites.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al. 2025 · 2025
Closest in time.