Fetching the paper…
Reading the bibliography…
We present an in-context learning agent for formal theorem-proving in environments like Lean and Coq.
Empirical explorations of the logic theory machine: a case study in heuristic
Newell, A., Shaw, J. C., and Simon, H. A · 1957
Earlier work this paper cites.
The use of explicit plans to guide inductive proofs
Bundy, A · 1988
Earlier work this paper cites.
Isabelle: A generic theorem prover
Paulson, L. C · 1994
Earlier work this paper cites.
The coq proof assistant a tutorial
Huet, G., Kahn, G., and Paulin-Mohring, C · 1997
Earlier work this paper cites.
E - a brainiac theorem prover
Schulz, S · 2002
Earlier work this paper cites.
Tps: A hybrid automatic-interactive system for developing proofs
Andrews, P. B. and Brown, C. E · 2006
Earlier work this paper cites.
Z3: an efficient smt solver
De Moura, L. and Bjørner, N · 2008
Earlier work this paper cites.
Formal verification of a realistic compiler
Leroy, X · 2009
Earlier work this paper cites.
Automatic proof and disproof in isabelle/hol
Blanchette, J. C., Bulwahn, L., and Nipkow, T · 2011
Earlier work this paper cites.
The Lean theorem prover (system description)
de Moura, L., Kong, S., Avigad, J., Van Doorn, F., and von Raumer, J · 2015
Earlier work this paper cites.
Three years of experience with sledgehammer, a practical link between automatic and interactive theorem provers
Paulson, L. and Blanchette, J · 2015
Earlier work this paper cites.
Gamepad: A learning environment for theorem proving
Huang, D., Dhariwal, P., Song, D., and Sutskever, I · 2019
Earlier work this paper cites.
Learning to prove theorems via interacting with proof assistants
Yang, K. and Deng, J · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Cited alongside, same era.
The lean mathematical library
mathlib Community, T · 2020
Cited alongside, same era.
Generative language modeling for automated theorem proving
Polu, S. and Sutskever, I · 2020
Cited alongside, same era.
Generating correctness proofs with neural networks
Sanchez-Stern, A., Alhessi, Y., Saul, L., and Lerner, S · 2020
Cited alongside, same era.
Proof artifact co-training for theorem proving with language models
Han, J. M., Rute, J., Wu, Y., Ayers, E. W., and Polu, S · 2021
Cited alongside, same era.
Measuring mathematical problem solving with the math dataset
React: Synergizing reasoning and acting in language models
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K., and Cao, Y · 2022
Later among the works it cites.
Llemma: An open language model for mathematics, 2023
Azerbayev, Z., Schoelkopf, H., Paster, K., Santos, M. D., McAleer, S., Jiang, A. Q., Deng, J., Biderman, S., and Welleck, S · 2023
Closest in time.
Baldur: whole-proof generation and repair with large language models
First, E., Rabe, M. N., Ringer, T., and Brun, Y · 2023
Closest in time.
Code llama: Open foundation models for code
Roziere, B., Gehring, J., Gloeckle, F., Sootla, S., Gat, I., Tan, X. E., Adi, Y., Liu, J., Remez, T., Rapin, J., et al · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hendrycks, D., Burns, C., Kadavath, S., Arora, A., Basart, S., Tang, E., Song, D., and Steinhardt, J · 2021
Cited alongside, same era.
Lisa: Language models of isabelle proofs
Jiang, A. Q., Li, W., Han, J. M., and Wu, Y · 2021
Cited alongside, same era.
Zero-shot text-to-image generation
Ramesh, A., Pavlov, M., Goh, G., Gray, S., Voss, C., Radford, A., Chen, M., and Sutskever, I · 2021
Cited alongside, same era.
Minif2f: a cross-system benchmark for formal olympiad-level mathematics
Zheng, K., Han, J. M., and Polu, S · 2021
Cited alongside, same era.
Hypertree proof search for neural theorem proving
Lample, G., Lacroix, T., Lachaux, M.-A., Rodriguez, A., Hayat, A., Lavril, T., Ebner, G., and Martinet, X · 2022
Cited alongside, same era.
Formal mathematics statement curriculum learning
Polu, S., Han, J. M., Zheng, K., Baksys, M., Babuschkin, I., and Sutskever, I · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Xia, F., Chi, E., Le, Q. V., Zhou, D., et al · 2022
Cited alongside, same era.
Schick, T., Dwivedi-Yu, J., Dessì, R., Raileanu, R., Lomeli, M., Zettlemoyer, L., Cancedda, N., and Scialom, T · 2023
Closest in time.
Reflexion: Language agents with verbal reinforcement learning
Shinn, N., Cassano, F., Labash, B., Gopinath, A., Narasimhan, K., and Yao, S · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozière, B., Goyal, N., Hambro, E., Azhar, F., et al · 2023
Closest in time.
Leandojo: Theorem proving with retrieval-augmented language models
Yang, K., Swope, A. M., Gu, A., Chalamala, R., Song, P., Yu, S., Godil, S., Prenger, R., and Anandkumar, A · 2023
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Yao, S., Yu, D., Zhao, J., Shafran, I., Griffiths, T. L., Cao, Y., and Narasimhan, K · 2023
Closest in time.
Decomposing the enigma: Subgoal-based demonstration learning for formal theorem proving
Zhao, X., Li, W., and Kong, L · 2023
Closest in time.
Lyra: Orchestrating dual correction in automated theorem proving
Zheng, C., Wang, H., Xie, E., Liu, Z., Sun, J., Xin, H., Shen, J., Li, Z., and Li, Y · 2023
Closest in time.
Solving olympiad geometry without human demonstrations
Trinh, T. H., Wu, Y., Le, Q. V., He, H., and Luong, T · 2024
Closest in time.