Fetching the paper…
Reading the bibliography…
We explore the use of expert iteration in the context of language modeling applied to formal mathematics.
Learning to reason in large theories without imitation
Bansal, K., Loos, S. M., Rabe, M. N., and Szegedy, C · 1905
Earlier work this paper cites.
Making proofs without modus ponens: An introduction to the combinatorics and complexity of cut elimination
Carbone, A. and Semmes, S · 1996
Earlier work this paper cites.
A survey of monte carlo tree search methods
Browne, C. B., Powley, E., Whitehouse, D., Lucas, S. M., Cowling, P. I., Rohlfshagen, P., Tavener, S., Perez, D., Samothrakis, S., and Colton, S · 2012
Earlier work this paper cites.
The lean theorem prover (system description)
de Moura, L., Kong, S., Avigad, J., Van Doorn, F., and von Raumer, J · 2015
Earlier work this paper cites.
Deepmath-deep sequence models for premise selection
Irving, G., Szegedy, C., Alemi, A. A., Eén, N., Chollet, F., and Urban, J · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., et al · 2016
Earlier work this paper cites.
Minif2f: a cross-system benchmark for formal olympiad-level mathematics
Zheng, K., Han, J. M., and Polu, S · 2016
Earlier work this paper cites.
Deep network guided proof search
Loos, S., Irving, G., Szegedy, C., and Kaliszyk, C · 2017
Earlier work this paper cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
Silver, D., Hubert, T., Schrittwieser, J., Antonoglou, I., Lai, M., Guez, A., Lanctot, M., Sifre, L., Kumaran, D., Graepel, T., et al · 2017
Earlier work this paper cites.
Premise selection for theorem proving by deep graph embedding
Wang, M., Tang, Y., Wang, J., and Deng, J · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2018
Earlier work this paper cites.
Gamepad: A learning environment for theorem proving
Huang, D., Dhariwal, P., Song, D., and Sutskever, I · 2018
Cited alongside, same era.
Learning a sat solver from single-bit supervision
Selsam, D., Lamm, M., Bünz, B., Liang, P., de Moura, L., and Dill, D. L · 2018
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Berner, C., Brockman, G., Chan, B., Cheung, V., Dębiak, P., Dennison, C., Farhi, D., Fischer, Q., Hashme, S., Hesse, C., et al · 2019
Cited alongside, same era.
Lean perfectoid spaces
Buzzard, K., Commelin, J., and Massot, P · 2019
Cited alongside, same era.
A style-based generator architecture for generative adversarial networks
Karras, T., Laine, S., and Aila, T · 2019
Cited alongside, same era.
Generative language modeling for automated theorem proving
Polu, S. and Sutskever, I · 2020
Later among the works it cites.
Mathematical reasoning via self-supervised skip-tree training
Rabe, M. N., Lee, D., Bansal, K., and Szegedy, C · 2020
Later among the works it cites.
Liquid tensor experiment
Scholze, P · 2020
Later among the works it cites.
First neural conjecturing datasets and experiments
Urban, J. and Jakubův, J · 2020
Later among the works it cites.
Int: An inequality benchmark for evaluating generalization in theorem proving
Wu, Y., Jiang, A. Q., Ba, J., and Grosse, R · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Metamath: A Computer Language for Pure Mathematics , 2019
Megill, N. D. and Wheeler, D. A · 2019
Cited alongside, same era.
Efficientnet: Rethinking model scaling for convolutional neural networks
Tan, M. and Le, Q · 2019
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Cited alongside, same era.
Learning to prove theorems via interacting with proof assistants
Yang, K. and Deng, J · 2019
Cited alongside, same era.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Cited alongside, same era.
Scaling laws for neural language models
Kaplan, J., McCandlish, S., Henighan, T., Brown, T. B., Chess, B., Child, R., Gray, S., Radford, A., Wu, J., and Amodei, D · 2020
Cited alongside, same era.
Modelling high-level mathematical reasoning in mechanised declarative proofs
Li, W., Yu, L., Wu, Y., and Paulson, L. C · 2020
Cited alongside, same era.
Subgoal search for complex reasoning tasks
Czechowski, K., Odrzygóźdź, T., Zbysiński, M., Zawalski, M., Olejnik, K., Wu, Y., Kucinski, L., and Miłoś, P · 2021
Later among the works it cites.
Training a first-order theorem prover from synthetic data
Firoiu, V., Aygun, E., Anand, A., Ahmed, Z., Glorot, X., Orseau, L., Zhang, L., Precup, D., and Mourad, S · 2021
Later among the works it cites.
Proof artifact co-training for theorem proving with language models
Han, J. M., Rute, J., Wu, Y., Ayers, E. W., and Polu, S · 2021
Later among the works it cites.
Measuring mathematical problem solving with the math dataset
Hendrycks, D., Burns, C., Kadavath, S., Arora, A., Basart, S., Tang, E., Song, D., and Steinhardt, J · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Ramesh, A., Pavlov, M., Goh, G., Gray, S., Voss, C., Radford, A., Chen, M., and Sutskever, I · 2021
Later among the works it cites.