Fetching the paper…
Reading the bibliography…
How can Transformers model and learn enumerative geometry? What is a robust procedure for using Transformers in abductive knowledge discovery within a mathematician-machine collaboration? In this work, we introduce a Transformer-based approach to computational enumerative geometry, specifically targeting the computation of $\psi$-class intersection numbers on the moduli space of curves.
Towards an enumerative geometry of the moduli space of curves
David Mumford · 1983
Earlier work this paper cites.
Two-dimensional gravity and intersection theory on moduli space
Edward Witten · 1990
Earlier work this paper cites.
Topological strings in d < 1 d<1
Robbert Dijkgraaf, Herman L. Verlinde, and Erik P. Verlinde · 1991
Earlier work this paper cites.
Intersection theory on the moduli space of curves and the matrix Airy function
Maxim Kontsevich · 1992
Earlier work this paper cites.
Gromov–Witten invariants in algebraic geometry
Kai Behrend · 1997
Earlier work this paper cites.
Machine-learning applications of algorithmic randomness
Volodya Vovk, Alexander Gammerman, and Craig Saunders · 1999
Earlier work this paper cites.
Generating functions for intersection numbers on moduli spaces of curves
Andrei Okounkov · 2002
Earlier work this paper cites.
Enumerative Geometry and String Theory , volume 32
Sheldon Katz · 2006
Earlier work this paper cites.
Invariants of algebraic curves and topological expansion
Bertrand Eynard and Nicolas Orantin · 2007
Earlier work this paper cites.
Weil–Petersson volumes and intersection theory on the moduli space of curves
Maryam Mirzakhani · 2007
Earlier work this paper cites.
Conformal prediction with neural networks
Harris Papadopoulos, Volodya Vovk, and Alex Gammerman · 2007
Earlier work this paper cites.
The on-line encyclopedia of integer sequences
Neil J. A. Sloane · 2007
Earlier work this paper cites.
Growth of the number of simple closed geodesies on hyperbolic surfaces
Maryam Mirzakhani · 2008
Earlier work this paper cites.
A tutorial on conformal prediction
Glenn Shafer and Vladimir Vovk · 2008
Earlier work this paper cites.
Determinantal formulae and loop equations, 2009
Michel Bergère and Bertrand Eynard · 2009
Earlier work this paper cites.
Cut-and-join operator representation for Kontsevich–Witten tau-function
Alexander Alexandrov · 2011
Earlier work this paper cites.
Predicting numbers: An ai approach to solving number series
Marco Ragni and Andreas Klein · 2011
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Language modeling with gated convolutional networks
Yann N. Dauphin, Angela Fan, Michael Auli, and David Grangier · 2017
Earlier work this paper cites.
Deep network guided proof search
Sarah Loos, Geoffrey Irving, Christian Szegedy, and Cezary Kaliszyk · 2017
Earlier work this paper cites.
Taming the waves: sine as activation function in deep neural networks, 2017
Giambattista Parascandolo, Heikki Huttunen, and Tuomas Virtanen · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Canonical Correlation Analysis
Hervé Abdi, Vincent Guillemot, Aida Eslami, and Derek Beaton · 2018
Earlier work this paper cites.
What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni · 2018
Earlier work this paper cites.
Airy structures and symplectic geometry of topological recursion
Maxim Kontsevich and Yan Soibelman · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Branes with brains: exploring string vacua with deep reinforcement learning
James Halverson, Brent Nelson, and Fabian Ruehle · 2019
Earlier work this paper cites.
JT gravity as a matrix integral, 2019
Phil Saad, Stephen H. Shenker, and Douglas Stanford · 2019
Earlier work this paper cites.
Analysing mathematical reasoning abilities of neural models
David Saxton, Edward Grefenstette, Felix Hill, and Pushmeet Kohli · 2019
Earlier work this paper cites.
Quiver mutations, Seiberg duality, and machine learning
Jiakang Bao, Sebastián Franco, Yang-Hui He, Edward Hirst, Gregg Musiker, and Yan Xiao · 2020
Earlier work this paper cites.
Principal neighbourhood aggregation for graph nets, 2020
Gabriele Corso, Luca Cavalleri, Dominique Beaini, Pietro Liò, and Petar Veličković · 2020
Cited alongside, same era.
Compositionality decomposed: How do neural networks generalise?
Dieuwke Hupkes, Verna Dankers, Mathijs Mul, and Elia Bruni · 2020
Cited alongside, same era.
Graph representations for higher-order logic and theorem proving
Aditya Paliwal, Sarah Loos, Markus Rabe, Kshitij Bansal, and Christian Szegedy · 2020
Cited alongside, same era.
Causal mediation analysis for interpreting neural NLP: the case of gender bias, 2020
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Simas Sakenis, Jason Huang, Yaron Singer, and Stuart Shieber · 2020
Cited alongside, same era.
Learning to prove theorems by learning to generate theorems
Mingzhe Wang and Jia Deng · 2020
Cited alongside, same era.
Exploration of neural machine translation in autoformalization of mathematics in mizar
Learning knot invariants across dimensions
Jessica Craven, Mark Hughes, Vishnu Jejjala, and Arjun Kar · 2023
Later among the works it cites.
Simplifying polylogarithms with machine learning
Aurélien Dersy, Matthew D. Schwartz, and Xiaoyuan Zhang · 2023
Later among the works it cites.
A new formula for intersection numbers, 2023
Bertrand Eynard and Dimitrios Mitsios · 2023
Later among the works it cites.
Resurgent large genus asymptotics of intersection numbers, 2023
Bertrand Eynard, Elba Garcia-Failde, Alessandro Giacchetto, Paolo Gregori, and Danilo Lewański · 2023
Later among the works it cites.
Searching for ribbons with machine learning, 2023
Sergei Gukov, James Halverson, Ciprian Manolescu, and Fabian Ruehle · 2023
Later among the works it cites.
Machine learning in physics and geometry
Yang-Hui He, Elli Heyes, and Edward Hirst · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qingxiang Wang, Chad Brown, Cezary Kaliszyk, and Josef Urban · 2020
Cited alongside, same era.
Neural networks fail to learn periodic functions and how to fix it
Liu Ziyin, Tilman Hartwig, and Masahito Ueda · 2020
Cited alongside, same era.
Large genus asymptotics for intersection numbers and principal strata volumes of quadratic differentials
Amol Aggarwal · 2021
Cited alongside, same era.
Moduli-dependent Calabi–Yau and SU(3)-structure metrics from Machine Learning
Lara B. Anderson, Mathis Gerdes, James Gray, Sven Krippendorf, Nikhil Raghuram, and Fabian Ruehle · 2021
Cited alongside, same era.
Advancing mathematics by guiding human intuition with AI
Alex Davies, Petar Veličković, Lars Buesing, Sam Blackwell, Daniel Zheng, Nenad Tomašev, Richard Tanburn, Peter Battaglia, Charles Blundell, András Juhász, et al · 2021
Cited alongside, same era.
Masur–Veech volumes, frequencies of simple closed geodesics, and intersection numbers of moduli spaces of curves
Vincent Delecroix, Élise Goujard, Peter Zograf, and Anton Zorich · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Cited alongside, same era.
Multilingual mathematical autoformalization, 2023
Albert Q. Jiang, Wenda Li, and Mateja Jamnik · 2023
Later among the works it cites.
Emergent world representations: Exploring a sequence model trained on a synthetic task
Kenneth Li, Aspen K. Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg · 2023
Later among the works it cites.
Transformers are sample-efficient world models
Vincent Micheli, Eloi Alonso, and François Fleuret · 2023
Later among the works it cites.
Conformal prediction for time series, 2023
Chen Xu and Yao Xie · 2023
Later among the works it cites.
Learn from failure: Fine-tuning LLMs with trial-and-error data for intuitionistic propositional logic proving
Chenyang An, Zhibo Chen, Qihao Ye, Emily First, Letian Peng, Jiayun Zhang, Zihan Wang, Sorin Lerner, and Jingbo Shang · 2024
Closest in time.
Generating triangulations and fibrations with reinforcement learning, 2024
Per Berglund, Giorgi Butbaia, Yang-Hui He, Elli Heyes, Edward Hirst, and Vishnu Jejjala · 2024
Closest in time.
Transforming the bootstrap: Using transformers to compute scattering amplitudes in planar 𝒩 = 4 \mathcal{N}=4 super Yang–Mills theory, 2024
Tianji Cai, Garrett W. Merz, François Charton, Niklas Nolte, Matthias Wilhelm, Kyle Cranmer, and Lance J. Dixon · 2024
Closest in time.
Machine learning detects terminal singularities
Tom Coates, Alexander M. Kasprzyk, and Sara Veneziale · 2024
Closest in time.
Vision transformers need registers
Timothée Darcet, Maxime Oquab, Julien Mairal, and Piotr Bojanowski · 2024
Closest in time.
Machine learning assisted exploration for affine Deligne–Lusztig varieties
Bin Dong, Xuhua He, Pengfei Jin, Felix Schremmer, and Qingchao Yu · 2024
Closest in time.
How do language models bind entities in context?
Jiahai Feng and Jacob Steinhardt · 2024
Closest in time.
A survey on self-supervised learning: Algorithms, applications, and future trends, 2024
Jie Gui, Tuo Chen, Jing Zhang, Qiong Cao, Zhenan Sun, Hao Luo, and Dacheng Tao · 2024
Closest in time.
ATG: Benchmarking automated theorem generation for generative language models, 2024
Xiaohan Lin, Qingxing Cao, Yinya Huang, Zhicheng Yang, Zhengying Liu, Zhenguo Li, and Xiaodan Liang · 2024
Closest in time.
KAN: Kolmogorov–Arnold networks, 2024
Ziming Liu, Yixuan Wang, Sachin Vaidya, Fabian Ruehle, James Halverson, Marin Soljačić, Thomas Y. Hou, and Max Tegmark · 2024
Closest in time.
SNIP: Bridging mathematical symbolic and numeric realms with unified pre-training
Kazem Meidani, Parshin Shojaee, Chandan K. Reddy, and Amir Barati Farimani · 2024
Closest in time.
Magnushammer: A transformer-based approach to premise selection
Maciej Mikuła, Szymon Tworkowski, Szymon Antoniak, Bartosz Piotrowski, Albert Q. Jiang, Jin Peng Zhou, Christian Szegedy, Łukasz Kuciński, Piotr Miłoś, and Yuhuai Wu · 2024
Closest in time.
A survey of deep learning and foundation models for time series forecasting, 2024
John A. Miller, Mohammed Aldosari, Farah Saeed, Nasid Habib Barna, Subas Rana, I. Budak Arpinar, and Ninghao Liu · 2024
Closest in time.
Mathematical discoveries from program search with large language models
Bernardino Romera-Paredes, Mohammadamin Barekatain, Alexander Novikov, Matej Balog, M. Pawan Kumar, Emilien Dupont, Francisco J. R. Ruiz, Jordan S. Ellenberg, Pengming Wang, Omar Fawzi, Pushmeet Kohli, and Alhussein Fawzi · 2024
Closest in time.
Solving olympiad geometry without human demonstrations
Trieu H. Trinh, Yuhuai Wu, Quoc V. Le, He He, and Thang Luong · 2024
Closest in time.
Grokked transformers are implicit reasoners: a mechanistic journey to the edge of generalization
Boshi Wang, Xiang Yue, Yu Su, and Huan Sun · 2024
Closest in time.
Lean workbook: A large-scale Lean problem set formalized from natural language math problems, 2024
Huaiyuan Ying, Zijian Wu, Yihan Geng, Jiayu Wang, Dahua Lin, and Kai Chen · 2024
Closest in time.
Transformer-based models are not yet perfect at learning to emulate structural recursion, 2024
Dylan Zhang, Curt Tigges, Zory Zhang, Stella Biderman, Maxim Raginsky, and Talia Ringer · 2024
Closest in time.
Towards best practices of activation patching in language models: metrics and methods
Fred Zhang and Neel Nanda · 2024
Closest in time.