Fetching the paper…
Reading the bibliography…
The dot product attention mechanism, originally designed for natural language processing tasks, is a cornerstone of modern Transformers.
Cluster decomposition properties of the s s matrix
Eyvind H. Wichmann and James H. Crichton · 1963
Earlier work this paper cites.
Exact ground state of a quantum mechanical antiferromagnet
B.S. Shastry and B. Sutherland · 1981
Earlier work this paper cites.
Magnetic order and disorder in the frustrated quantum heisenberg antiferromagnet in two dimensions
H. J. Schulz, T. A.L. Ziman, and D. Poilblanc · 1996
Earlier work this paper cites.
Finite-size scaling of the ground-state parameters of the two-dimensional heisenberg model
Anders W. Sandvik · 1997
Earlier work this paper cites.
Green function monte carlo with stochastic reconfiguration
Sandro Sorella · 1998
Earlier work this paper cites.
Why natural gradient?
S. Amari and S.C. Douglas · 1998
Earlier work this paper cites.
Numerical study of the two-dimensional heisenberg model using a green function monte carlo technique with a fixed number of walkers
Matteo Calandra Buonaura and Sandro Sorella · 1998
Earlier work this paper cites.
What is quantum field theory, and what did we think it was?
Steven Weinberg · 1999
Earlier work this paper cites.
Wave function optimization in the variational monte carlo method
Sandro Sorella · 2005
Earlier work this paper cites.
Hierarchical structures induce long-range dynamical correlations in written texts
E. Alvarez-Lacalle, B. Dorow, J.-P. Eckmann, and E. Moses · 2006
Earlier work this paper cites.
Computational studies of quantum spin systems
A.W. Sandvik · 2010
Earlier work this paper cites.
On the origin of long-range correlations in texts
Eduardo G. Altmann, Giampaolo Cristadoro, and Mirko Degli Esposti · 2012
Earlier work this paper cites.
Direct evidence for a gapless Z 2 {Z}_{2} spin liquid by frustrating néel antiferromagnetism
Wen-Jun Hu, Federico Becca, Alberto Parola, and Sandro Sorella · 2013
Earlier work this paper cites.
Plaquette ordered phase and quantum phase diagram in the spin- 1 2 \frac{1}{2} J 1 − J 2 {J}_{1}\text{$-$}{J}_{2} square heisenberg model
Shou-Shu Gong, Wei Zhu, D. N. Sheng, Olexei I. Motrunich, and Matthew P. A. Fisher · 2014
Earlier work this paper cites.
Quantum spin liquids: a review
Lucile Savary and Leon Balents · 2016
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A.N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Solving the quantum many-body problem with artificial neural networks
G. Carleo and M. Troyer · 2017
Earlier work this paper cites.
Quantum Monte Carlo Approaches for Correlated Systems
F. Becca and S. Sorella · 2017
Earlier work this paper cites.
4-spin plaquette singlet state in the shastry–sutherland compound srcu2(bo3)2
M. E. Zayed, Ch. Rüegg, J. Larrea J., A. M. Läuchli, C. Panagopoulos, S. S. Saxena, M. Ellerby, D. F. McMorrow, Th. Strässle, S. Klotz, G. Hamel, R. A. Sadykov, V. Pomjakushin, M. Boehm, M. Jiménez–Ruiz, A. Schneidewind, E. Pomjakushina, M. Stingaciu, K. Conder, and H. M. Rønnow · 2017
Earlier work this paper cites.
Improving language understanding by generative pre-training, 2018
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever · 2018
Earlier work this paper cites.
JAX: composable transformations of Python+NumPy programs, 2018
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang · 2018
Earlier work this paper cites.
Neural-network quantum states, string-bond states, and chiral topological states
Ivan Glasser, Nicola Pancotti, Moritz August, Ivan D. Rodriguez, and J. Ignacio Cirac · 2018
Cited alongside, same era.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Cited alongside, same era.
What does bert look at? an analysis of bert’s attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning · 2019
Fermionic wave functions from neural-network constrained hidden states
Javier Robledo Moreno, Giuseppe Carleo, Antoine Georges, and James Stokes · 2022
Later among the works it cites.
NetKet 3: Machine Learning Toolbox for Many-Body Quantum Systems
Filippo Vicentini, Damian Hofmann, Attila Szabó, Dian Wu, Christopher Roth, Clemens Giuliani, Gabriel Pescia, Jannes Nys, Vladimir Vargas-Calderón, Nikita Astrakhantsev, and Giuseppe Carleo · 2022
Later among the works it cites.
Variational monte carlo with large patched transformers
Kyle Sprague and Stefanie Czischek · 2023
Later among the works it cites.
Transformer wave function for the shastry-sutherland model: emergence of a spin-liquid phase
Luciano Loris Viteritti, Riccardo Rende, Alberto Parola, Sebastian Goldt, and Federico Becca · 2023
Later among the works it cites.
Gauge-invariant and anyonic-symmetric autoregressive neural network for quantum lattice models
Di Luo, Zhuo Chen, Kaiwen Hu, Zhizhen Zhao, Vera Mikyoung Hur, and Bryan K. Clark · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Fisher information and natural gradient learning in random deep networks
Shunichi Amari, Ryo Karakida, and Masafumi Oizumi · 2019
Cited alongside, same era.
Modern Quantum Mechanics
J. J. Sakurai and Jim Napolitano · 2020
Cited alongside, same era.
Geometry of learning neural quantum states
Chae-Yeun Park and Michael J. Kastoryano · 2020
Cited alongside, same era.
Ab initio solution of the many-electron schrödinger equation with deep neural networks
David Pfau, James S. Spencer, Alexander G. D. G. Matthews, and W. M. C. Foulkes · 2020
Cited alongside, same era.
On layer normalization in the transformer architecture
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tie-Yan Liu · 2020
Cited alongside, same era.
Do transformers need deep long-range memory?
Jack Rae and Ali Razavi · 2020
Cited alongside, same era.
Highly accurate protein structure prediction with alphafold
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Žídek, Anna Potapenko, Alex Bridgland, Clemens Meyer, Simon A. A. Kohl, Andrew J. Ballard, Andrew Cowie, Bernardino Romera-Paredes, Stanislav Nikolov, Rishub Jain, Jonas Adler, Trevor Back, Stig Petersen, David Reiman, Ellen Clancy, Michal Zielinski, Martin Steinegger, Michalina Pacholska, Tamas Berghammer, Sebastian Bodenstein, David Silver, Oriol Vinyals, Andrew W. Senior, Koray Kavukcuoglu, Pushmeet Kohli, and Demis Hassabis · 2021
Cited alongside, same era.
Later among the works it cites.
Transformer variational wave functions for frustrated quantum spin systems
Luciano Loris Viteritti, Riccardo Rende, and Federico Becca · 2023
Later among the works it cites.
A self-attention ansatz for ab-initio quantum chemistry
Ingrid von Glehn, James S. Spencer, and David Pfau · 2023
Later among the works it cites.
Efficient optimization of deep neural quantum states toward machine precision
Ao Chen and Markus Heyl · 2023
Later among the works it cites.
High-accuracy variational monte carlo for frustrated magnets with deep neural networks
Christopher Roth, Attila Szabó, and Allan H. MacDonald · 2023
Later among the works it cites.
Neural-network quantum states for ultra-cold fermi gases
Jane Kim, Gabriel Pescia, Bryce Fore, Jannes Nys, Giuseppe Carleo, Stefano Gandolfi, Morten Hjorth-Jensen, and Alessandro Lovato · 2023
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2023
Later among the works it cites.
Optimizing design choices for neural quantum states
Moritz Reh, Markus Schmitt, and Martin Gärttner · 2023
Later among the works it cites.
The nlp task effectiveness of long-range transformers, 2023
Guanghui Qin, Yukun Feng, and Benjamin Van Durme · 2023
Later among the works it cites.
OpenAI · 2024
Closest in time.
Language models for quantum simulation
Roger G. Melko and Juan Carrasquilla · 2024
Closest in time.
A simple linear algebra identity to optimize large-scale neural network quantum states
Riccardo Rende, Luciano Loris Viteritti, Lorenzo Bardone, Federico Becca, and Sebastian Goldt · 2024
Closest in time.
From architectures to applications: A review of neural quantum states
Hannah Lange, Anka Van de Walle, Atiye Abedinnia, and Annabelle Bohrdt · 2024
Closest in time.
Accurate neural quantum states for interacting lattice bosons
Zakari Denis and Giuseppe Carleo · 2024
Closest in time.
Ab-initio variational wave functions for the time-dependent many-electron schrödinger equation
Jannes Nys, Gabriel Pescia, and Giuseppe Carleo · 2024
Closest in time.
Mapping of attention mechanisms to a generalized potts model
Riccardo Rende, Federica Gerace, Alessandro Laio, and Sebastian Goldt · 2024
Closest in time.
Fine-tuning neural network quantum states
Riccardo Rende, Sebastian Goldt, Federico Becca, and Luciano Loris Viteritti · 2024
Closest in time.