Fetching the paper…
Reading the bibliography…
In this paper, we apply the self-attention from the state-of-the-art Transformer in Attention Is All You Need for the first time to a data-driven operator learning problem related to partial differential equations.
Survey of the stability of linear finite difference equations
Peter D Lax and Robert D Richtmyer · 1956
Earlier work this paper cites.
On the partial difference equations of mathematical physics
Richard Courant, Kurt Friedrichs, and Hans Lewy · 1967
Earlier work this paper cites.
Multiple-precision zero-finding methods and the complexity of elementary function evaluation
Richard P Brent · 1976
Earlier work this paper cites.
Finite element approximation of the Navier-Stokes equations
Vivette Girault and P-A Raviart · 1979
Earlier work this paper cites.
An identification problem for an elliptic equation in two variables
Giovanni Alessandrini · 1986
Earlier work this paper cites.
The Fourier transform and its applications , volume 31999
Ronald Newbold Bracewell and Ronald N Bracewell · 1986
Earlier work this paper cites.
Equivalence of Nyström’s method and Fourier methods for the numerical solution of Fredholm integral equations
Jean-Paul Berrut and Manfred R Trummer · 1987
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
Kurt Hornik, Maxwell Stinchcombe, and Halbert White · 1989
Earlier work this paper cites.
Universal approximation to nonlinear operators by neural networks with arbitrary activation functions and its application to dynamical systems
Tianping Chen and Hong Chen · 1995
Earlier work this paper cites.
Elliptic Partial Differential Equations of Second Order
D. Gilbarg and N.S. Trudinger · 2001
Earlier work this paper cites.
The finite element method for elliptic problems
Philippe G Ciarlet · 2002
Earlier work this paper cites.
Identification of discontinuous coefficients in elliptic problems using total variation regularization
Tony F Chan and Xue-Cheng Tai · 2003
Earlier work this paper cites.
Finite element methods for Maxwell’s equations
Peter Monk et al · 2003
Earlier work this paper cites.
Theory and Practice of Finite Elements
Alexandre Ern and Jean-Luc Guermond · 2004
Earlier work this paper cites.
Matplotlib: A 2d graphics environment
J. D. Hunter · 2007
Earlier work this paper cites.
Python for scientific computing
Travis E. Oliphant · 2007
Earlier work this paper cites.
The mathematical theory of finite element methods , volume 15
Susanne C Brenner and Ridgway Scott · 2008
Earlier work this paper cites.
i i FEM: an integrated finite element methods package in MATLAB
Long Chen · 2008
Earlier work this paper cites.
Python 3 Reference Manual
Guido Van Rossum and Fred L. Drake · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
An Introduction to the Mathematical Theory of Inverse Problems
A. Kirsch · 2011
Earlier work this paper cites.
Linear and nonlinear functional analysis with applications , volume 130
Philippe G Ciarlet · 2013
Earlier work this paper cites.
Multi-Grid Methods and Applications
W. Hackbusch · 2013
Earlier work this paper cites.
Chebfun guide, 2014
Tobin A Driscoll, Nicholas Hale, and Lloyd N Trefethen · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyung Hyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
plotly, 2015
Plotly Technologies Inc · 2015
Earlier work this paper cites.
Banach space projections and Petrov–Galerkin estimates
Ari Stern · 2015
Earlier work this paper cites.
Geometric Integration Theory
H. Whitney · 2015
Earlier work this paper cites.
A cheap linear attention mechanism with fast lookups and fixed-size representations
Alexandre de Brébisson and Pascal Vincent · 2016
Earlier work this paper cites.
Convolutional neural networks for steady flow approximation
Xiaoxiao Guo, Wei Li, and Francesco Iorio · 2016
Cited alongside, same era.
Identity mappings in deep residual networks
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Instance normalization: The missing ingredient for fast stylization
Dmitry Ulyanov, Andrea Vedaldi, and Victor Lempitsky · 2016
Cited alongside, same era.
Solving ill-posed inverse problems using iterative deep neural networks
Jonas Adler and Ozan Öktem · 2017
Cited alongside, same era.
OpenNMT: Open-Source Toolkit for Neural Machine Translation
Guillaume Klein, Yoon Kim, Yuntian Deng, Jean Senellart, and Alexander M. Rush · 2017
Cited alongside, same era.
Array programming with NumPy
Charles R. Harris, K. Jarrod Millman, Stéfan J. van der Walt, Ralf Gommers, Pauli Virtanen, David Cournapeau, Eric Wieser, Julian Taylor, Sebastian Berg, Nathaniel J. Smith, Robert Kern, Matti Picus, Stephan Hoyer, Marten H. van Kerkwijk, Matthew Brett, Allan Haldane, Jaime Fernández del Río, Mark Wiebe, Pearu Peterson, Pierre Gérard-Marchant, Kevin Sheppard, Tyler Reddy, Warren Weckesser, Hameer Abbasi, Christoph Gohlke, and Travis E. Oliphant · 2020
Later among the works it cites.
Transformers are RNNs: Fast autoregressive transformers with linear attention
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and François Fleuret · 2020
Later among the works it cites.
Approximation rates for neural networks with general activation functions
Jonathan W Siegel and Jinchao Xu · 2020
Later among the works it cites.
Lagrangian fluid simulation with continuous convolutions
Benjamin Ummenhofer, Lukas Prantl, Nils Thuerey, and Vladlen Koltun · 2020
Later among the works it cites.
Linformer: Self-attention with linear complexity
Sinong Wang, Belinda Z Li, Madian Khabsa, Han Fang, and Hao Ma · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Prajit Ramachandran, Barret Zoph, and Quoc V Le · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Algebraic multigrid methods
Jinchao Xu and Ludmil Zikatanov · 2017
Cited alongside, same era.
Adaptive sampled softmax with kernel based sampling
Guy Blanc and Steffen Rendle · 2018
Cited alongside, same era.
Resnet with one-neuron hidden layers is a universal approximator
Hongzhou Lin and Stefanie Jegelka · 2018
Cited alongside, same era.
A probabilistic framework for multi-view feature learning with many-to-many associations via neural networks
Akifumi Okuno, Tetsuya Hada, and Hidetoshi Shimodaira · 2018
Cited alongside, same era.
Bayesian deep convolutional encoder–decoder networks for surrogate modeling and uncertainty quantification
Yinhao Zhu and Nicholas Zabaras · 2018
Cited alongside, same era.
Later among the works it cites.
On layer normalization in the transformer architecture
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tieyan Liu · 2020
Later among the works it cites.
Are transformers universal approximators of sequence-to-sequence functions?
Chulhee Yun, Srinadh Bhojanapalli, Ankit Singh Rawat, Sashank Reddi, and Sanjiv Kumar · 2020
Later among the works it cites.
A virtual finite element method for two dimensional Maxwell interface problems with a background unfitted mesh
Shuhao Cao, Long Chen, and Ruchi Guo · 2021
Closest in time.
Rethinking attention with Performers
Krzysztof Marcin Choromanski, Valerii Likhosherstov, David Dohan, Xingyou Song, Andreea Gane, Tamas Sarlos, Peter Hawkins, Jared Quincy Davis, Afroz Mohiuddin, Lukasz Kaiser, David Benjamin Belanger, Lucy J Colwell, and Adrian Weller · 2021
Closest in time.
The devil is in the detail: Simple tricks improve systematic generalization of transformers
Róbert Csordás, Kazuki Irie, and Jürgen Schmidhuber · 2021
Closest in time.
Construct deep neural networks based on direct sampling methods for solving electrical impedance tomography
Ruchi Guo and Jiahua Jiang · 2021
Closest in time.
Multiwavelet-based operator learning for differential equations
Gaurav Gupta, Xiongye Xiao, and Paul Bogdan · 2021
Closest in time.
LieTransformer: equivariant self-attention for Lie groups
Michael J Hutchinson, Charline Le Lan, Sheheryar Zaidi, Emilien Dupont, Yee Whye Teh, and Hyunjik Kim · 2021
Closest in time.
draw.io, 2021
JGraph · 2021
Closest in time.
Jiahua Jiang, Yi Li, and Ruchi Guo · 2021
Closest in time.
Highly accurate protein structure prediction with AlphaFold
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Žídek, Anna Potapenko, et al · 2021
Closest in time.
Physics-informed machine learning
George Em Karniadakis, Ioannis G. Kevrekidis, Lu Lu, Paris Perdikaris, Sifan Wang, and Liu Yang · 2021
Closest in time.
FNet: Mixing tokens with Fourier transforms
James Lee-Thorp, Joshua Ainslie, Ilya Eckstein, and Santiago Ontanon · 2021
Closest in time.
Fourier neural operator for parametric partial differential equations
Zongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede liu, Kaushik Bhattacharya, Andrew Stuart, and Anima Anandkumar · 2021
Closest in time.
The random feature model for input-output maps between banach spaces
Nicholas H Nelsen and Andrew M Stuart · 2021
Closest in time.
FMMformer: Efficient and Flexible Transformer via Decomposed Near-field and Far-field Attention
Tan M. Nguyen, Vai Suliafu, Stanley J. Osher, Long Chen, and Bao Wang · 2021
Closest in time.
Random feature attention
Hao Peng, Nikolaos Pappas, Dani Yogatama, Roy Schwartz, Noah Smith, and Lingpeng Kong · 2021
Closest in time.
Rethinking neural operations for diverse tasks
Nicholas Roberts, Mikhail Khodak, Tri Dao, Liam Li, Christopher Ré, and Ameet Talwalkar · 2021
Closest in time.
Linear transformers are secretly fast weight programmers, 2021
Imanol Schlag, Kazuki Irie, and Jürgen Schmidhuber · 2021
Closest in time.
Efficient attention: Attention with linear complexities
Zhuoran Shen, Mingyuan Zhang, Haiyu Zhao, Shuai Yi, and Hongsheng Li · 2021
Closest in time.
Implicit kernel attention
Kyungwoo Song, Yohan Jung, Dongjun Kim, and Il-Chul Moon · 2021
Closest in time.
MLP-mixer: An all-MLP architecture for vision
Ilya Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer, Xiaohua Zhai, Thomas Unterthiner, Jessica Yung, Andreas Steiner, Daniel Keysers, Jakob Uszkoreit, et al · 2021
Closest in time.
seaborn: statistical data visualization
Michael L. Waskom · 2021
Closest in time.
Transformers are deep infinite-dimensional non-Mercer binary kernel machines
Matthew A Wright and Joseph E Gonzalez · 2021
Closest in time.
Nyströmformer: A Nyström-based algorithm for approximating self-attention
Yunyang Xiong, Zhanpeng Zeng, Rudrasis Chakraborty, Mingxing Tan, Glenn Fung, Yin Li, and Vikas Singh · 2021
Closest in time.