Fetching the paper…
Reading the bibliography…
We study the training process of Deep Neural Networks (DNNs) from the Fourier analysis perspective.
Machine learning from a continuous viewpoint
Weinan E, Chao Ma, and Lei Wu · 1912
Earlier work this paper cites.
The Fourier transform and its applications , volume 31999
Ronald Newbold Bracewell and Ronald N Bracewell · 1986
Earlier work this paper cites.
Convergence estimates for multigrid algorithms without regularity assumptions
James H Bramble, Joseph E Pasciak, Jun Ping Wang, and Jinchao Xu · 1991
Earlier work this paper cites.
Model reduction with memory and the machine learning of dynamical systems
E Weinan, Chao Ma, and Jianchun Wang · 1991
Earlier work this paper cites.
The mnist database of handwritten digits
Yann LeCun · 1998
Earlier work this paper cites.
Almost linear vc dimension bounds for piecewise polynomial networks
Peter L Bartlett, Vitaly Maiorov, and Ron Meir · 1999
Earlier work this paper cites.
Multigrid
Ulrich Trottenberg, Cornelius W Oosterlee, and Anton Schuller · 2000
Earlier work this paper cites.
Cifar-10 (canadian institute for advanced research)
Alex Krizhevsky, Vinod Nair, and Geoffrey Hinton · 2010
Earlier work this paper cites.
Partial differential equations
Lawrence C Evans · 2010
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Train faster, generalize better: Stability of stochastic gradient descent
Moritz Hardt, Benjamin Recht, and Yoram Singer · 2015
Earlier work this paper cites.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang · 2016
Cited alongside, same era.
An efficient multigrid strategy for large-scale molecular mechanics optimization
Jingrun Chen and Carlos J García-Cervera · 2017
Cited alongside, same era.
Deep learning-based numerical methods for high-dimensional parabolic partial differential equations and backward stochastic differential equations
Weinan E, Jiequn Han, and Arnulf Jentzen · 2017
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Cited alongside, same era.
Failures of gradient-based deep learning
Shai Shalev-Shwartz, Ohad Shamir, and Shaked Shammah · 2017
The implicit bias of gradient descent on separable data
Daniel Soudry, Elad Hoffer, Mor Shpigel Nacson, Suriya Gunasekar, and Nathan Srebro · 2018
Later among the works it cites.
Training behavior of deep neural network in frequency domain
Zhi-Qin J Xu, Yaoyu Zhang, and Yanyang Xiao · 2019
Closest in time.
On the spectral bias of deep neural networks
Nasim Rahaman, Devansh Arpit, Aristide Baratin, Felix Draxler, Min Lin, Fred A Hamprecht, Yoshua Bengio, and Aaron Courville · 2019
Closest in time.
Switchnet: a neural network model for forward and inverse scattering problems
Yuehaw Khoo and Lexing Ying · 2019
Closest in time.
Relu deep neural networks and linear finite elements
Juncai He, Lin Li, Jinchao Xu, and Chunyue Zheng · 2019
Closest in time.
A multiscale neural network based on hierarchical matrices
Yuwei Fan, Lin Lin, Lexing Ying, and Leonardo Zepeda-Núnez · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nati Srebro · 2017
Cited alongside, same era.
Towards understanding generalization of deep learning: Perspective of loss landscapes
Lei Wu, Zhanxing Zhu, and Weinan E · 2017
Cited alongside, same era.
A closer look at memorization in deep networks
Devansh Arpit, Stanislaw Jastrzbski, Nicolas Ballas, David Krueger, Emmanuel Bengio, Maxinder S Kanwal, Tegan Maharaj, Asja Fischer, Aaron Courville, Yoshua Bengio, et al · 2017
Cited alongside, same era.
Deep potential: A general representation of a many-body potential energy surface
Jiequn Han, Linfeng Zhang, Roberto Car, et al · 2018
Cited alongside, same era.
Are efficient deep representations learnable?
Maxwell Nye and Andrew Saxe · 2018
Cited alongside, same era.
The deep ritz method: A deep learning-based numerical algorithm for solving variational problems
Weinan E and Bing Yu · 2018
Cited alongside, same era.
A priori estimates of the population risk for two-layer neural networks
Weinan E, Chao Ma, and Lei Wu
Cited in the paper.
Closest in time.
Uniformly accurate machine learning-based hydrodynamic models for kinetic equations
Jiequn Han, Chao Ma, Zheng Ma, and E Weinan · 2019
Closest in time.
Data-driven, physics-based feature extraction from fluid flow fields using convolutional neural networks
Carlos Michelen Strofer, Jin-Long Wu, Heng Xiao, and Eric Paterson · 2019
Closest in time.
Theory of the frequency principle for general deep neural networks
Tao Luo, Zheng Ma, Zhi-Qin John Xu, and Yaoyu Zhang · 2019
Closest in time.
Frequency-aware reconstruction of fluid simulations with generative networks
Simon Biland, Vinicius C Azevedo, Byungsoo Kim, and Barbara Solenthaler · 2019
Closest in time.
A phase shift deep neural network for high frequency wave equations in inhomogeneous media
Wei Cai, Xiaoguang Li, and Lizuo Liu · 2019
Closest in time.