Fetching the paper…
Reading the bibliography…
Connections between Deep Neural Networks (DNNs) training and optimal control theory has attracted considerable attention as a principled tool of algorithmic design.
The theory of dynamic programming
Richard Bellman · 1954
Earlier work this paper cites.
Advantages of differential dynamic programming over newton’s method for discrete-time optimal control problems
Li-zhi Liao and Christine A Shoemaker · 1992
Earlier work this paper cites.
Respect the unstable
Gunter Stein · 2003
Earlier work this paper cites.
A generalized iterative lqg method for locally-optimal feedback control of constrained nonlinear stochastic systems
Emanuel Todorov and Weiwei Li · 2005
Earlier work this paper cites.
Cooperative stochastic differential games
David WK Yeung and Leon A Petrosjan · 2006
Earlier work this paper cites.
Deconvolutional networks
Matthew D Zeiler, Dilip Krishnan, Graham W Taylor, and Rob Fergus · 2010
Earlier work this paper cites.
Synthesis and stabilization of complex behaviors through online trajectory optimization
Yuval Tassa, Tom Erez, and Emanuel Todorov · 2012
Earlier work this paper cites.
Neural networks for machine learning lecture 6a overview of mini-batch gradient descent
Geoffrey Hinton, Nitish Srivastava, and Kevin Swersky · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Matthew D Zeiler and Rob Fergus · 2014
Earlier work this paper cites.
Optimizing neural networks with kronecker-factored approximate curvature
James Martens and Roger Grosse · 2015
Earlier work this paper cites.
Robust trajectory optimization: A cooperative stochastic game theoretic approach
Yunpeng Pan, Evangelos Theodorou, and Kaivalya Bakshi · 2015
Earlier work this paper cites.
Samuel S Schoenholz, Justin Gilmer, Surya Ganguli, and Jascha Sohl-Dickstein · 2016
Earlier work this paper cites.
Grammars for games: a gradient-based, game-theoretic framework for optimization in deep learning
David Balduzzi · 2016
Earlier work this paper cites.
A kronecker-factored approximate fisher matrix for convolution layers
Roger Grosse and James Martens · 2016
Earlier work this paper cites.
Differential dynamic programming for time-delayed systems
David D Fan and Evangelos A Theodorou · 2016
Earlier work this paper cites.
A guide to convolution arithmetic for deep learning
Vincent Dumoulin and Francesco Visin · 2016
Cited alongside, same era.
Yiping Lu, Aoxiao Zhong, Quanzheng Li, and Bin Dong · 2017
Cited alongside, same era.
Opening the black box of deep neural networks via information
Ravid Shwartz-Ziv and Naftali Tishby · 2017
Cited alongside, same era.
A proposal on machine learning via dynamical systems
E Weinan · 2017
Cited alongside, same era.
Stable architectures for deep neural networks
Eldad Haber and Lars Ruthotto · 2017
Cited alongside, same era.
Analysing neural network topologies: a game theoretic approach
Julian Stier, Gabriele Gianini, Michael Granitzer, and Konstantin Ziegler · 2018
Later among the works it cites.
Fast approximate natural gradient descent in a kronecker factored eigenbasis
Thomas George, César Laurent, Xavier Bouthillier, Nicolas Ballas, and Pascal Vincent · 2018
Later among the works it cites.
Ole: Orthogonal low-rank embedding-a plug and play geometric loss for deep learning
José Lezama, Qiang Qiu, Pablo Musé, and Guillermo Sapiro · 2018
Later among the works it cites.
Deep reinforcement learning that matters
Peter Henderson, Riashat Islam, Philip Bachman, Joelle Pineau, Doina Precup, and David Meger · 2018
Later among the works it cites.
Hamiltonian neural networks
Samuel Greydanus, Misko Dzamba, and Jason Yosinski · 2019
Later among the works it cites.
Symplectic ode-net: Learning hamiltonian dynamics with control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stochastic modified equations and adaptive stochastic gradient algorithms
Qianxiao Li, Cheng Tai, and Weinan E · 2017
Cited alongside, same era.
On convergence and stability of gans
Naveen Kodali, Jacob Abernethy, James Hays, and Zsolt Kira · 2017
Cited alongside, same era.
The numerics of gans
Lars Mescheder, Sebastian Nowozin, and Andreas Geiger · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Cited alongside, same era.
Stochastic training of residual networks: a differential equation viewpoint
Qi Sun, Yunzhe Tao, and Qiang Du · 2018
Cited alongside, same era.
Reversible architectures for arbitrarily deep residual neural networks
Bo Chang, Lili Meng, Eldad Haber, Lars Ruthotto, David Begert, and Elliot Holtham · 2018
Cited alongside, same era.
Neural ordinary differential equations
Tian Qi Chen, Yulia Rubanova, Jesse Bettencourt, and David K Duvenaud · 2018
Cited alongside, same era.
Yaofeng Desmond Zhong, Biswadip Dey, and Amit Chakraborty · 2019
Later among the works it cites.
Neural sde: Stabilizing neural ode networks with stochastic noise
Xuanqing Liu, Tesi Xiao, Si Si, Qin Cao, Sanjiv Kumar, and Cho-Jui Hsieh · 2019
Later among the works it cites.
Deep learning theory review: An optimal control and dynamical systems perspective
Guan-Horng Liu and Evangelos A Theodorou · 2019
Later among the works it cites.
You only propagate once: Accelerating adversarial training via maximal principle
Dinghuai Zhang, Tianyuan Zhang, Yiping Lu, Zhanxing Zhu, and Bin Dong · 2019
Later among the works it cites.
Cross-entropy loss and low-rank features have responsibility for adversarial examples
Kamil Nar, Orhan Ocal, S Shankar Sastry, and Kannan Ramchandran · 2019
Later among the works it cites.
Splitting steepest descent for growing neural architectures
Lemeng Wu, Dilin Wang, and Qiang Liu · 2019
Later among the works it cites.
Differential dynamic programming neural optimizer
Guan-Horng Liu, Tianrong Chen, and Evangelos A Theodorou · 2020
Closest in time.
Yiping Lu, Chao Ma, Yulong Lu, Jianfeng Lu, and Lexing Ying · 2020
Closest in time.
Layer-parallel training of deep residual neural networks
Stefanie Gunther, Lars Ruthotto, Jacob B Schroder, Eric C Cyr, and Nicolas R Gauger · 2020
Closest in time.
Neuron shapley: Discovering the responsible neurons
Amirata Ghorbani and James Zou · 2020
Closest in time.