Fetching the paper…
Reading the bibliography…
Implicit deep learning has received increasing attention recently due to the fact that it generalizes the recursive prediction rules of many commonly used neural network architectures.
Introductory functional analysis with applications , volume 1
Erwin Kreyszig · 1978
Earlier work this paper cites.
A learning rule for asynchronous perceptrons with feedback in a combinatorial environment
LB ALMEIDA · 1987
Earlier work this paper cites.
Generalization of back propagation to recurrent and higher order neural networks
Fernando Pineda · 1987
Earlier work this paper cites.
Elementary functional analysis , volume 253
Barbara MacCluer · 2008
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Andrew M Saxe, James L McClelland, and Surya Ganguli · 2013
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Hamed Karimi, Julie Nutini, and Mark Schmidt · 2016
Earlier work this paper cites.
Optnet: Differentiable optimization as a layer in neural networks
Brandon Amos and J Zico Kolter · 2017
Earlier work this paper cites.
Differentiable learning of submodular models
Josip Djolonga and Andreas Krause · 2017
Earlier work this paper cites.
A convergence analysis of gradient descent for deep linear neural networks
Sanjeev Arora, Nadav Cohen, Noah Golowich, and Wei Hu · 2018
Earlier work this paper cites.
Trellis networks for sequence modeling
Shaojie Bai, J Zico Kolter, and Vladlen Koltun · 2018
Earlier work this paper cites.
Neural ordinary differential equations
Ricky TQ Chen, Yulia Rubanova, Jesse Bettencourt, and David Duvenaud · 2018
Cited alongside, same era.
End-to-end differentiable physics for learning and control
Filipe de Avila Belbute-Peres, Kevin Smith, Kelsey Allen, Josh Tenenbaum, and J Zico Kolter · 2018
Cited alongside, same era.
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals, Jakob Uszkoreit, and Łukasz Kaiser · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Cited alongside, same era.
Learning overparameterized neural networks via stochastic gradient descent on structured data
Laurent El Ghaoui, Fangda Gu, Bertrand Travacca, Armin Askari, and Alicia Y Tsai · 2019
Later among the works it cites.
Deep declarative networks: A new hope
Stephen Gould, Richard Hartley, and Dylan Campbell · 2019
Later among the works it cites.
Ffjord: Free-form continuous dynamics for scalable reversible generative models
Will Grathwohl, Ricky T. Q. Chen, Jesse Bettencourt, Ilya Sutskever, and David Duvenaud · 2019
Later among the works it cites.
Satnet: Bridging deep learning and logical reasoning using a differentiable satisfiability solver
Po-Wei Wang, Priya Donti, Bryan Wilder, and Zico Kolter · 2019
Later among the works it cites.
An improved analysis of training over-parameterized deep neural networks
Difan Zou and Quanquan Gu · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuanzhi Li and Yingyu Liang · 2018
Cited alongside, same era.
Reviving and improving recurrent back-propagation
Renjie Liao, Yuwen Xiong, Ethan Fetaya, Lisa Zhang, KiJung Yoon, Xaq Pitkow, Raquel Urtasun, and Richard Zemel · 2018
Cited alongside, same era.
High-Dimensional Probability: An Introduction with Applications in Data Science
Roman Vershynin · 2018
Cited alongside, same era.
A convergence theory for deep learning via over-parameterization
Zeyuan Allen-Zhu, Yuanzhi Li, and Zhao Song · 2019
Cited alongside, same era.
On exact computation with an infinitely wide neural net
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, Ruslan Salakhutdinov, and Ruosong Wang · 2019
Cited alongside, same era.
Shaojie Bai, J Zico Kolter, and Vladlen Koltun · 2019
Cited alongside, same era.
Recurrent stacking of layers for compact neural machine translation models
Raj Dabre and Atsushi Fujita · 2019
Cited alongside, same era.
Gradient descent finds global minima of deep neural networks
Simon Du, Jason Lee, Haochuan Li, Liwei Wang, and Xiyu Zhai · 2019
Cited alongside, same era.
Later among the works it cites.
Multiscale deep equilibrium models
Shaojie Bai, Vladlen Koltun, and J Zico Kolter · 2020
Later among the works it cites.
Global convergence of deep networks with one wide layer followed by pyramidal topology
Quynh Nguyen and Marco Mondelli · 2020
Later among the works it cites.
Toward moderate overparameterization: Global convergence guarantees for training shallow neural networks
Samet Oymak and Mahdi Soltanolkotabi · 2020
Later among the works it cites.
Scalable differentiable physics for learning and control
Yi-Ling Qiao, Junbang Liang, Vladlen Koltun, and Ming C Lin · 2020
Later among the works it cites.
Gradient descent optimizes over-parameterized deep relu networks
Difan Zou, Yuan Cao, Dongruo Zhou, and Quanquan Gu · 2020
Later among the works it cites.
On the theory of implicit deep learning: Global convergence with implicit layers
Kenji Kawaguchi · 2021
Closest in time.
On the proof of global convergence of gradient descent for deep relu networks with linear widths
Quynh Nguyen · 2021
Closest in time.