Fetching the paper…
Reading the bibliography…
In recent years, implicit deep learning has emerged as a method to increase the effective depth of deep neural networks.
Adjustment of an inverse matrix corresponding to a change in one element of a given matrix
J. Sherman and W. J. Morrison · 1950
Earlier work this paper cites.
A Class of Methods for Solving Nonlinear Simultaneous Equations
C. G. Broyden · 1965
Earlier work this paper cites.
On the local and superlinear convergence of quasi-Newton methods
C. G. Broyden, J. E. jun. Dennis, and J. J. More · 1973
Earlier work this paper cites.
On the global convergence of Broyden’s method
J. More and J. Trangenstein · 1976
Earlier work this paper cites.
Parallel quasi-Newton methods for unconstrained optimization
R. H. Byrd, R. B. Schnabel, and G. A. Shultz · 1988
Earlier work this paper cites.
A tool for the analysis of quasi-Newton methods with application to unconstrained minimization
R. H. Byrd and J. Nocedal · 1989
Earlier work this paper cites.
On the limited memory BFGS method for large scale optimization
D. C. Liu and J. Nocedal · 1989
Earlier work this paper cites.
Convergence of quasi-Newton matrices generated by the symmetric rank one update
A. R. Conn, N. I. Gould, and P. L. Toint · 1991
Earlier work this paper cites.
A theoretical and experimental study of the symmetric rank-one update
H. Fayez Khalfan, R. H. Byrd, and R. B. Schnabel · 1993
Earlier work this paper cites.
NewsWeeder: Learning to Filter Netnews
K. Lang · 1995
Earlier work this paper cites.
Convergence of Broyden-Like Matrix
D. Li, J. Zeng, and S. Zhou · 1998
Earlier work this paper cites.
A derivative-free line search and global convergence of Broyden-like method for nonlinear equations
D. Li and M. Fukushima · 2000
Earlier work this paper cites.
Quasi-Newton Methods
J. Nocedal and S. Wright · 2006
Earlier work this paper cites.
Matplotlib: A 2D graphics environment
J. D. Hunter · 2007
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Learning Multiple Layers of Features from Tiny Images
A. Krizhevsky · 2009
Cited alongside, same era.
On the local convergence of adjoint Broyden methods
S. Schlenkrich, A. Griewank, and A. Walther · 2010
Cited alongside, same era.
Scikit-learn: Machine Learning in Python Gaël Varoquaux Bertrand Thirion Vincent Dubourg Alexandre Passos PEDREGOSA, VAROQUAUX, GRAMFORT ET AL. Matthieu Perrot
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Cited alongside, same era.
Random Search for Hyper-Parameter Optimization Yoshua Bengio
J. Bergstra and Y. Bengio · 2012
Cited alongside, same era.
The Implicit Function Theorem: History, Theory, and Applications
S. G. Krantz and H. R. Parks · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Deep equilibrium models
S. Bai, J. Z. Kolter, and V. Koltun · 2019
Later among the works it cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Köpf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Later among the works it cites.
Multiscale deep equilibrium models
S. Bai, V. Koltun, and J. Z. Kolter · 2020
Later among the works it cites.
Array programming with NumPy
C. R. Harris, K. J. Millman, S. J. van der Walt, R. Gommers, P. Virtanen, D. Cournapeau, E. Wieser, J. Taylor, S. Berg, N. J. Smith, R. Kern, M. Picus, S. Hoyer, M. H. van Kerkwijk, M. Brett, A. Haldane, J. F. del Río, M. Wiebe, P. Peterson, P. Gérard-Marchant, K. Sheppard, T. Reddy, W. Weckesser, H. Abbasi, C. Gohlke, and T. E. Oliphant · 2020
Later among the works it cites.
Optimizing Millions of Hyperparameters by Implicit Differentiation
J. Lorraine, P. Vicol, and D. Duvenaud · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. P. Kingma and J. L. Ba · 2015
Cited alongside, same era.
Training Deep Nets with Sublinear Memory Cost
T. Chen, B. Xu, C. Zhang, and C. Guestrin · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Hyperparameter optimization with approximate gradient
F. Pedregosa · 2016
Cited alongside, same era.
Benefits of depth in neural networks
M. Telgarsky · 2016
Cited alongside, same era.
OptNet: Differentiable Optimization as a Layer in Neural Networks
B. Amos and J. Zico Kolter · 2017
Cited alongside, same era.
UCI machine learning repository, 2017
D. Dua and C. Graff · 2017
Cited alongside, same era.
F. Mannel · 2020
Later among the works it cites.
SciPy 1.0: fundamental algorithms for scientific computing in Python
P. Virtanen, R. Gommers, T. E. Oliphant, M. Haberland, T. Reddy, D. Cournapeau, E. Burovski, P. Peterson, W. Weckesser, J. Bright, S. J. van der Walt, M. Brett, J. Wilson, K. J. Millman, N. Mayorov, A. R. Nelson, E. Jones, R. Kern, E. Larson, C. J. Carey, I. Polat, Y. Feng, E. W. Moore, J. VanderPlas, D. Laxalde, J. Perktold, R. Cimrman, I. Henriksen, E. A. Quintero, C. R. Harris, A. M. Archibald, A. H. Ribeiro, F. Pedregosa, P. van Mulbregt, A. Vijaykumar, A. P. Bardelli, A. Rothberg, A. Hilboll, A. Kloeckner, A. Scopatz, A. Lee, A. Rokem, C. N. Woods, C. Fulton, C. Masson, C. Häggström, C. Fitzgerald, D. A. Nicholson, D. R. Hagen, D. V. Pasechnik, E. Olivetti, E. Martin, E. Wieser, F. Silva, F. Lenders, F. Wilhelm, G. Young, G. A. Price, G. L. Ingold, G. E. Allen, G. R. Lee, H. Audren, I. Probst, J. P. Dietrich, J. Silterra, J. T. Webber, J. Slavič, J. Nothman, J. Buchner, J. Kulick, J. L. Schönberger, J. V. de Miranda Cardoso, J. Reimer, J. Harrington, J. L. C. Rodríguez, J. Nunez-Iglesias, J. Kuczynski, K. Tritz, M. Thoma, M. Newville, M. Kümmerer, M. Bolingbroke, M. Tartre, M. Pak, N. J. Smith, N. Nowaczyk, N. Shebanov, O. Pavlyk, P. A. Brodtkorb, P. Lee, R. T. McGibbon, R. Feldbauer, S. Lewis, S. Tygier, S. Sievert, S. Vigna, S. Peterson, S. More, T. Pudlik, T. Oshima, T. J. Pingel, T. P. Robitaille, T. Spura, T. R. Jones, T. Cera, T. Leslie, T. Zito, T. Krauss, U. Upadhyay, Y. O. Halchenko, and Y. Vázquez-Baeza · 2020
Later among the works it cites.
Second-Order Optimization for Non-Convex Machine Learning: An Empirical Study
P. Xu, F. Roosta-Khorasani, and M. W. Mahoney · 2020
Later among the works it cites.
https://www.csie.ntu.edu.tw/˜cjlin/libsvmtools/datasets/
Libsvm datasets · 2021
Closest in time.
Quasi-Newton Methods for Machine Learning: Forget the Past, Just Sample
A. S. Berahas, M. Jahani, P. Richtárik, and M. Takáč · 2021
Closest in time.
Fixed Point Networks: Implicit Depth Models with Jacobian-Free Backprop
S. W. Fung, H. Heaton, Q. Li, D. Mckenzie, S. Osher, and W. Yin · 2021
Closest in time.
garrettj403/SciencePlots, Feb. 2021
J. D. Garrett and H.-H. Peng · 2021
Closest in time.
Deep Equilibrium Architectures for Inverse Problems in Imaging
D. Gilton, G. Ongie, and R. Willett · 2021
Closest in time.
Feasibility-based Fixed Point Networks
H. Heaton, S. W. Fung, A. Gibali, and W. Yin · 2021
Closest in time.
Momentum Residual Neural Networks
M. E. Sander, P. Ablin, M. Blondel, and G. Peyré · 2021
Closest in time.