Fetching the paper…
Reading the bibliography…
We propose a new class of implicit networks, the multiscale deep equilibrium model (MDEQ), suited to large-scale and highly hierarchical pattern recognition domains.
Adjustment of an inverse matrix corresponding to a change in one element of a given matrix
J. Sherman and W. J. Morrison · 1950
Earlier work this paper cites.
A class of methods for solving nonlinear simultaneous equations
C. G. Broyden · 1965
Earlier work this paper cites.
The Laplacian pyramid as a compact image code
P. Burt and E. Adelson · 1983
Earlier work this paper cites.
Generalization of back propagation to recurrent and higher order neural networks
F. J. Pineda · 1988
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel · 1989
Earlier work this paper cites.
On the limited memory BFGS method for large scale optimization
D. C. Liu and J. Nocedal · 1989
Earlier work this paper cites.
A learning rule for asynchronous perceptrons with feedback in a combinatorial environment
L. B. Almeida · 1990
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and F. Li · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Earlier work this paper cites.
The implicit function theorem: History, theory, and applications
S. G. Krantz and H. R. Parks · 2012
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Learning hierarchical features for scene labeling
C. Farabet, C. Couprie, L. Najman, and Y. LeCun · 2013
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Earlier work this paper cites.
Deeply-supervised nets
C.-Y. Lee, S. Xie, P. Gallagher, Z. Zhang, and Z. Tu · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Earlier work this paper cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. E. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Earlier work this paper cites.
TensorFlow: A system for large-scale machine learning
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al · 2016
Cited alongside, same era.
Attention to scale: Scale-aware semantic image segmentation
L.-C. Chen, Y. Yang, J. Wang, W. Xu, and A. L. Yuille · 2016
Cited alongside, same era.
The Cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Cited alongside, same era.
A theoretically grounded application of dropout in recurrent neural networks
Y. Gal and Z. Ghahramani · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Stacked hourglass networks for human pose estimation
A. Newell, K. Yang, and J. Deng · 2016
End-to-end differentiable physics for learning and control
F. de Avila Belbute-Peres, K. Smith, K. Allen, J. Tenenbaum, and J. Z. Kolter · 2018
Later among the works it cites.
Multi-scale dense networks for resource efficient image classification
G. Huang, D. Chen, T. Li, F. Wu, L. van der Maaten, and K. Q. Weinberger · 2018
Later among the works it cites.
Reviving and improving recurrent back-propagation
R. Liao, Y. Xiong, E. Fetaya, L. Zhang, K. Yoon, X. Pitkow, R. Urtasun, and R. Zemel · 2018
Later among the works it cites.
MobileNetV2: Inverted residuals and linear bottlenecks
M. Sandler, A. G. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen · 2018
Later among the works it cites.
FishNet: A versatile backbone for image, region, and pixel level prediction
S. Sun, J. Pang, J. Shi, S. Yi, and W. Ouyang · 2018
Later among the works it cites.
Group normalization
Y. Wu and K. He · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
T. Salimans and D. P. Kingma · 2016
Cited alongside, same era.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Cited alongside, same era.
S. Zagoruyko and N. Komodakis · 2016
Cited alongside, same era.
OptNet: Differentiable optimization as a layer in neural networks
B. Amos and J. Z. Kolter · 2017
Cited alongside, same era.
Rethinking atrous convolution for semantic image segmentation
L.-C. Chen, G. Papandreou, F. Schroff, and H. Adam · 2017
Cited alongside, same era.
Differentiable learning of submodular models
J. Djolonga and A. Krause · 2017
Cited alongside, same era.
Later among the works it cites.
PSANet: Point-wise spatial attention network for scene parsing
H. Zhao, Y. Zhang, S. Liu, J. Shi, C. Change Loy, D. Lin, and J. Jia · 2018
Later among the works it cites.
UNet++: A nested U-Net architecture for medical image segmentation
Z. Zhou, M. M. R. Siddiquee, N. Tajbakhsh, and J. Liang · 2018
Later among the works it cites.
Universal transformers
M. Dehghani, S. Gouws, O. Vinyals, J. Uszkoreit, and Ł. Kaiser · 2019
Later among the works it cites.
Augmented neural ODEs
E. Dupont, A. Doucet, and Y. W. Teh · 2019
Later among the works it cites.
L. El Ghaoui, F. Gu, B. Travacca, and A. Askari · 2019
Later among the works it cites.
Deep declarative networks: A new hope
S. Gould, R. Hartley, and D. Campbell · 2019
Later among the works it cites.
FFJORD: Free-form continuous dynamics for scalable reversible generative models
W. Grathwohl, R. T. Chen, J. Betterncourt, I. Sutskever, and D. Duvenaud · 2019
Later among the works it cites.
Structured knowledge distillation for dense prediction
Y. Liu, C. Shu, J. Wang, and C. Shen · 2019
Later among the works it cites.
Gated-scnn: Gated shape cnns for semantic segmentation
T. Takikawa, D. Acuna, V. Jampani, and S. Fidler · 2019
Later among the works it cites.
SATNet: Bridging deep learning and logical reasoning using a differentiable satisfiability solver
P.-W. Wang, P. Donti, B. Wilder, and Z. Kolter · 2019
Later among the works it cites.
Scalable differentiable physics for learning and control
Y.-L. Qiao, J. Liang, V. Koltun, and M. C. Lin · 2020
Closest in time.
Deep high-resolution representation learning for visual recognition
J. Wang, K. Sun, T. Cheng, B. Jiang, C. Deng, Y. Zhao, D. Liu, Y. Mu, M. Tan, X. Wang, W. Liu, and B. Xiao · 2020
Closest in time.
Self-training with noisy student improves ImageNet classification
Q. Xie, E. Hovy, M.-T. Luong, and Q. V. Le · 2020
Closest in time.