Fetching the paper…
Reading the bibliography…
Dropout as a common regularizer to prevent overfitting in deep neural networks has been less effective in convolutional layers than in fully connected layers.
S. Lloyd, “Least squares quantization in pcm,”
1982
Earlier work this paper cites.
P. J. Rousseeuw, “Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,”
1987
Earlier work this paper cites.
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel, “Backpropagation applied to handwritten zip code recognition,”
1989
Earlier work this paper cites.
A. Banerjee and R. N. Dave, “Validating clusters using the Hopkins statistic,” in
2004
Earlier work this paper cites.
L. Fei-Fei, R. Fergus, and P. Perona, “One-shot learning of object categories,”
2006
Earlier work this paper cites.
J. Fan and R. Li, “Statistical challenges with high dimensionality: Feature selection in knowledge discovery,” in
2006
Earlier work this paper cites.
T. E. Oliphant, “Python for scientific computing,”
2007
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” Master’s Thesis, University of Toronto, 2009
2009
Earlier work this paper cites.
B. Lake, R. Salakhutdinov, J. Gross, and J. Tenenbaum, “One shot learning of simple visual concepts,” in
2011
Earlier work this paper cites.
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng, “Reading digits in natural images with unsupervised feature learning,” in
2011
Earlier work this paper cites.
F. Pedregosa, G. Varoquaux, A. Gramfort
2011
Earlier work this paper cites.
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in
2012
Earlier work this paper cites.
K. P. Murphy,
2012
Earlier work this paper cites.
L. Wan, M. Zeiler, S. Zhang, Y. Le Cun, and R. Fergus, “Regularization of neural networks using dropconnect,” in
2013
Earlier work this paper cites.
J. Ba and B. Frey, “Adaptive dropout for training deep neural networks,”
2013
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: A simple way to prevent neural networks from overfitting,”
2014
Earlier work this paper cites.
J. Tompson, R. Goroshin, A. Jain, Y. LeCun, and C. Bregler, “Efficient object localization using convolutional networks,” in
2015
Cited alongside, same era.
1000 Genomes Project Consortium, “A global reference for human genetic variation,”
2015
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in
2015
Cited alongside, same era.
G. Huang, Y. Sun, Z. Liu, D. Sedra, and K. Q. Weinberger, “Deep networks with stochastic depth,” in
2016
Cited alongside, same era.
A. Veit, M. J. Wilber, and S. Belongie, “Residual networks behave like ensembles of relatively shallow networks,” 2016
2016
Cited alongside, same era.
J. Snell, K. Swersky, and R. Zemel, “Prototypical networks for few-shot learning,”
Z. Chen, J. Niu, X. Liu, and S. Tang, “Selectscale: Mining more patterns from images via selective and soft dropout,” in
2020
Closest in time.
B. Shi, D. Zhang, Q. Dai, Z. Zhu, Y. Mu, and J. Wang, “Informative dropout for robust representation learning: A shape-bias perspective,” in
2020
Closest in time.
A. Pal, C. Lane, R. Vidal, and B. D. Haeffele, “On the regularization properties of structured dropout,” in
2020
Closest in time.
Y. Zeng, T. Dai, and S.-T. Xia, “Corrdrop: Correlation based dropout for convolutional neural networks,” in
2020
Closest in time.
V. Dodballapur, R. Calisa, Y. Song, and W. Cai, “Automatic dropout for deep neural networks,” in
2020
Closest in time.
M. V. Ramos, “Dropblock,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
S. Ravi and H. Larochelle, “Optimization as a model for few-shot learning,” in
2017
Cited alongside, same era.
L. Kaiser, O. Nachum, A. Roy, and S. Bengio, “Learning to remember rare events,” in
2017
Cited alongside, same era.
G. Ghiasi, T.-Y. Lin, and Q. V. Le, “Dropblock: A regularization method for convolutional networks,” in
2018
Cited alongside, same era.
J. Cavazza, P. Morerio, B. Haeffele, C. Lane, V. Murino, and R. Vidal, “Dropout as a low-rank regularizer for matrix factorization,” in
2018
Cited alongside, same era.
P. Mianjy, R. Arora, and R. Vidal, “On the implicit bias of dropout,” in
2018
Cited alongside, same era.
A. Clapés, O. Bilici, D. Temirova, E. Avots, G. Anbarjafari, and S. Escalera, “From apparent to real age: gender, age, ethnic, makeup, and expression bias analysis in real age estimation,” in
2018
Cited alongside, same era.
2020
Closest in time.
S. Mostafa, D. Mondal, M. Beck, C. Bidinosti, C. Henry, and I. Stavness, “Visualizing feature maps for model selection in convolutional neural networks,” in
2021
Closest in time.
A. Zunino, S. A. Bargal, P. Morerio, J. Zhang, S. Sclaroff, and V. Murino, “Excitation dropout: Encouraging plasticity in deep neural networks,”
2021
Closest in time.
H. Pham and Q. V. Le, “Autodropout: Learning dropout patterns to regularize deep networks,” in
2021
Closest in time.
X. Fan, S. Zhang, K. Tanwisuth, X. Qian, and M. Zhou, “Contextual dropout: An efficient sample-dependent dropout module,” in
2021
Closest in time.
H. Lee, Y. Park, H. Seo, and M. Kang, “Self-knowledge distillation via dropout,”
2022
Closest in time.
B. Li, Y. Hu, X. Nie, C. Han, X. Jiang, T. Guo, and L. Liu, “Dropkey,”
2022
Closest in time.
Z. Liu, Z. Xu, J. Jin, Z. Shen, and T. Darrell, “Dropout reduces underfitting,”
2023
Closest in time.
Y. Liu, C. Matsoukas, F. Strand, H. Azizpour, and K. Smith, “Patchdropout: Economizing vision transformers using patch dropout,” in
2023
Closest in time.