Fetching the paper…
Reading the bibliography…
Generalization Performance of Deep Learning models trained using Empirical Risk Minimization can be improved significantly by using Data Augmentation strategies such as simple transformations, or using Mixed Samples.
Dropout: A simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R · 1958
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G. E., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R · 1958
Earlier work this paper cites.
Transformation invariance in pattern recognition-tangent distance and tangent propagation
Simard, P. Y., LeCun, Y., Denker, J. S., and Victorri, B · 1996
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. B. and Haffner, P · 1998
Earlier work this paper cites.
The vicinal risk minimization principle and the svms
Vapnik, V. N · 2000
Earlier work this paper cites.
Model compression
Buciluǎ, C., Caruana, R., and Niculescu-Mizil, A · 2006
Earlier work this paper cites.
Visualizing data using t-sne
Maaten, L. v. d. and Hinton, G · 2008
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Weight uncertainty in neural networks, 2015
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Earlier work this paper cites.
Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding, 2015
Han, S., Mao, H., and Dally, W. J · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network, 2015
Hinton, G., Vinyals, O., and Dean, J · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift, 2015
Ioffe, S. and Szegedy, C · 2015
Earlier work this paper cites.
Deep residual learning for image recognition, 2015
Kaiming He, Xiangyu Zhang, S. R. J. S · 2015
Earlier work this paper cites.
Improved regularization of convolutional neural networks with cutout, 2017
DeVries, T. and Taylor, G. W · 2017
Earlier work this paper cites.
On calibration of modern neural networks, 2017
Guo, C., Pleiss, G., Sun, Y., and Weinberger, K. Q · 2017
Cited alongside, same era.
Learning from noisy labels with distillation
Li, Y., Yang, J., Song, Y., Cao, L., Luo, J., and Li, L.-J · 2017
Cited alongside, same era.
Implicit regularization in deep learning
Neyshabur, B · 2017
Cited alongside, same era.
Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results, 2017
Tarvainen, A. and Valpola, H · 2017
Cited alongside, same era.
mixup: Beyond empirical risk minimization, 2017
Zhang, H., Cisse, M., Dauphin, Y. N., and Lopez-Paz, D · 2017
Cited alongside, same era.
On the convergence of mirror descent beyond stochastic convex programming, 2017
A group-theoretic framework for data augmentation, 2019
Chen, S., Dobriban, E., and Lee, J. H · 2019
Later among the works it cites.
On the efficacy of knowledge distillation
Cho, J. H. and Hariharan, B · 2019
Later among the works it cites.
Data augmentation revisited: Rethinking the distribution gap between clean and augmented data, 2019
He, Z., Xie, L., Chen, X., Zhang, Y., Wang, Y., and Tian, Q · 2019
Later among the works it cites.
Knowledge distillation with adversarial samples supporting decision boundary
Heo, B., Lee, M., Yun, S., and Choi, J. Y · 2019
Later among the works it cites.
Data augmentation instead of explicit regularization, 2019
Hernandez-Garcia, A. and Konig, P · 2019
Later among the works it cites.
When does label smoothing help?, 2019
Müller, R., Kornblith, S., and Hinton, G · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhou, Z., Mertikopoulos, P., Bambos, N., Boyd, S., and Glynn, P · 2017
Cited alongside, same era.
Cinic-10 is not imagenet or cifar-10, 2018
Darlow, L. N., Crowley, E. J., Antoniou, A., and Storkey, A. J · 2018
Cited alongside, same era.
Born again neural networks, 2018
Furlanello, T., Lipton, Z. C., Tschannen, M., Itti, L., and Anandkumar, A · 2018
Cited alongside, same era.
Dropblock: A regularization method for convolutional networks, 2018
Golnaz Ghiasi, Tsung-Yi Lin, Q. V. L · 2018
Cited alongside, same era.
The lottery ticket hypothesis: Finding sparse, trainable neural networks, 2018
Jonathan Frankle, M. C · 2018
Cited alongside, same era.
Robustness may be at odds with accuracy, 2018
Tsipras, D., Santurkar, S., Engstrom, L., Turner, A., and Madry, A · 2018
Cited alongside, same era.
Between-class learning for image classification, 2018
Yuji Tokozume, Yoshitaka Ushiku, T. H · 2018
Cited alongside, same era.
Human uncertainty makes classification more robust
Peterson, J., Battleday, R., Griffiths, T., and Russakovsky, O · 2019
Later among the works it cites.
Towards understanding knowledge distillation
Phuong, M. and Lampert, C · 2019
Later among the works it cites.
On mixup training: Improved calibration and predictive uncertainty for deep neural networks, 2019
Thulasidasan, S., Chennupati, G., Bilmes, J., Bhattacharya, T., and Michalak, S · 2019
Later among the works it cites.
Self-training with noisy student improves imagenet classification, 2019
Xie, Q., Luong, M.-T., Hovy, E., and Le, Q. V · 2019
Later among the works it cites.
Cutmix: Regularization strategy to train strong classifiers with localizable features
Yun, S., Han, D., Chun, S., Oh, S. J., Yoo, Y., and Choe, J · 2019
Later among the works it cites.
Affinity and diversity: Quantifying mechanisms of data augmentation, 2020
Gontijo-Lopes, R., Smullin, S. J., Cubuk, E. D., and Dyer, E · 2020
Closest in time.
Understanding and enhancing mixed sample data augmentation, 2020
Harris, E., Marcu, A., Painter, M., Niranjan, M., Prügel-Bennett, A., and Hare, J · 2020
Closest in time.