Fetching the paper…
Reading the bibliography…
MixUp is an effective data augmentation method to regularize deep neural networks via random linear interpolations between pairs of samples and their labels.
1908
Earlier work this paper cites.
Y. Bengio, S. Bengio, and J. Cloutier, Learning a synaptic learning rule . Université de Montréal, Département d’informatique et de recherche opérationnelle, 1990
1990
Earlier work this paper cites.
G. E. Hinton and D. van Camp, “Keeping the neural networks simple by minimizing the description length of the weights,” in Proceedings of the Sixth Annual ACM Conference on Computational Learning Theory, COLT 1993, Santa Cruz, CA, USA, July 26-28, 1993. , 1993, pp. 5–13
1993
Earlier work this paper cites.
V. R. de Sa, “Learning classification with unlabeled data,” in NeurIPS , 1994
1994
Earlier work this paper cites.
S. Thrun and L. Pratt, “Learning to learn: Introduction and overview,” in Learning to learn . Springer, 1998
1998
Earlier work this paper cites.
B. Colson, P. Marcotte, and G. Savard, “An overview of bilevel optimization,” Annals OR , vol. 153, no. 1, pp. 235–256, 2007
2007
Earlier work this paper cites.
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng, “Reading digits in natural images with unsupervised feature learning,” 2011
2011
Earlier work this paper cites.
J. Snoek, H. Larochelle, and R. P. Adams, “Practical bayesian optimization of machine learning algorithms,” in NeurIPS , 2012
2012
Earlier work this paper cites.
J. Snoek, H. Larochelle, and R. P. Adams, “Practical bayesian optimization of machine learning algorithms,” in Advances in Neural Information Processing Systems 25: 26th Annual Conference on Neural Information Processing Systems 2012. Proceedings of a meeting held December 3-6, 2012, Lake Tahoe, Nevada, United States. , 2012, pp. 2960–2968
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
D.-H. Lee, “Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks,” in Workshop on Challenges in Representation Learning, ICML , 2013
2013
Earlier work this paper cites.
F. Shen, C. Shen, A. van den Hengel, and Z. Tang, “Approximate least trimmed sum of squares fitting and applications in image analysis,” IEEE Trans. Image Processing , vol. 22, no. 5, pp. 1836–1847, 2013
2013
Earlier work this paper cites.
J. Bergstra, D. Yamins, and D. D. Cox, “Making a science of model search: Hyperparameter optimization in hundreds of dimensions for vision architectures,” in Proceedings of the 30th International Conference on Machine Learning, ICML 2013, Atlanta, GA, USA, 16-21 June 2013 , 2013, pp. 115–123
2013
Earlier work this paper cites.
C. Thornton, F. Hutter, H. H. Hoos, and K. Leyton-Brown, “Auto-weka: combined selection and hyperparameter optimization of classification algorithms,” in The 19th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, KDD 2013, Chicago, IL, USA, August 11-14, 2013 , 2013, pp. 847–855
2013
Earlier work this paper cites.
N. Srivastava, G. E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” Journal of Machine Learning Research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Earlier work this paper cites.
M. Oquab, L. Bottou, I. Laptev, and J. Sivic, “Is object localization for free? - weakly-supervised learning with convolutional neural networks,” in CVPR , 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7-9, 2015, Conference Track Proceedings , 2015
2015
Earlier work this paper cites.
F. Shen, C. Shen, W. Liu, and H. T. Shen, “Supervised discrete hashing,” in IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2015, Boston, MA, USA, June 7-12, 2015 , 2015, pp. 37–45
2015
Cited alongside, same era.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. S. Bernstein, A. C. Berg, and F. Li, “Imagenet large scale visual recognition challenge,” International Journal of Computer Vision , vol. 115, no. 3, pp. 211–252, 2015
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Cited alongside, same era.
F. Pedregosa, “Hyperparameter optimization with approximate gradient,” in Proceedings of the 33nd International Conference on Machine Learning, ICML 2016, New York City, NY, USA, June 19-24, 2016 , 2016, pp. 737–746
2016
Cited alongside, same era.
K. Saito, Y. Ushiku, and T. Harada, “Asymmetric tri-training for unsupervised domain adaptation,” in Proceedings of the 34th International Conference on Machine Learning, ICML 2017, Sydney, NSW, Australia, 6-11 August 2017 , 2017, pp. 2988–2997
2017
Later among the works it cites.
M. Cisse, P. Bojanowski, E. Grave, Y. Dauphin, and N. Usunier, “Parseval networks: Improving robustness to adversarial examples,” in ICML . JMLR. org, 2017
2017
Later among the works it cites.
I. Loshchilov and F. Hutter, “SGDR: stochastic gradient descent with warm restarts,” in ICLR , 2017
2017
Later among the works it cites.
T. Miyato, S.-i. Maeda, S. Ishii, and M. Koyama, “Virtual adversarial training: a regularization method for supervised and semi-supervised learning,” IEEE transactions on pattern analysis and machine intelligence , 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
J. Luketina, T. Raiko, M. Berglund, and K. Greff, “Scalable gradient-based tuning of continuous regularization hyperparameters,” in Proceedings of the 33nd International Conference on Machine Learning, ICML 2016, New York City, NY, USA, June 19-24, 2016 , 2016, pp. 2952–2960
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Identity mappings in deep residual networks,” in ECCV , 2016
2016
Cited alongside, same era.
S. Zagoruyko and N. Komodakis, “Wide residual networks,” in BMVC , 2016
2016
Cited alongside, same era.
T. Durand, T. Mordan, N. Thome, and M. Cord, “WILDCAT: weakly supervised learning of deep convnets for image classification, pointwise localization and segmentation,” in CVPR , 2017
2017
Cited alongside, same era.
A. Tarvainen and H. Valpola, “Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results,” in NeurIPS , 2017
2017
Cited alongside, same era.
X. Gastaldi, “Shake-shake regularization of 3-branch residual networks,” in 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Workshop Track Proceedings , 2017
2017
Cited alongside, same era.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” Commun. ACM , vol. 60, no. 6, pp. 84–90, 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Y. Tokozume, Y. Ushiku, and T. Harada, “Between-class learning for image classification,” in CVPR , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
N. Mishra, M. Rohaninejad, X. Chen, and P. Abbeel, “A simple neural attentive meta-learner,” in 6th International Conference on Learning Representations, ICLR 2018, Vancouver, BC, Canada, April 30 - May 3, 2018, Conference Track Proceedings , 2018
2018
Later among the works it cites.
T. Munkhdalai, X. Yuan, S. Mehri, and A. Trischler, “Rapid adaptation with conditionally shifted neurons,” in Proceedings of the 35th International Conference on Machine Learning, ICML 2018, Stockholmsmässan, Stockholm, Sweden, July 10-15, 2018 , 2018, pp. 3661–3670
2018
Later among the works it cites.
M. Ren, W. Zeng, B. Yang, and R. Urtasun, “Learning to reweight examples for robust deep learning,” in ICML , 2018
2018
Later among the works it cites.
S. Xie, Z. Zheng, L. Chen, and C. Chen, “Learning semantic representations for unsupervised domain adaptation,” in Proceedings of the 35th International Conference on Machine Learning, ICML 2018, Stockholmsmässan, Stockholm, Sweden, July 10-15, 2018 , 2018, pp. 5419–5428
2018
Later among the works it cites.
Y. Tsuzuku, I. Sato, and M. Sugiyama, “Lipschitz-margin training: Scalable certification of perturbation invariance for deep neural networks,” in NeurIPS , 2018
2018
Later among the works it cites.
C. Liu, B. Zoph, M. Neumann, J. Shlens, W. Hua, L. Li, L. Fei-Fei, A. L. Yuille, J. Huang, and K. Murphy, “Progressive neural architecture search,” in Computer Vision - ECCV 2018 - 15th European Conference, Munich, Germany, September 8-14, 2018, Proceedings, Part I , 2018, pp. 19–35
2018
Later among the works it cites.
A. Oliver, A. Odena, C. A. Raffel, E. D. Cubuk, and I. J. Goodfellow, “Realistic evaluation of deep semi-supervised learning algorithms,” in NeurIPS , 2018
2018
Later among the works it cites.