Fetching the paper…
Reading the bibliography…
Convolutional neural networks trained without supervision come close to matching performance with supervised pre-training, but sometimes at the cost of an even higher number of parameters.
Y. LeCun, J. S. Denker, and S. A. Solla, “Optimal brain damage,” in Advances in Neural Information Processing Systems (NIPS) , 1990
1990
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR) , 2009
2009
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” Tech. Rep., 2009
2009
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (voc) challenge,” International Journal of Computer Vision (IJCV) , 2010
2010
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in Neural Information Processing Systems (NIPS) , 2012
2012
Earlier work this paper cites.
D.-H. Lee, “Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks,” in Workshop on Challenges in Representation Learning, (ICML) , 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva, “Learning deep features for scene recognition using places database,” in Advances in Neural Information Processing Systems (NIPS) , 2014
2014
Earlier work this paper cites.
S. Han, J. Pool, J. Tran, and W. Dally, “Learning both weights and connections for efficient neural network,” in Advances in Neural Information Processing Systems (NIPS) , 2015
2015
Earlier work this paper cites.
W. Chen, J. Wilson, S. Tyree, K. Weinberger, and Y. Chen, “Compressing neural networks with the hashing trick,” in Proceedings of the International Conference on Machine Learning (ICML) , 2015
2015
Earlier work this paper cites.
C. Doersch, A. Gupta, and A. A. Efros, “Unsupervised visual representation learning by context prediction,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
X. Wang and A. Gupta, “Unsupervised learning of visual representations using videos,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Earlier work this paper cites.
M. Noroozi and P. Favaro, “Unsupervised learning of visual representations by solving jigsaw puzzles,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2016
2016
Earlier work this paper cites.
Y. Guo, A. Yao, and Y. Chen, “Dynamic network surgery for efficient dnns,” in Advances in Neural Information Processing Systems (NIPS) , 2016
2016
Earlier work this paper cites.
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi, “Xnor-net: Imagenet classification using binary convolutional neural networks,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2016
2016
Cited alongside, same era.
R. Zhang, P. Isola, and A. A. Efros, “Colorful image colorization,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2016
2016
Cited alongside, same era.
A. Dosovitskiy, P. Fischer, J. T. Springenberg, M. Riedmiller, and T. Brox, “Discriminative unsupervised feature learning with exemplar convolutional neural networks,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
T. Miyato, S.-i. Maeda, M. Koyama, and S. Ishii, “Virtual adversarial training: a regularization method for supervised and semi-supervised learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , 2018
2018
Later among the works it cites.
M. Caron, P. Bojanowski, J. Mairal, and A. Joulin, “Unsupervised pre-training of image features on non-curated data,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
P. Goyal, D. Mahajan, A. Gupta, and I. Misra, “Scaling and benchmarking self-supervised visual representation learning,” Proceedings of the International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Li, A. Kadav, I. Durdanovic, H. Samet, and H. P. Graf, “Pruning filters for efficient convnets,” in International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
X. Dong, S. Chen, and S. Pan, “Learning to prune deep neural networks via layer-wise optimal brain surgeon,” in Advances in Neural Information Processing Systems (NIPS) , 2017
2017
Cited alongside, same era.
K. Ullrich, E. Meeds, and M. Welling, “Soft weight-sharing for neural network compression,” in International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
D. Pathak, R. Girshick, P. Dollár, T. Darrell, and B. Hariharan, “Learning features by watching objects move,” in Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR) , 2017
2017
Cited alongside, same era.
P. Bojanowski and A. Joulin, “Unsupervised learning by predicting noise,” in Proceedings of the International Conference on Machine Learning (ICML) , 2017
2017
Cited alongside, same era.
C. Doersch and A. Zisserman, “Multi-task self-supervised visual learning,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2017
2017
Cited alongside, same era.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” 2017
2017
Cited alongside, same era.
C. Louizos, M. Welling, and D. P. Kingma, “Learning sparse neural networks through l _ 0 l\_0 regularization,” in International Conference on Learning Representations (ICLR) , 2018
2018
Cited alongside, same era.
2019
Later among the works it cites.
J. Frankle and M. Carbin, “The lottery ticket hypothesis: Finding sparse, trainable neural networks,” in International Conference on Learning Representations (ICLR) , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
Z. Liu, M. Sun, T. Zhou, G. Huang, and T. Darrell, “Rethinking the value of network pruning,” in International Conference on Learning Representations (ICLR) , 2019
2019
Later among the works it cites.
A. S. Morcos, H. Yu, M. Paganini, and Y. Tian, “One ticket to win them all: generalizing lottery ticket initializations across datasets and optimizers,” Advances in Neural Information Processing Systems (NeurIPS) , 2019
2019
Later among the works it cites.
A. Prakash, J. Storer, D. Florencio, and C. Zhang, “Repr: Improved training of convolutional filters,” in Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Later among the works it cites.
H. Zhou, J. Lan, R. Liu, and J. Yosinski, “Deconstructing lottery tickets: Zeros, signs, and the supermask,” in ”Workshop on Identifying and Understanding Deep Learning Phenomena (ICML)” , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
L. Jing and Y. Tian, “Self-supervised visual feature learning with deep neural networks: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , 2019
2019
Later among the works it cites.
X. Zhai, A. Oliver, A. Kolesnikov, and L. Beyer, “S4l: Self-supervised semi-supervised learning,” Proceedings of the International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
H. Yu, S. Edunov, Y. Tian, and A. S. Morcos, “Playing the lottery with rewards and multiple languages: lottery tickets in RL and NLP,” International Conference on Learning Representations (ICLR) , 2020
2020
Closest in time.