Fetching the paper…
Reading the bibliography…
We present the 2017 WebVision Challenge, a public image recognition challenge designed for deep learning based on web images without instance-level human annotation.
G. Griffin, A. Holub, and P. Perona, “Caltech-256 object category dataset,” California Institute of Technology, Tech. Rep., 2007
2007
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, and L. Fei-Fei, “Imagenet: a large-scale hierarchical image database,” in CVPR , 2008
2008
Earlier work this paper cites.
S. Vijayanarasimhan and K. Grauman, “Keywords to visual categories: Multiple-instance learning for weakly supervised object categorization,” in CVPR , 2008
2008
Earlier work this paper cites.
A. Torralba, R. Fergus, and W. T. Freeman, “80 million tiny images: a large dataset for non-parametric object and scene recognition,” T-PAMI , vol. 30, no. 11, pp. 1958–1970, 2008
2008
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. Williams, J. Winn, and A. Zisserman, “The PASCAL visual object classes (VOC) challenge,” IJCV , vol. 88, no. 2, pp. 303–338, 2010
2010
Earlier work this paper cites.
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P. Manzagol, “Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion,” JMLR , vol. 11, pp. 3371––3408, 2010
2010
Earlier work this paper cites.
F. Schroff, A. Criminisi, and A. Zisserman, “Harvesting image databases from the web,” T-PAMI , vol. 33, no. 4, pp. 756–766, 2011
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in NIPS , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in NIPS , 2014, pp. 568–576
2014
Earlier work this paper cites.
W. Li, L. Niu, and D. Xu, “Exploiting privileged information from web data for image categorization,” in ECCV , 2014
2014
Earlier work this paper cites.
M. Oquab, L. Bottou, I. Laptev, and J. Sivic, “Learning and transferring mid-level image representations using convolutional neural networks,” in 1717–1724 , 2014, p. June
2014
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. E. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in CVPR , 2015, pp. 1–9
2015
Cited alongside, same era.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in CVPR , 2015, pp. 3431–3440
2015
Cited alongside, same era.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in NIPS , 2015, pp. 91–99
2015
Cited alongside, same era.
A. Dosovitskiy, P. Fischer, E. Ilg, P. Häusser, C. Hazirbas, V. Golkov, P. van der Smagt, D. Cremers, and T. Brox, “Flownet: Learning optical flow with convolutional networks,” in ICCV , 2015, pp. 2758–2766
2015
Cited alongside, same era.
D. Tran, L. D. Bourdev, R. Fergus, L. Torresani, and M. Paluri, “Learning spatiotemporal features with 3d convolutional networks,” in ICCV , 2015, pp. 4489–4497
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Later among the works it cites.
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Van Gool, “Temporal segment networks: Towards good practices for deep action recognition,” in ECCV , 2016, pp. 20–36
2016
Later among the works it cites.
C. Dong, C. C. Loy, K. He, and X. Tang, “Image super-resolution using deep convolutional networks,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 38, no. 2, pp. 295–307, 2016
2016
Later among the works it cites.
M. Norooz and P. Favaro, “Unsupervised learning of visual representations by solving jigsaw puzzles,” in ECCV , 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
A. Radford, L. Metz, and S. Chintalan, “Unsupervised representation learning with deep convolutional generative adversarial networks,” in arXiv:/1511.06434 , 2015
2015
Cited alongside, same era.
C. Doersch, A. Gupta, and A. A. Efros, “Unsupervised visual representation learning by context prediction,” in ICCV , 2015
2015
Cited alongside, same era.
X. Wang and A. Gupta, “Unsupervised learning of visual representations using videos,” in ICCV , 2015
2015
Cited alongside, same era.
P. Agrawal, J. Carreira, and J. Malik, “Learning to see by moving,” in ICCV , 2015
2015
Cited alongside, same era.
X. Chen and A. Gupta, “Webly supervised learning of convolutional networks,” in ICCV , 2015
2015
Cited alongside, same era.
2016
Later among the works it cites.
A. Joulin, L. van der Maaten, A. Jabri, and N. Vasilache, “Learning visual features from large weakly supervised data,” in ECCV , 2016
2016
Later among the works it cites.
A. Owens, J. Wu, J. Mcdermott, A. Torralba, and W. Freeman, “Unsupervised learning of visual representations by solving jigsaw puzzles,” in ECCV , 2016
2016
Later among the works it cites.
J. Krause, B. Sapp, A. Howard, H. Zhou, A. Toshev, T. Duerig, J. Philbin, and L. Fei-Fei, “The unreasonable effectiveness of noisy data for fine-grained recognition,” in ECCV , 2016
2016
Later among the works it cites.
R. B. Girshick, J. Donahue, T. Darrell, and J. Malik, “Region-based convolutional networks for accurate object detection and segmentation,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 38, no. 1, pp. 142–158, 2016
2016
Later among the works it cites.
E. Shelhamer, J. Long, and T. Darrell, “Fully convolutional networks for semantic segmentation,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 39, no. 4, pp. 640–651, 2017
2017
Closest in time.