Fetching the paper…
Reading the bibliography…
In this paper, we present a study on learning visual recognition models from large scale noisy web data.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Learning object categories from google’s image search
R. Fergus, L. Fei-Fei, P. Perona, and A. Zisserman · 2005
Earlier work this paper cites.
Caltech-256 object category dataset
G. Griffin, A. Holub, and P. Perona · 2007
Earlier work this paper cites.
Imagenet: a large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, and L. Fei-Fei · 2008
Earlier work this paper cites.
80 million tiny images: a large dataset for non-parametric object and scene recognition
A. Torralba, R. Fergus, and W. T. Freeman · 2008
Earlier work this paper cites.
Keywords to visual categories: Multiple-instance learning for weakly supervised object categorization
S. Vijayanarasimhan and K. Grauman · 2008
Earlier work this paper cites.
Exploiting weakly-labeled web images to improve object classification: a domain adaptation approach
A. Bergamo and L. Torresani · 2010
Earlier work this paper cites.
The PASCAL visual object classes (VOC) challenge
M. Everingham, L. Van Gool, C. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P. Manzagol · 2010
Earlier work this paper cites.
Sun database: Large-scale scene recognition from abbey to zoo
J. Xiao, J. Hays, K. Ehinger, A. Oliva, and A. Torralba · 2010
Earlier work this paper cites.
Domain adaptation for object recognition: An unsupervised approach
R. Gopalan, R. Li, and R. Chellappa · 2011
Earlier work this paper cites.
Hmdb: a large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
What you saw is not what you get: Domain adaptation using asymmetric kernel transforms
B. Kulis, K. Saenko, and T. Darrell · 2011
Earlier work this paper cites.
Harvesting image databases from the web
F. Schroff, A. Criminisi, and A. Zisserman · 2011
Earlier work this paper cites.
Unbiased look at dataset bias
A. Torralba and A. A. Efros · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Cited alongside, same era.
Exploiting privileged information from web data for image categorization
W. Li, L. Niu, and D. Xu · 2014
Cited alongside, same era.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollar, and C. L. Zitnick · 2014
Cited alongside, same era.
Learning and transferring mid-level image representations using convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2014
Cited alongside, same era.
Two-stream convolutional networks for action recognition in videos
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. E. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Later among the works it cites.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. D. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Later among the works it cites.
Unsupervised learning of visual representations using videos
X. Wang and A. Gupta · 2015
Later among the works it cites.
Image super-resolution using deep convolutional networks
C. Dong, C. C. Loy, K. He, and X. Tang · 2016
Later among the works it cites.
Region-based convolutional networks for accurate object detection and segmentation
R. B. Girshick, J. Donahue, T. Darrell, and J. Malik · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Learning to see by moving
P. Agrawal, J. Carreira, and J. Malik · 2015
Cited alongside, same era.
Webly supervised learning of convolutional networks
X. Chen and A. Gupta · 2015
Cited alongside, same era.
Unsupervised visual representation learning by context prediction
C. Doersch, A. Gupta, and A. A. Efros · 2015
Cited alongside, same era.
Flownet: Learning optical flow with convolutional networks
A. Dosovitskiy, P. Fischer, E. Ilg, P. Häusser, C. Hazirbas, V. Golkov, P. van der Smagt, D. Cremers, and T. Brox · 2015
Cited alongside, same era.
Activitynet: A large-scale video benchmark for human activity understanding
B. G. Fabian Caba Heilbron, Victor Escorcia and J. C. Niebles · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Learning visual features from large weakly supervised data
A. Joulin, L. van der Maaten, A. Jabri, and N. Vasilache · 2016
Later among the works it cites.
Openimages: A public dataset for large-scale multi-label and multi-class image classification
I. Krasin, T. Duerig, N. Alldrin, A. Veit, S. Abu-El-Haija, S. Belongie, D. Cai, Z. Feng, V. Ferrari, V. Gomes, A. Gupta, D. Narayanan, C. Sun, G. Chechik, and K. Murphy · 2016
Later among the works it cites.
The unreasonable effectiveness of noisy data for fine-grained recognition
J. Krause, B. Sapp, A. Howard, H. Zhou, A. Toshev, T. Duerig, J. Philbin, and L. Fei-Fei · 2016
Later among the works it cites.
Unsupervised learning of visual representations by solving jigsaw puzzles
M. Norooz and P. Favaro · 2016
Later among the works it cites.
Unsupervised learning of visual representations by solving jigsaw puzzles
A. Owens, J. Wu, J. Mcdermott, A. Torralba, and W. Freeman · 2016
Later among the works it cites.
Temporal segment networks: Towards good practices for deep action recognition
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Van Gool · 2016
Later among the works it cites.
Places: An image database for deep scene understanding
B. Zhou, A. Khosla, À. Lapedriza, A. Torralba, and A. Oliva · 2016
Later among the works it cites.
Fully convolutional networks for semantic segmentation
E. Shelhamer, J. Long, and T. Darrell · 2017
Closest in time.