Fetching the paper…
Reading the bibliography…
We propose `Hide-and-Seek', a weakly-supervised framework that aims to improve object localization in images and action localization in videos.
Unsupervised Learning of Models for Recognition
M. Weber, M. Welling, and P. Perona · 2000
Earlier work this paper cites.
Object Class Recognition by Unsupervised Scale-Invariant Learning
R. Fergus, P. Perona, and A. Zisserman · 2003
Earlier work this paper cites.
Weakly supervised learning of part-based spatial models for visual object recognition
D. J. Crandall and D. P. Huttenlocher · 2006
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszalek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
Automatic annotation of human actions in video
O. Duchenne, I. Laptev, J. Sivic, F. Bach, and J. Ponce · 2009
Earlier work this paper cites.
Automatic attribute discovery and characterization from noisy web data
T. Berg, A. Berg, and J. Shih · 2010
Earlier work this paper cites.
Efficient activity detection with max-subgraph search
C. Y. Chen and K. Grauman · 2012
Earlier work this paper cites.
Discovering localized attributes for fine-grained recognition
K. Duan, D. Parikh, D. Crandall, and K. Grauman · 2012
Earlier work this paper cites.
Imagenet Classification with Deep Convolutional Neural Networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Earlier work this paper cites.
Learning Object Class Detectors from Weakly Annotated Video
A. Prest, C. Leistner, J. Civera, C. Schmid, and V. Ferrari · 2012
Earlier work this paper cites.
In Defence of Negative Mining for Annotating Weakly Labelled Data
P. Siva, C. Russell, and T. Xiang · 2012
Earlier work this paper cites.
Towards understanding action recognition
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black · 2013
Earlier work this paper cites.
Regularization of neural network using dropconnect
L. Wan, M. Zeiler, S. Zhang, Y. LeCun, and R. Fergus · 2013
Earlier work this paper cites.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
Earlier work this paper cites.
Weakly supervised learning for attribute localization in outdoor scenes
S. Wang, J. Joo, Y. Wang, and S. C. Zhu · 2013
Earlier work this paper cites.
Weakly supervised object detection with posterior regularization
H. Bilen, M. Pedersoli, and T. Tuytelaars · 2014
Earlier work this paper cites.
Weakly supervised action labeling in videos under ordering constraints
P. Bojanowski, R. Lajugie, F. Bach, I. Laptev, J. Ponce, C. Schmid, and J. Sivic · 2014
Earlier work this paper cites.
Multi-fold MIL Training for Weakly Supervised Object Localization
R. Cinbis, J. Verbeek, and C. Schmid · 2014
Earlier work this paper cites.
Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Simultaneous detection and segmentation
B. Hariharan, P. Arbeláez, R. Girshick, and J. Malik · 2014
Earlier work this paper cites.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Earlier work this paper cites.
Efficient feature extraction, encoding and classification for action recognition
V. Kantorov and I. Laptev · 2014
Cited alongside, same era.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Cited alongside, same era.
Hipster wars: Discovering elements of fashion styles
M. Kiapour, K. Yamaguchi, A. C. Berg, and T. L. Berg · 2014
Cited alongside, same era.
Deep inside convolutional networks: Visualising image classification models and saliency maps
K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Cited alongside, same era.
On Learning to Localize Objects with Minimal Supervision
H. O. Song, R. Girshick, S. Jegelka, J. Mairal, Z. Harchaoui, and T. Darrell · 2014
Cited alongside, same era.
Weakly-supervised discovery of visual pattern configurations
H. O. Song, Y. J. Lee, S. Jegelka, and T. Darrell · 2014
Temporal localization of fine-grained actions in videos by domain transfer from web images
C. Sun, S. Shetty, R. Sukthankar, and R. Nevatia · 2015
Later among the works it cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Later among the works it cites.
Efficient object localization using convolutional networks
J. Tompson, R. Goroshin, A. Jain, Y. LeCun, and C. Bregler · 2015
Later among the works it cites.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Later among the works it cites.
Discovering the spatial extent of relative attributes
F. Xiao and Y. J. Lee · 2015
Later among the works it cites.
Self-taught object localization with deep networks
L. Bazzani, B. A., D. Anguelov, and L. Torresani · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Weakly supervised object localization with latent category learning
C. Wang, W. Ren, K. Huang, and T. Tan · 2014
Cited alongside, same era.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Cited alongside, same era.
PANDA: Pose Aligned Networks for Deep Attribute Modeling
N. Zhang, M. Paluri, M. Ranzato, T. Darrell, and L. Bourdev · 2014
Cited alongside, same era.
Weakly supervised object localization with multi-fold multiple instance learning
R. Cinbis, J. Verbeek, and C. Schmid · 2015
Cited alongside, same era.
Convolutional feature masking for joint object and stuff segmentation
J. Dai, K. He, and J. Sun · 2015
Cited alongside, same era.
Later among the works it cites.
Weakly supervised deep detection networks
H. Bilen and A. Vedaldi · 2016
Later among the works it cites.
Connectionist temporal modeling for weakly supervised action labeling
D.-A. Huang, L. Fei-Fei, and J. C. Niebles · 2016
Later among the works it cites.
Contextlocnet: Context-aware deep network models for weakly supervised localization
V. Kantorov, M. Oquab, M. Cho, and I. Laptev · 2016
Later among the works it cites.
Weakly supervised object boundaries
A. Khoreva, R. Benenson, M. Omran, M. Hein, and B. Schiele · 2016
Later among the works it cites.
Ssd: Single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, and A. C. Berg · 2016
Later among the works it cites.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krähenbühl, J. Donahue, T. Darrell, and A. Efros · 2016
Later among the works it cites.
Temporal action localization in untrimmed videos via multi-stage cnns
Z. Shou, D. Wang, and S.-F. Chang · 2016
Later among the works it cites.
End-to-end localization and ranking for relative attributes
K. K. Singh and Y. J. Lee · 2016
Later among the works it cites.
Track and transfer: Watching videos to simulate strong human supervision for weakly-supervised object detection
K. K. Singh, F. Xiao, and Y. J. Lee · 2016
Later among the works it cites.
Walk and learn: Facial attribute representation learning from egocentric video and contextual data
J. Wang, Y. Cheng, and R. Schmidt Feris · 2016
Later among the works it cites.
End-to-end learning of action detection from frame glimpses in videos
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei · 2016
Later among the works it cites.
Learning deep features for discriminative localization
B. Zhou, A. Khosla, L. A., A. Oliva, and A. Torralba · 2016
Later among the works it cites.
A-fast-rcnn: Hard positive generation via adversary for object detection
X. Wang, A. Shrivastava, and A. Gupta · 2017
Closest in time.
Object region mining with adversarial erasing: A simple classification to semantic segmentation approach
Y. Wei, J. Feng, X. Liang, M.-M. Cheng, Y. Zhao, and S. Yan · 2017
Closest in time.