Fetching the paper…
Reading the bibliography…
In this paper, we investigate a weakly-supervised object detection framework.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Localizing objects while learning their appearance
T. Deselaers, B. Alexe, and V. Ferrari · 2010
Earlier work this paper cites.
The PASCAL Visual Object Classes (VOC) challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Torch7: A Matlab-like environment for machine learning
R. Collobert, K. Kavukcuoglu, and C. Farabet · 2011
Earlier work this paper cites.
Learning object class detectors from weakly annotated video
A. Prest, C. Leistner, J. Civera, C. Schmid, and V. Ferrari · 2012
Earlier work this paper cites.
In defence of negative mining for annotating weakly labelled data
P. Siva, C. Russell, and T. Xiang · 2012
Earlier work this paper cites.
UCF101: A dataset of 101 human action classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Earlier work this paper cites.
Fast object segmentation in unconstrained video
A. Papazoglou and V. Ferrari · 2013
Earlier work this paper cites.
Unsupervised object discovery and segmentation in videos
S. Schulter, C. Leistner, P. M. Roth, and H. Bischof · 2013
Earlier work this paper cites.
Selective search for object recognition
J. Uijlings, K. van de Sande, T. Gevers, and A. Smeulders · 2013
Earlier work this paper cites.
Weakly supervised detection with posterior regularization
H. Bilen, M. Pedersoli, and T. Tuytelaars · 2014
Earlier work this paper cites.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Efficient image and video co-localization with Frank-Wolfe algorithm
A. Joulin, K. Tang, and L. Fei-Fei · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
T. Lin, M. Maire, S. J. Belongie, L. D. Bourdev, R. B. Girshick, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
On learning to localize objects with minimal supervision
H. O. Song, R. B. Girshick, S. Jegelka, J. Mairal, Z. Harchaoui, T. Darrell, et al · 2014
Earlier work this paper cites.
Weakly supervised object localization with latent category learning
C. Wang, W. Ren, K. Huang, and T. Tan · 2014
Earlier work this paper cites.
Edge Boxes: Locating object proposals from edges
C. L. Zitnick and P. Dollár · 2014
Cited alongside, same era.
Fast R-CNN
R. Girshick · 2015
Cited alongside, same era.
Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Cited alongside, same era.
ActivityNet: A large-scale video benchmark for human activity understanding
F. C. Heilbron, V. Escorcia, B. Ghanem, and J. C. Niebles · 2015
Cited alongside, same era.
Unsupervised object discovery and tracking in video collections
S. Kwak, M. Cho, I. Laptev, J. Ponce, and C. Schmid · 2015
Cited alongside, same era.
Towards computational baby learning: A weakly-supervised approach for object detection
X. Liang, S. Liu, Y. Wei, L. Liu, L. Lin, and S. Yan · 2015
Object detection from video tubelets with convolutional neural networks
K. Kang, W. Ouyang, H. Li, and X. Wang · 2016
Later among the works it cites.
ContextLocNet: Context-aware deep network models for weakly supervised localization
V. Kantorov, M. Oquab, M. Cho, and I. Laptev · 2016
Later among the works it cites.
Semantic object parsing with graph LSTM
X. Liang, X. Shen, J. Feng, L. Lin, and S. Yan · 2016
Later among the works it cites.
Semantic object parsing with local-global long short-term memory
X. Liang, X. Shen, D. Xiang, J. Feng, L. Lin, and S. Yan · 2016
Later among the works it cites.
SSD: Single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. E. Reed, C. Fu, and A. C. Berg · 2016
Later among the works it cites.
You only look once: Unified, real-time object detection
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Beyond short snippets: Deep networks for video classification
J. Y. Ng, M. J. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Cited alongside, same era.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
ImageNet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Cited alongside, same era.
Convolutional LSTM network: A machine learning approach for precipitation nowcasting
X. Shi, Z. Chen, H. Wang, D. Yeung, W. Wong, and W. Woo · 2015
Cited alongside, same era.
Unsupervised learning of video representations using LSTMs
N. Srivastava, E. Mansimov, and R. Salakhutdinov · 2015
Cited alongside, same era.
Improved semantic representations from tree-structured long short-term memory networks
K. S. Tai, R. Socher, and C. D. Manning · 2015
Cited alongside, same era.
Later among the works it cites.
Hollywood in homes: Crowdsourcing data collection for activity understanding
G. A. Sigurdsson, G. Varol, X. Wang, A. Farhadi, I. Laptev, and A. Gupta · 2016
Later among the works it cites.
Track and transfer: Watching videos to simulate strong human supervision for weakly-supervised object detection
K. K. Singh, F. Xiao, and Y. J. Lee · 2016
Later among the works it cites.
Video object discovery and co-segmentation with extremely weak supervision
L. Wang, G. Hua, R. Sukthankar, J. Xue, Z. Niu, and N. Zheng · 2016
Later among the works it cites.
End-to-end learning of action detection from frame glimpses in videos
S. Yeung, O. Russakovsky, G. Mori, and F. Li · 2016
Later among the works it cites.
Video summarization with long short-term memory
K. Zhang, W.-L. Chao, F. Sha, and K. Grauman · 2016
Later among the works it cites.
Deep self-taught learning for weakly supervised object localization
Z. Jie, Y. Wei, X. Jin, J. Feng, and W. Liu · 2017
Closest in time.
Interpretable structure-evolving LSTM
X. Liang, L. Lin, X. Shen, J. Feng, S. Yan, and E. Xing · 2017
Closest in time.
Revisiting unreasonable effectiveness of data in deep learning era
C. Sun, A. Shrivastava, S. Singh, and A. Gupta · 2017
Closest in time.
Revealing event saliency in unconstrained video collection
D. Zhang, J. Han, L. Jiang, S. Ye, and X. Chang · 2017
Closest in time.
Co-saliency detection via a self-paced multiple-instance learning framework
D. Zhang, D. Meng, and J. Han · 2017
Closest in time.