Fetching the paper…
Reading the bibliography…
We introduce a new large-scale data set of video URLs with densely-sampled object bounding box annotations called YouTube-BoundingBoxes (YT-BB).
Columbia object image library (coil-20)
S. A. Nene, S. K. Nayar, H. Murase, et al · 1996
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
WordNet: An Electronic Lexical Database
C. Fellbaum · 1998
Earlier work this paper cites.
A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics
D. Martin, C. Fowlkes, D. Tal, and J. Malik · 2001
Earlier work this paper cites.
Sharing features: efficient boosting procedures for multiclass object detection
A. Torralba, K. P. Murphy, and W. T. Freeman · 2004
Earlier work this paper cites.
Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories
L. Fei-Fei, R. Fergus, and P. Perona · 2007
Earlier work this paper cites.
Caltech-256 object category dataset
G. Griffin, A. Holub, and P. Perona · 2007
Earlier work this paper cites.
A discriminative kernel-based model to rank images from text queries
D. Grangier and S. Bengio · 2008
Earlier work this paper cites.
Get another label? improving data quality and data mining
V. Sheng, F. Provost, and P. Ipeirotis · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Pedestrian detection: A benchmark
P. Dollár, C. Wojek, B. Schiele, and P. Perona · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Quality management on amazon mechanical turk
P. G. Ipeirotis, F. Provost, and J. Wang · 2010
Earlier work this paper cites.
Sun database: Large-scale scene recognition from abbey to zoo
J. Xiao, J. Hays, K. A. Ehinger, A. Oliva, and A. Torralba · 2010
Earlier work this paper cites.
Amazon’s mechanical turk a new source of inexpensive, yet high-quality, data?
M. Buhrmester, T. Kwang, and S. D. Gosling · 2011
Earlier work this paper cites.
Opportunities for crowdsourcing research on amazon mechanical turk
J. J. Chen, N. J. Menezes, A. D. Bradley, and T. North · 2011
Earlier work this paper cites.
Shepherding the crowd: managing and providing feedback to crowd workers
S. Dow, A. Kulkarni, B. Bunge, T. Nguyen, S. Klemmer, and B. Hartmann · 2011
Earlier work this paper cites.
HMDB: a large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Human computation must be reproducible
P. Paritosh · 2012
Cited alongside, same era.
Learning object class detectors from weakly annotated video
A. Prest, C. Leistner, J. Civera, C. Schmid, and V. Ferrari · 2012
Cited alongside, same era.
Ucf101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Cited alongside, same era.
Keep it simple: Reward and task design in crowdsourcing
A. Finnerty, P. Kucherbaev, S. Tranquillini, and G. Convertino · 2013
Cited alongside, same era.
Zero-shot learning by convex combination of semantic embeddings
M. Norouzi, T. Mikolov, S. Bengio, Y. Singer, J. Shlens, A. Frome, G. Corrado, and J. Dean · 2013
Cited alongside, same era.
The pascal visual object classes challenge: A retrospective
M. Everingham, S. A. Eslami, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2015
Later among the works it cites.
THUMOS challenge: Action recognition with a large number of classes
A. Gorban, H. Idrees, Y.-G. Jiang, A. Roshan Zamir, I. Laptev, M. Shah, and R. Sukthankar · 2015
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Later among the works it cites.
The visual object tracking vot2015 challenge results
M. Kristan, J. Matas, A. Leonardis, M. Felsberg, L. Cehovin, G. Fernandez, T. Vojir, G. Hager, G. Nebehay, and R. Pflugfelder · 2015
Later among the works it cites.
MOTChallenge 2015: Towards a benchmark for multi-target tracking
L. Leal-Taixé, A. Milan, I. Reid, S. Roth, and K. Schindler · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
O. Russakovsky, J. Deng, Z. Huang, A. C. Berg, and L. Fei-Fei · 2013
Cited alongside, same era.
Discriminative segment annotation in weakly labeled video
K. Tang, R. Sukthankar, J. Yagnik, and L. Fei-Fei · 2013
Cited alongside, same era.
Pay by the bit: an information-theoretic metric for collective human judgment
T. P. Waterhouse · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, Ç. Gülçehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Cited alongside, same era.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Cited alongside, same era.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Cited alongside, same era.
Deeptrack: Learning discriminative feature representations online for robust visual tracking
H. Li, Y. Li, and F. Porikli · 2015
Later among the works it cites.
Beyond short snippets: Deep networks for video classification
J. Y.-H. Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Later among the works it cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Later among the works it cites.
ImageNet Large Scale Visual Recognition Challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Later among the works it cites.
Rethinking the inception architecture for computer vision
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna · 2015
Later among the works it cites.
Visual tracking with fully convolutional networks
L. Wang, W. Ouyang, X. Wang, and H. Lu · 2015
Later among the works it cites.
Youtube-8m: A large-scale video classification benchmark, 2016
S. Abu-El-Haija, N. Kothari, J. Lee, P. Natsev, G. Toderici, B. Varadarajan, and S. Vijayanarasimhan · 2016
Later among the works it cites.
Trecvid 2016: Evaluating video search, video event detection, localization, and hyperlinking
G. Awad, J. Fiscus, M. Michel, D. Joy, W. Kraaij, A. F. Smeaton, G. Quénot, M. Eskevich, R. Aly, G. J. F. Jones, R. Ordelman, B. Huet, and M. Larson · 2016
Later among the works it cites.
http://github.com/BVLC/caffe/wiki/Model-Zoo
Caffe Model Zoo · 2016
Later among the works it cites.
Speed/accuracy trade-offs for modern convolutional object detectors
J. Huang, V. Rathod, C. Sun, M. Zhu, A. Korattikara, A. Fathi, I. Fischer, Z. Wojna, Y. Song, S. Guadarrama, et al · 2016
Later among the works it cites.
The YouTube Sports-1M Dataset
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2016
Later among the works it cites.
Youtube-Objects dataset
A. Prest, C. Leistner, J. Civera, C. Schmid, and V. Ferrari · 2016
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
C. Szegedy, S. Ioffe, and V. Vanhoucke · 2016
Later among the works it cites.
http://github.com/tensorflow/models/tree/master/slim
TensorFlow-Slim image classification library · 2016
Later among the works it cites.