Fetching the paper…
Reading the bibliography…
In this work, we develop a monocular SLAM-aware object recognition system that is able to achieve considerably stronger recognition performance, as compared to classical object recognition systems that function on a frame-by-frame basis.
Video google: A text retrieval approach to object matching in videos
J. Sivic and A. Zisserman · 2003
Earlier work this paper cites.
Visual categorization with bags of keypoints
G. Csurka, C. Dance, L. Fan, J. Willamowski, and C. Bray · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection
N. Dalal and B. Triggs · 2005
Earlier work this paper cites.
Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories
S. Lazebnik, C. Schmid, and J. Ponce · 2006
Earlier work this paper cites.
Towards multi-view object class detection
A. Thomas, V. Ferrar, B. Leibe, T. Tuytelaars, B. Schiel, and L. Van Gool · 2006
Earlier work this paper cites.
Image classification using random forests and ferns
A. Bosch, A. Zisserman, and X. Muoz · 2007
Earlier work this paper cites.
Constrained parametric min-cuts for automatic object segmentation
J. Carreira and C. Sminchisescu · 2010
Earlier work this paper cites.
Combining monoSLAM with object recognition for scene augmentation using a wearable camera
R. O. Castle, G. Klein, and D. W. Murray · 2010
Earlier work this paper cites.
Efficient multi-view object recognition and full pose estimation
A. Collet and S. S. Srinivasa · 2010
Earlier work this paper cites.
The PASCAL Visual Object Classes (VOC) Challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Object detection with discriminatively trained part-based models
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan · 2010
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
H. Jégou, M. Douze, C. Schmid, and P. Pérez · 2010
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Cited alongside, same era.
Image classification using super-vector coding of local image descriptors
X. Zhou, K. Yu, T. Zhang, and T. S. Huang · 2010
Cited alongside, same era.
Semantic structure from motion
S. Y. Bao and S. Savarese · 2011
Cited alongside, same era.
Hierarchical matching pursuit for image classification: Architecture and fast algorithms
L. Bo, X. Ren, and D. Fox · 2011
Cited alongside, same era.
The devil is in the details: an evaluation of recent feature encoding methods
K. Chatfield, V. Lempitsky, A. Vedaldi, and A. Zisserman · 2011
Cited alongside, same era.
Towards semantic SLAM using a monocular camera
J. Civera, D. Gálvez-López, L. Riazuelo, J. D. Tardós, and J. Montiel · 2011
Cited alongside, same era.
SLAM++: Simultaneous localisation and mapping at the level of objects
R. F. Salas-Moreno, R. A. Newcombe, H. Strasdat, P. H. Kelly, and A. J. Davison · 2013
Later among the works it cites.
Selective search for object recognition
J. R. Uijlings, K. E. van de Sande, T. Gevers, and A. W. Smeulders · 2013
Later among the works it cites.
BING: Binarized normed gradients for objectness estimation at 300fps
M.-M. Cheng, Z. Zhang, W.-Y. Lin, and P. Torr · 2014
Later among the works it cites.
LSD-SLAM: Large-scale direct monocular SLAM
J. Engel, T. Schöps, and D. Cremers · 2014
Later among the works it cites.
SVO: Fast semi-direct monocular visual odometry
C. Forster, M. Pizzoli, and D. Scaramuzza · 2014
Later among the works it cites.
Learning rich features from RGB-D images for object detection and segmentation
S. Gupta, R. Girshick, P. Arbelaez, and J. Malik · 2014
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A large-scale hierarchical multi-view RGB-D object dataset
K. Lai, L. Bo, X. Ren, and D. Fox · 2011
Cited alongside, same era.
Semantic structure from motion with points, regions, and objects
S. Y. Bao, M. Bagra, Y.-W. Chao, and S. Savarese · 2012
Cited alongside, same era.
Detection-based object labeling in 3D scenes
K. Lai, L. Bo, X. Ren, and D. Fox · 2012
Cited alongside, same era.
All about VLAD
R. Arandjelovic and A. Zisserman · 2013
Cited alongside, same era.
Fast, accurate detection of 100,000 object classes on a single machine
T. Dean, M. A. Ruzon, M. Segal, J. Shlens, S. Vijayanarasimhan, and J. Yagnik · 2013
Cited alongside, same era.
Revisiting the VLAD image representation
J. Delhumeau, P.-H. Gosselin, H. Jégou, and P. Pérez · 2013
Cited alongside, same era.
Later among the works it cites.
How good are detection proposals, really?
J. Hosang, R. Benenson, and B. Schiele · 2014
Later among the works it cites.
Unsupervised feature learning for 3D scene labeling
K. Lai, L. Bo, and D. Fox · 2014
Later among the works it cites.
Fisher and VLAD with FLAIR
K. E. van de Sande, C. G. Snoek, and A. W. Smeulders · 2014
Later among the works it cites.
Edge boxes: Locating object proposals from edges
C. L. Zitnick and P. Dollár · 2014
Later among the works it cites.
ORB-SLAM: a versatile and accurate monocular SLAM system
R. Mur-Artal, J. Montiel, and J. D. Tardos · 2015
Closest in time.
ImageNet Large Scale Visual Recognition Challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Closest in time.