Fetching the paper…
Reading the bibliography…
We propose a Convolutional Neural Network (CNN)-based model "RotationNet," which takes multi-view images of an object as input and jointly estimates its pose and object category.
Appearance-based active object recognition
H. Borotschnig, L. Paletta, M. Prantl, and A. Pinz · 2000
Earlier work this paper cites.
Active object recognition by view integration and reinforcement learning
L. Paletta and A. Pinz · 2000
Earlier work this paper cites.
On visual similarity based 3D model retrieval
D.-Y. Chen, X.-P. Tian, Y.-T. Shen, and M. Ouhyoung · 2003
Earlier work this paper cites.
3D generic object categorization, localization and pose estimation
S. Savarese and L. Fei-Fei · 2007
Earlier work this paper cites.
A large-scale hierarchical multi-view rgb-d object dataset
K. Lai, L. Bo, X. Ren, and D. Fox · 2011
Earlier work this paper cites.
A scalable tree-based approach for joint object and pose recognition
K. Lai, L. Bo, X. Ren, and D. Fox · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Unsupervised feature learning for rgb-d based object recognition
L. Bo, X. Ren, and D. Fox · 2013
Earlier work this paper cites.
Joint object and pose recognition using homeomorphic manifold analysis
H. Zhang, T. El-Gaaly, A. M. Elgammal, and Z. Jiang · 2013
Earlier work this paper cites.
Untangling object-view manifold for multiview recognition and pose estimation
A. Bakry and A. Elgammal · 2014
Earlier work this paper cites.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Inferring unseen views of people
C.-Y. Chen and K. Grauman · 2014
Earlier work this paper cites.
Lsd-slam: Large-scale direct monocular slam
J. Engel, T. Schöps, and D. Cremers · 2014
Earlier work this paper cites.
Multi-view perceptron: a deep model for learning face identity and view representations
Z. Zhu, P. Luo, X. Wang, and X. Tang · 2014
Earlier work this paper cites.
Voxnet: A 3d convolutional neural network for real-time object recognition
D. Maturana and S. Scherer · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Cited alongside, same era.
Deeppano: Deep panoramic representation for 3-d shape recognition
B. Shi, S. Bai, Z. Zhou, and X. Bai · 2015
Cited alongside, same era.
Multi-view convolutional neural networks for 3D shape recognition
H. Su, S. Maji, E. Kalogerakis, and E. G. Learned-Miller · 2015
Cited alongside, same era.
Render for cnn: Viewpoint estimation in images using cnns trained with rendered 3D model views
H. Su, C. R. Qi, Y. Li, and L. J. Guibas · 2015
Cited alongside, same era.
3D-assisted feature synthesis for novel views of an object
H. Su, F. Wang, E. Yi, and L. J. Guibas · 2015
Cited alongside, same era.
3d shapenets: A deep representation for volumetric shapes
Z. Wu, S. Song, A. Khosla, F. Yu, L. Zhang, X. Tang, and J. Xiao · 2015
Deep learning with sets and point clouds
S. Ravanbakhsh, J. Schneider, and B. Poczos · 2016
Closest in time.
Deep learning 3D shape surfaces using geometry images
A. Sinha, J. Bai, and K. Ramani · 2016
Closest in time.
Learning a probabilistic latent space of object shapes via 3D generative-adversarial modeling
J. Wu, C. Zhang, T. Xue, W. T. Freeman, and J. B. Tenenbaum · 2016
Closest in time.
Beam search for learning a deep convolutional neural network of 3d shapes
X. Xu and S. Todorovic · 2016
Closest in time.
Generative and discriminative voxel modeling with convolutional neural networks
A. Brock, T. Lim, J. Ritchie, and N. Weston · 2017
Closest in time.
Escape from cells: Deep kd-networks for the recognition of 3d point cloud models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gift: A real-time and scalable 3d shape search engine
S. Bai, X. Bai, Z. Zhou, Z. Zhang, and L. J. Latecki · 2016
Cited alongside, same era.
A comparative analysis and study of multiview cnn models for joint object categorization and pose estimation
M. Elhoseiny, T. El-Gaaly, A. Bakry, and A. Elgammal · 2016
Cited alongside, same era.
Pointnet: A 3d convolutional neural network for real-time object class recognition
A. Garcia-Garcia, F. Gomez-Donoso, J. Garcia-Rodriguez, S. Orts-Escolano, M. Cazorla, and J. Azorin-Lopez · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Fusionnet: 3d object classification using multiple data representations
V. Hegde and R. Zadeh · 2016
Cited alongside, same era.
Pairwise decomposition of image sequences for active multi-view recognition
E. Johns, S. Leutenegger, and A. J. Davison · 2016
Cited alongside, same era.
R. Klokov and V. Lempitsky · 2017
Closest in time.
Learning 3d object categories by looking around them
D. Novotny, D. Larlus, and A. Vedaldi · 2017
Closest in time.
Pointnet: Deep learning on point sets for 3D classification and segmentation
C. R. Qi, H. Su, K. Mo, and L. J. Guibas · 2017
Closest in time.
Orientation-boosted voxel nets for 3D object recognition
N. Sedaghat, M. Zolfaghari, and T. Brox · 2017
Closest in time.
Exploiting the panorama representation for convolutional neural network classification and retrieval
K. Sfikas, T. Theoharis, and I. Pratikakis · 2017
Closest in time.
Dynamic edge-conditioned filters in convolutional neural networks on graphs
M. Simonovsky and N. Komodakis · 2017
Closest in time.
Dominant set clustering and pooling for multi-view 3d object recognition
C. Wang, M. Pelillo, and K. Siddiqi · 2017
Closest in time.
Deep learning for 3d shape classification from multiple depth maps
P. Zanuttigh and L. Minto · 2017
Closest in time.
Lightnet: A lightweight 3D convolutional neural network for real-time 3D object recognition
S. Zhi, Y. Liu, X. Li, and Y. Guo · 2017
Closest in time.
Unsupervised learning of depth and ego-motion from video
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe · 2017
Closest in time.