Fetching the paper…
Reading the bibliography…
We address the task of aligning CAD models to a video sequence of a complex scene containing multiple objects.
Self-calibration and metric reconstruction inspite of varying and unknown intrinsic camera parameters
M. Pollefeys, R. Koch, and L. Van Gool · 1999
Earlier work this paper cites.
Recognizing objects in range data using regional point descriptors
A. Frome, D. Huber, R. Kolluri, T. Bülow, and J. Malik · 2004
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
A search-classify approach for cluttered indoor scene understanding
L. Nan, K. Xie, and A. Sharf · 2012
Earlier work this paper cites.
An interactive approach to semantic modeling of indoor scenes with an RGBD camera
T. Shao, W. Xu, K. Zhou, J. Wang, D. Li, and B. Guo · 2012
Earlier work this paper cites.
Real-time 3d reconstruction at scale using voxel hashing
M. Nießner, M. Zollhöfer, S. Izadi, and M. Stamminger · 2013
Earlier work this paper cites.
SLAM++: Simultaneous localisation and mapping at the level of objects
R. F. Salas-Moreno, R. A. Newcombe, H. Strasdat, P. H. Kelly, and A. J. Davison · 2013
Earlier work this paper cites.
Towards linear-time incremental structure from motion
C. Wu · 2013
Earlier work this paper cites.
Detecting people looking at each other in videos
M. J. Marin-Jimenez, A. Zisserman, M. Eichner, and V. Ferrari · 2014
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
A. X. Chang, T. Funkhouser, L. Guibas, P. Hanrahan, Q. Huang, Z. Li, S. Savarese, M. Savva, S. Song, H. Su, et al · 2015
Earlier work this paper cites.
Fast R-CNN
R. Girshick · 2015
Earlier work this paper cites.
Database-assisted object retrieval for real-time 3d reconstruction
Y. Li, A. Dai, L. Guibas, and M. Nießner · 2015
Earlier work this paper cites.
Fully convolutional models for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Earlier work this paper cites.
ORB-SLAM: a versatile and accurate monocular slam system
R. Mur-Artal, J. M. M. Montiel, and J. D. Tardos · 2015
Earlier work this paper cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Earlier work this paper cites.
3D-R2N2: A unified approach for single and multi-view 3D object reconstruction
C. B. Choy, D. Xu, J. Gwak, K. Chen, and S. Savarese · 2016
Earlier work this paper cites.
Learning a predictable and generative vector representation for objects
R. Girdhar, D. Fouhey, M. Rodriguez, and A. Gupta · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Structure-from-motion revisited
J. L. Schönberger and J.-M. Frahm · 2016
Cited alongside, same era.
Learning a probabilistic latent space of object shapes via 3D generative-adversarial modeling
J. Wu, C. Zhang, T. Xue, W. T. Freeman, and J. B. Tenenbaum · 2016
Cited alongside, same era.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner · 2017
Cited alongside, same era.
Bundlefusion: Real-time globally consistent 3d reconstruction using on-the-fly surface reintegration
A. Dai, M. Nießner, M. Zollhöfer, S. Izadi, and C. Theobalt · 2017
Cited alongside, same era.
A point set generation network for 3d object reconstruction from a single image
H. Fan, H. Su, and L. J. Guibas · 2017
Cited alongside, same era.
Mask R-CNN
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Shapemask: Learning to segment novel objects by refining shape priors
W. Kuo, A. Angelova, J. Malik, and T.-Y. Lin · 2019
Later among the works it cites.
Occupancy networks: Learning 3d reconstruction in function space
L. Mescheder, M. Oechsle, M. Niemeyer, S. Nowozin, and A. Geiger · 2019
Later among the works it cites.
Deepsdf: Learning continuous signed distance functions for shape representation
J. J. Park, P. Florence, J. Straub, R. Newcombe, and S. Lovegrove · 2019
Later among the works it cites.
What do single-view 3d reconstruction networks learn?
M. Tatarchenko, S. R. Richter, R. Ranftl, Z. Li, V. Koltun, and T. Brox · 2019
Later among the works it cites.
SceneCAD: Predicting object alignments and layouts in RGB-D scans
A. Avetisyan, T. Khanova, C. Choy, D. Dash, A. Dai, and M. Nießner · 2020
Closest in time.
Bsp-net: Generating compact meshes via binary space partitioning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Fully convolutional instance-aware semantic segmentation
Y. Li, H. Qi, J. Dai, X. Ji, and Y. Wei · 2017
Cited alongside, same era.
DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. Yuille · 2018
Cited alongside, same era.
Visual-inertial object detection and mapping
X. Fei and S. Soatto · 2018
Cited alongside, same era.
Holistic 3D scene parsing and reconstruction from a single RGB image
S. Huang, S. Qi, Y. Zhu, Y. Xiao, Y. Xu, and S.-C. Zhu · 2018
Cited alongside, same era.
3D-RCNN: Instance-level 3d object reconstruction via render-and-compare
A. Kundu, Y. Li, and J. M. Rehg · 2018
Cited alongside, same era.
3d-lmnet: Latent embedding matching for accurate and diverse 3d point cloud reconstruction from a single image
P. Mandikal, N. K. L., M. Agarwal, and V. B. Radhakrishnan · 2018
Cited alongside, same era.
Z. Chen, A. Tagliasacchi, and H. Zhang · 2020
Closest in time.
Scene recomposition by learning-based icp
H. Izadinia and S. M. Seitz · 2020
Closest in time.
Mask2CAD: 3D shape prediction by learning to segment and retrieve
W. Kuo, A. Angelova, T.-Y. Lin, and A. Dai · 2020
Closest in time.
Mo-ltr: Multiple object localization, tracking, and reconstruction from monocular RGB videos
K. Li, H. Rezatofighi, and I. Reid · 2020
Closest in time.
Total3dunderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image
Y. Nie, X. Han, S. Guo, Y. Zheng, J. Chang, and J. J. Zhang · 2020
Closest in time.
CoReNet: Coherent 3D scene reconstruction from a single RGB image
S. Popov, P. Bauszat, and V. Ferrari · 2020
Closest in time.
Associative3d: Volumetric reconstruction from sparse views
S. Qian, L. Jin, and D. F. Fouhey · 2020
Closest in time.
Accelerating 3d deep learning with pytorch3d
N. Ravi, J. Reizenstein, D. Novotny, T. Gordon, W.-Y. Lo, J. Johnson, and G. Gkioxari · 2020
Closest in time.
Frodo: From detections to 3d objects
M. Runz, K. Li, M. Tang, L. Ma, C. Kong, T. Schmidt, I. Reid, L. Agapito, J. Straub, S. Lovegrove, et al · 2020
Closest in time.
Pix2vox++: multi-scale context-aware 3d object reconstruction from single and multiple images
H. Xie, H. Yao, S. Zhang, S. Zhou, and W. Sun · 2020
Closest in time.
Perceiving 3d human-object spatial arrangements from a single image in the wild
J. Y. Zhang, S. Pepose, H. Joo, D. Ramanan, J. Malik, and A. Kanazawa · 2020
Closest in time.
Deepvideomvs: Multi-view stereo on video with recurrent spatio-temporal fusion
A. Duzceker, S. Galliani, C. Vogel, P. Speciale, M. Dusmanu, and M. Pollefeys · 2021
Closest in time.
Odam: Object detection, association, and mapping using posed rgb video
K. Li, D. DeTone, Y. F. S. Chen, M. Vo, I. Reid, H. Rezatofighi, C. Sweeney, J. Straub, and R. Newcombe · 2021
Closest in time.