Fetching the paper…
Reading the bibliography…
This paper studies the complex task of simultaneous multi-object 3D reconstruction, 6D pose and size estimation from a single-view RGB-D observation.
C. Ferrari and J. F. Canny, “Planning Optimal Grasps.” in ICRA , vol. 3, no. 4, 1992, p. 6
1992
Earlier work this paper cites.
L. Van der Maaten and G. Hinton, “Visualizing data using t-SNE,” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
A. Segal, D. Haehnel, and S. Thrun, “Generalized-ICP,” in Robotics: science and systems , vol. 2, no. 4. Seattle, WA, 2009, p. 435
2009
Earlier work this paper cites.
R. Girshick, J. Donahue, T. Darrell, and J. Malik, “Rich feature hierarchies for accurate object detection and semantic segmentation,” in Computer Vision and Pattern Recognition , 2014
2014
Earlier work this paper cites.
A. Tejani, D. Tang, R. Kouskouridas, and T.-K. Kim, “Latent-class hough forests for 3d object detection and pose estimation,” in European Conference on Computer Vision . Springer, 2014, pp. 462–477
2014
Earlier work this paper cites.
X. Chen, K. Kundu, Y. Zhu, A. G. Berneshawi, H. Ma, S. Fidler, and R. Urtasun, “3D object proposals for accurate object class detection,” in Advances in Neural Information Processing Systems . Citeseer, 2015, pp. 424–432
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster R-CNN: Towards real-time object detection with region proposal networks,” Advances in neural information processing systems , vol. 28, pp. 91–99, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
C. G. Cifuentes, J. Issac, M. Wüthrich, S. Schaal, and J. Bohg, “Probabilistic articulated real-time tracking for robot manipulation,” IEEE Robotics and Automation Letters , vol. 2, no. 2, pp. 577–584, 2016
2016
Earlier work this paper cites.
C. B. Choy, D. Xu, J. Gwak, K. Chen, and S. Savarese, “3d-r2n2: A unified approach for single and multi-view 3d object reconstruction,” in European conference on computer vision . Springer, 2016, pp. 628–644
2016
Earlier work this paper cites.
W. Kehl, F. Milletari, F. Tombari, S. Ilic, and N. Navab, “Deep learning of local rgb-d patches for 3d object detection and 6d pose estimation,” in European conference on computer vision . Springer, 2016, pp. 205–220
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
W. Kehl, F. Manhardt, F. Tombari, S. Ilic, and N. Navab, “SDD-6D: Making RGB-based 3D detection and 6d pose estimation great again,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 1521–1529
2017
Earlier work this paper cites.
M. Rad and V. Lepetit, “BB8: A scalable, accurate, robust to partial occlusion method for predicting the 3d poses of challenging objects without using depth,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 3828–3836
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-CNN,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 2961–2969
2017
Earlier work this paper cites.
H. Fan, H. Su, and L. J. Guibas, “A point set generation network for 3d object reconstruction from a single image,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 605–613
2017
Earlier work this paper cites.
J. Varley, C. DeChant, A. Richardson, J. Ruales, and P. Allen, “Shape completion enabled robotic grasping,” in 2017 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 2017, pp. 2442–2447
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, and S. Belongie, “Feature pyramid networks for object detection,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 2117–2125
2017
Earlier work this paper cites.
C. R. Qi, H. Su, K. Mo, and L. J. Guibas, “PointNet: Deep learning on point sets for 3d classification and segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 652–660
2017
Earlier work this paper cites.
C. R. Qi, W. Liu, C. Wu, H. Su, and L. J. Guibas, “Frustum Pointnets for 3D object detection from RGB-D data,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 918–927
2018
Earlier work this paper cites.
D. Kappler, F. Meier, J. Issac, J. Mainprice, C. G. Cifuentes, M. Wüthrich, V. Berenz, S. Schaal, N. Ratliff, and J. Bohg, “Real-time perception meets reactive motion generation,” IEEE Robotics and Automation Letters , vol. 3, no. 3, pp. 1864–1871, 2018
2018
Earlier work this paper cites.
Y. Xiang, T. Schmidt, V. Narayanan, and D. Fox, “PoseCNN: A convolutional neural network for 6d object pose estimation in cluttered scenes,” 2018
2018
Earlier work this paper cites.
B. Tekin, S. N. Sinha, and P. Fua, “Real-time seamless single shot 6d object pose prediction,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 292–301
2018
Earlier work this paper cites.
H. Kato, Y. Ushiku, and T. Harada, “Neural 3D mesh renderer,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 3907–3916
2018
Cited alongside, same era.
T. Groueix, M. Fisher, V. G. Kim, B. Russell, and M. Aubry, “AtlasNet: A Papier-Mâché Approach to Learning 3D Surface Generation,” in Proceedings IEEE Conf. on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Cited alongside, same era.
B. Yang, S. Rosa, A. Markham, N. Trigoni, and H. Wen, “Dense 3d object reconstruction from a single depth view,” in TPAMI , 2018
2018
Cited alongside, same era.
W. Yuan, T. Khot, D. Held, C. Mertz, and M. Hebert, “PCN: Point completion network,” in 3D Vision (3DV), 2018 International Conference on , 2018
2018
Cited alongside, same era.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala, “Pytorch: An imperative style, high-performance deep learning library,” in Advances in Neural Information Processing Systems 32 , H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alché-Buc, E. Fox, and R. Garnett, Eds. Curran Associates, Inc., 2019, pp. 8024–8035
2019
Later among the works it cites.
Y. Nie, X. Han, S. Guo, Y. Zheng, J. Chang, and J. J. Zhang, “Total3DUnderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2020
2020
Later among the works it cites.
W. Kuo, A. Angelova, T.-Y. Lin, and A. Dai, “Mask2CAD: 3D shape prediction by learning to segment and retrieve,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part III 16 . Springer, 2020, pp. 260–277
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
H. Law and J. Deng, “CornerNet: Detecting objects as paired keypoints,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 734–750
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Y. Yang, C. Feng, Y. Shen, and D. Tian, “FoldingNet: Point cloud auto-encoder via deep grid deformation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 206–215
2018
Cited alongside, same era.
L. Mescheder, M. Oechsle, M. Niemeyer, S. Nowozin, and A. Geiger, “Occupancy networks: Learning 3D reconstruction in function space,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 4460–4470
2019
Cited alongside, same era.
S. Peng, Y. Liu, Q. Huang, X. Zhou, and H. Bao, “PVNet: Pixel-wise voting network for 6dof pose estimation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 4561–4570
2019
Cited alongside, same era.
C. Wang, D. Xu, Y. Zhu, R. Martín-Martín, C. Lu, L. Fei-Fei, and S. Savarese, “Densefusion: 6d object pose estimation by iterative dense fusion,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 3343–3352
2019
Cited alongside, same era.
G. Gkioxari, J. Malik, and J. Johnson, “Mesh R-CNN,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 9785–9795
2019
Cited alongside, same era.
M. Niemeyer, L. Mescheder, M. Oechsle, and A. Geiger, “Differentiable volumetric rendering: Learning implicit 3d representations without 3d supervision,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 3504–3515
2020
Later among the works it cites.
M. Tian, M. H. Ang, and G. H. Lee, “Shape prior deformation for categorical 6d object pose and size estimation,” in European Conference on Computer Vision . Springer, 2020, pp. 530–546
2020
Later among the works it cites.
M. Sundermeyer, Z.-C. Marton, M. Durner, and R. Triebel, “Augmented autoencoders: Implicit 3D orientation learning for 6d object detection,” International Journal of Computer Vision , vol. 128, no. 3, pp. 714–729, 2020
2020
Later among the works it cites.
X. Zhou, V. Koltun, and P. Krähenbühl, “Tracking objects as points,” in European Conference on Computer Vision . Springer, 2020, pp. 474–490
2020
Later among the works it cites.
M. Runz, K. Li, M. Tang, L. Ma, C. Kong, T. Schmidt, I. Reid, L. Agapito, J. Straub, S. Lovegrove et al. , “Frodo: From detections to 3d objects,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 14 720–14 729
2020
Later among the works it cites.
D. Chen, J. Li, Z. Wang, and K. Xu, “Learning canonical shape space for category-level 6d object pose and size estimation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 11 973–11 982
2020
Later among the works it cites.
Y. Wang, Z. Xu, H. Shen, B. Cheng, and L. Yang, “Centermask: single shot instance segmentation with point representation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 9313–9321
2020
Later among the works it cites.
Z. Tian, C. Shen, and H. Chen, “Conditional convolutions for instance segmentation,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part I 16 . Springer, 2020, pp. 282–298
2020
Later among the works it cites.
X. Wang, T. Kong, C. Shen, Y. Jiang, and L. Li, “Solo: Segmenting objects by locations,” in European Conference on Computer Vision . Springer, 2020, pp. 649–665
2020
Later among the works it cites.
Y. Sun, Q. Bao, W. Liu, Y. Fu, and T. Mei, “Centerhmr: a bottom-up single-shot method for multi-person 3d mesh recovery from a single image,” arXiv e-prints , pp. arXiv–2008, 2020
2020
Later among the works it cites.
X. Chen, Z. Dong, J. Song, A. Geiger, and O. Hilliges, “Category level object pose estimation via neural analysis-by-synthesis,” in European Conference on Computer Vision . Springer, 2020, pp. 139–156
2020
Later among the works it cites.
C. Wang, R. Martín-Martín, D. Xu, J. Lv, C. Lu, L. Fei-Fei, S. Savarese, and Y. Zhu, “6-pack: Category-level 6d pose tracker with anchor-based keypoints,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 10 059–10 066
2020
Later among the works it cites.
Z. Jiang, Y. Zhu, M. Svetlik, K. Fang, and Y. Zhu, “Synergies between affordance and geometry: 6-Dof grasp detection via implicit representations,” Robotics: science and systems , 2021
2021
Later among the works it cites.
M. Laskey, B. Thananjeyan, K. Stone, T. Kollar, and M. Tjersland, “SimNet: Enabling robust unknown object manipulation from pure synthetic data via stereo,” in 5th Annual Conference on Robot Learning , 2021
2021
Later among the works it cites.
C. Zhang, Z. Cui, Y. Zhang, B. Zeng, M. Pollefeys, and S. Liu, “Holistic 3D scene understanding from a single image with implicit representation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 8833–8842
2021
Later among the works it cites.
F. Engelmann, K. Rematas, B. Leibe, and V. Ferrari, “From points to multi-object 3d reconstruction,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 4588–4597
2021
Later among the works it cites.
Y. Sun, Q. Bao, W. Liu, Y. Fu, B. Michael J., and T. Mei, “Monocular, one-stage, regression of multiple 3d people,” in ICCV , October 2021
2021
Later among the works it cites.
T. Lee, B.-U. Lee, M. Kim, and I. S. Kweon, “Category-level metric scale object shape and pose estimation,” IEEE Robotics and Automation Letters , vol. 6, no. 4, pp. 8575–8582, 2021
2021
Later among the works it cites.
B. Wen and K. E. Bekris, “Bundletrack: 6d pose tracking for novel objects without instance or category-level 3d models,” in IEEE/RSJ International Conference on Intelligent Robots and Systems , 2021
2021
Later among the works it cites.