Fetching the paper…
Reading the bibliography…
In this paper, we rethink the problem of scene reconstruction from an embodied agent's perspective: While the classic view focuses on the reconstruction accuracy, our new perspective emphasizes the underlying functions and constraints such that the reconstructed scenes provide \em{actionable} information for simulating \em{interactions} with agents.
Houghton Mifflin, 1950
J. J. Gibson, The perception of the visual world · 1950
Earlier work this paper cites.
Houghton Mifflin, 1966
J. J. Gibson, The senses considered as perceptual systems · 1966
Earlier work this paper cites.
J. J. Moré, “The levenberg-marquardt algorithm: implementation and theory,” in Numerical analysis
1978
Earlier work this paper cites.
R. Jonker and A. Volgenant, “A shortest augmenting path algorithm for dense and sparse linear assignment problems,” Computing
1987
Earlier work this paper cites.
K. Ikeuchi and M. Hebert, “Task oriented vision,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
1992
Earlier work this paper cites.
S. Minton, M. D. Johnston, A. B. Philips, and P. Laird, “Minimizing conflicts: a heuristic repair method for constraint satisfaction and scheduling problems,” Artificial intelligence
1992
Earlier work this paper cites.
G. Malandain and J.-D. Boissonnat, “Computing the diameter of a point set,” International Journal of Computational Geometry & Applications
2002
Earlier work this paper cites.
Cambridge university press, 2003
R. Hartley and A. Zisserman, Multiple view geometry in computer vision · 2003
Earlier work this paper cites.
N. P. Koenig and A. Howard, “Design and use paradigms for gazebo, an open-source multi-robot simulator.,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2004
Earlier work this paper cites.
S.-C. Zhu and D. Mumford, “A stochastic grammar of images,” Foundations and Trends® in Computer Graphics and Vision
2007
Earlier work this paper cites.
L. P. Kaelbling and T. Lozano-Pérez, “Hierarchical task and motion planning in the now,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2011
Earlier work this paper cites.
L. F. Yu, S. K. Yeung, C. K. Tang, D. Terzopoulos, T. F. Chan, and S. J. Osher, “Make it home: automatic optimization of furniture arrangement,” ACM Transactions on Graphics (TOG)
2011
Earlier work this paper cites.
Y. Zhao and S.-C. Zhu, “Image parsing with stochastic scene grammar,” in Proceedings of Advances in Neural Information Processing Systems (NeurIPS)
2011
Earlier work this paper cites.
A. Pronobis and P. Jensfelt, “Large-scale semantic mapping and reasoning with heterogeneous modalities,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2012
Earlier work this paper cites.
B. Zheng, Y. Zhao, J. C. Yu, K. Ikeuchi, and S.-C. Zhu, “Beyond point clouds: Scene understanding by reasoning geometry and physics,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2013
Earlier work this paper cites.
Y. Zhao and S.-C. Zhu, “Scene parsing by integrating function, geometry and appearance models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2013
Earlier work this paper cites.
E. Rohmer, S. P. Singh, and M. Freese, “V-rep: A versatile and scalable robot simulation framework,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2013
Earlier work this paper cites.
Y. Taguchi, Y.-D. Jian, S. Ramalingam, and C. Feng, “Point-plane slam for hand-held 3d sensors,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2013
Earlier work this paper cites.
B. Zheng, Y. Zhao, C. Y. Joey, K. Ikeuchi, and S.-C. Zhu, “Detecting potential falling objects by inferring human action and natural disturbance,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2014
Earlier work this paper cites.
S. Srivastava, E. Fang, L. Riano, R. Chitnis, S. Russell, and P. Abbeel, “Combined task and motion planning through an extensible planner-independent interface layer,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Proceedings of European Conference on Computer Vision (ECCV)
2014
Earlier work this paper cites.
Y. Zhu, Y. Zhao, and S.-C. Zhu, “Understanding tools: Task-oriented object modeling, learning and recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2015
Earlier work this paper cites.
A. Myers, C. L. Teo, C. Fermüller, and Y. Aloimonos, “Affordance detection of tool parts from geometric features,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2015
Earlier work this paper cites.
B. Zheng, Y. Zhao, J. Yu, K. Ikeuchi, and S.-C. Zhu, “Scene understanding by reasoning stability and safety,” International Journal of Robotics Research (IJRR)
2015
Earlier work this paper cites.
2015
Cited alongside, same era.
Y. Zhu, C. Jiang, Y. Zhao, D. Terzopoulos, and S.-C. Zhu, “Inferring forces and learning human utilities from videos,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2016
Cited alongside, same era.
C. Cadena, L. Carlone, H. Carrillo, Y. Latif, D. Scaramuzza, J. Neira, I. Reid, and J. J. Leonard, “Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age,” Transactions on Robotics (T-RO)
2016
Cited alongside, same era.
B.-S. Hua, Q.-H. Pham, D. T. Nguyen, M.-K. Tran, L.-F. Yu, and S.-K. Yeung, “Scenenn: A scene meshes dataset with annotations,” in Proceedings of International Conference on 3D Vision (3DV)
2016
Cited alongside, same era.
A. Kirillov, K. He, R. Girshick, C. Rother, and P. Dollár, “Panoptic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Later among the works it cites.
X. Xie, H. Liu, Z. Zhang, Y. Qiu, F. Gao, S. Qi, Y. Zhu, and S.-C. Zhu, “Vrgym: A virtual testbed for physical and interactive ai,” in Proceedings of the ACM Turing Celebration Conference-China
2019
Later among the works it cites.
Q.-H. Pham, B.-S. Hua, T. Nguyen, and S.-K. Yeung, “Real-time progressive 3d semantic segmentation for indoor scenes,” in Winter Conference on Applications of Computer Vision (WACV)
2019
Later among the works it cites.
S. Yang and S. Scherer, “Monocular object and plane slam in structured environments,” Robotics and Automation Letters (RA-L)
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. McCormac, A. Handa, A. Davison, and S. Leutenegger, “Semanticfusion: Dense 3d semantic mapping with convolutional neural networks,” in Proceedings of International Conference on Robotics and Automation (ICRA)
2017
Cited alongside, same era.
E. Kolve, R. Mottaghi, D. Gordon, Y. Zhu, A. Gupta, and A. Farhadi, “Ai2-thor: An interactive 3d environment for visual ai,” 2017
2017
Cited alongside, same era.
S. Song, F. Yu, A. Zeng, A. X. Chang, M. Savva, and T. Funkhouser, “Semantic scene completion from a single depth image,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2017
Cited alongside, same era.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner, “Scannet: Richly-annotated 3d reconstructions of indoor scenes,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2017
Cited alongside, same era.
M. Edmonds, F. Gao, X. Xie, H. Liu, S. Qi, Y. Zhu, B. Rothrock, and S.-C. Zhu, “Feeling the force: Integrating force and pose for fluent discovery through imitation learning to open medicine bottles,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2017
Cited alongside, same era.
S. Huang, S. Qi, Y. Xiao, Y. Zhu, Y. N. Wu, and S.-C. Zhu, “Cooperative holistic scene understanding: Unifying 3d object, layout, and camera pose estimation,” in Proceedings of Advances in Neural Information Processing Systems (NeurIPS)
2018
Cited alongside, same era.
Z. Wang, C. R. Garrett, L. P. Kaelbling, and T. Lozano-Pérez, “Active model learning and diverse action sampling for task and motion planning,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2018
Cited alongside, same era.
J. McCormac, R. Clark, M. Bloesch, A. Davison, and S. Leutenegger, “Fusion++: Volumetric object-level slam,” in Proceedings of International Conference on 3D Vision (3DV)
2018
Cited alongside, same era.
L. Yi, W. Zhao, H. Wang, M. Sung, and L. J. Guibas, “Gspn: Generative shape proposal network for 3d instance segmentation in point cloud,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Later among the works it cites.
Q.-H. Pham, T. Nguyen, B.-S. Hua, G. Roig, and S.-K. Yeung, “Jsis3d: joint semantic-instance segmentation of 3d point clouds with multi-task pointwise networks and multi-value conditional random fields,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Later among the works it cites.
A. Avetisyan, M. Dahnert, A. Dai, M. Savva, A. X. Chang, and M. Nießner, “Scan2cad: Learning cad model alignment in rgb-d scans,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Later among the works it cites.
I. Armeni, Z.-Y. He, J. Gwak, A. R. Zamir, M. Fischer, J. Malik, and S. Savarese, “3d scene graph: A structure for unified semantics, 3d space, and camera,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Later among the works it cites.
M. Edmonds, F. Gao, H. Liu, X. Xie, S. Qi, B. Rothrock, Y. Zhu, Y. N. Wu, H. Lu, and S.-C. Zhu, “A tale of two explanations: Enhancing human trust by explaining robot behavior,” Science Robotics
2019
Later among the works it cites.
H. Liu, C. Zhang, Y. Zhu, C. Jiang, and S.-C. Zhu, “Mirroring without overimitation: Learning functionally equivalent manipulation actions,” in Proceedings of AAAI Conference on Artificial Intelligence (AAAI)
2019
Later among the works it cites.
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick, “Detectron2.” https://github.com/facebookresearch/detectron2 , 2019
2019
Later among the works it cites.
D.-C. Hoang, A. J. Lilienthal, and T. Stoyanov, “Panoptic 3d mapping and object pose estimation using adaptively weighted semantic information,” Robotics and Automation Letters (RA-L)
2020
Later among the works it cites.
F. Xia, W. B. Shen, C. Li, P. Kasimbeg, M. E. Tchapmi, A. Toshev, R. Martín-Martín, and S. Savarese, “Interactive gibson benchmark: A benchmark for interactive navigation in cluttered environments,” Robotics and Automation Letters (RA-L)
2020
Later among the works it cites.
F. Xiang, Y. Qin, K. Mo, Y. Xia, H. Zhu, F. Liu, M. Liu, H. Jiang, Y. Yuan, H. Wang, et al
2020
Later among the works it cites.
2020
Later among the works it cites.
K. Wada, E. Sucar, S. James, D. Lenton, and A. J. Davison, “Morefusion: Multi-object reasoning for 6d pose estimation from volumetric fusion,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2020
Later among the works it cites.
Z. Sui, H. Chang, N. Xu, and O. Chadwicke Jenkins, “Geofusion: Geometric consistency informed scene estimation in dense clutter,” Robotics and Automation Letters (RA-L)
2020
Later among the works it cites.
A. Avetisyan, T. Khanova, C. Choy, D. Dash, A. Dai, and M. Nießner, “Scenecad: Predicting object alignments and layouts in rgb-d scans,” in The European Conference on Computer Vision (ECCV)
2020
Later among the works it cites.
J. Wald, H. Dhamo, N. Navab, and F. Tombari, “Learning 3d semantic scene graphs from 3d indoor reconstructions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2020
Later among the works it cites.
A. Rosinol, A. Gupta, M. Abate, J. Shi, and L. Carlone, “3D dynamic scene graphs: Actionable spatial perception with places, objects, and humans,” in Proceedings of Robotics: Science and Systems (RSS)
2020
Later among the works it cites.
S. Qi, B. Jia, S. Huang, P. Wei, and S.-C. Zhu, “A generalized earley parser for human activity parsing and prediction,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)
2020
Later among the works it cites.
Z. Zhang, Y. Zhu, and S.-C. Zhu, “Graph-based hierarchical knowledge representation for robot task transfer from virtual to physical world,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2020
Later among the works it cites.
T. Yuan, H. Liu, L. Fan, Z. Zheng, T. Gao, Y. Zhu, and S.-C. Zhu, “Joint inference of states, robot knowledge, and human (false-) beliefs,” in Proceedings of International Conference on Intelligent Robots and Systems (IROS)
2020
Later among the works it cites.