Fetching the paper…
Reading the bibliography…
3D open-vocabulary scene graph methods are a promising map representation for embodied agents, however many current approaches are computationally expensive.
Seeded region growing
R. Adams and L. Bischof · 1994
Earlier work this paper cites.
3D is here: Point Cloud Library (PCL)
Radu Bogdan Rusu and Steve Cousins · 2011
Earlier work this paper cites.
Contextually guided semantic labeling and search for three-dimensional point clouds
Abhishek Anand, Hema Swetha Koppula, Thorsten Joachims, and Ashutosh Saxena · 2013
Earlier work this paper cites.
3D Semantic Parsing of Large-Scale Indoor Spaces
Iro Armeni, Ozan Sener, Amir R. Zamir, Helen Jiang, Ioannis Brilakis, Martin Fischer, and Silvio Savarese · 2016
Earlier work this paper cites.
A Review of Point Clouds Segmentation and Classification Algorithms
E. Grilli, F. Menna, and F. Remondino · 2017
Earlier work this paper cites.
FutureMapping: The Computational Structure of Spatial AI Systems, 2018
Andrew J. Davison · 2018
Earlier work this paper cites.
3D Semantic Segmentation with Submanifold Sparse Convolutional Networks
Benjamin Graham, Martin Engelcke, and Laurens van der Maaten · 2018
Earlier work this paper cites.
Pointwise Convolutional Neural Networks
Binh-Son Hua, Minh-Khoi Tran, and Sai-Kit Yeung · 2018
Earlier work this paper cites.
4D Spatio-Temporal ConvNets: Minkowski Convolutional Neural Networks
Christopher Choy, JunYoung Gwak, and Silvio Savarese · 2019
Earlier work this paper cites.
Deep High-Resolution Representation Learning for Human Pose Estimation
Ke Sun, Bin Xiao, Dong Liu, and Jingdong Wang · 2019
Earlier work this paper cites.
3D-MPA: Multi Proposal Aggregation for 3D Semantic Instance Segmentation
Francis Engelmann, Martin Bokeloh, Alireza Fathi, Bastian Leibe, and Matthias Nießner · 2020
Earlier work this paper cites.
Bayesian Spatial Kernel Smoothing for Scalable Dense Semantic Mapping
Lu Gan, Ray Zhang, Jessy W. Grizzle, Ryan M. Eustice, and Maani Ghaffari · 2020
Earlier work this paper cites.
OccuSeg: Occupancy-Aware 3D Instance Segmentation
Lei Han, Tian Zheng, Lan Xu, and Lu Fang · 2020
Earlier work this paper cites.
PointGroup: Dual-Set Point Grouping for 3D Instance Segmentation
Li Jiang, Hengshuang Zhao, Shaoshuai Shi, Shu Liu, Chi-Wing Fu, and Jiaya Jia · 2020
Earlier work this paper cites.
Kimera: an Open-Source Library for Real-Time Metric-Semantic Localization and Mapping
Antoni Rosinol, Marcus Abate, Yun Chang, and Luca Carlone · 2020
Earlier work this paper cites.
VMNet: Voxel-Mesh Network for Geodesic-Aware 3D Semantic Segmentation
Zeyu Hu, Xuyang Bai, Jiaxiang Shang, Runze Zhang, Jiayu Dong, Xin Wang, Guangyuan Sun, Hongbo Fu, and Chiew-Lan Tai · 2021
Cited alongside, same era.
Learning Transferable Visual Models From Natural Language Supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Cited alongside, same era.
S-Graphs+: Real-time Localization and Mapping leveraging Hierarchical Representations, 2022
Hriday Bavle, Jose Luis Sanchez-Lopez, Muhammad Shaheer, Javier Civera, and Holger Voos · 2022
Cited alongside, same era.
Hydra: A Real-time Spatial Perception Engine for 3D Scene Graph Construction and Optimization
Nathan Hughes, Yun Chang, and Luca Carlone · 2022
Cited alongside, same era.
Deep Learning-Based 3D Instance and Semantic Segmentation: A Review
Siddiqui Muhammad Yasir and Hyunsik Ahn · 2022
Active Open-Vocabulary Recognition: Let Intelligent Moving Mitigate CLIP Limitations
Lei Fan, Jianxiong Zhou, Xiaoying Xing, and Ying Wu · 2024
Closest in time.
ConceptGraphs: Open-Vocabulary 3D Scene Graphs for Perception and Planning
Qiao Gu, Alihusein Kuwajerwala, Sacha Morin, Krishna Murthy Jatavallabhula, Bipasha Sen, Aditya Agarwal, Corban Rivera, William Paul, Kirsty Ellis, Rama Chellappa, Chuang Gan, Celso Miguel de Melo, Joshua B. Tenenbaum, Antonio Torralba, Florian Shkurti, and Liam Paull · 2024
Closest in time.
Foundations of spatial perception for robotics: Hierarchical representations and real-time systems
N. Hughes, Y. Chang, S. Hu, R. Talak, R. Abdulhai, J. Strader, and L. Carlone · 2024
Closest in time.
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation, 2024
Hanxiao Jiang, Binghao Huang, Ruihai Wu, Zhuoran Li, Shubham Garg, Hooshang Nayyeri, Shenlong Wang, and Yunzhu Li · 2024
Closest in time.
Language-EXtended Indoor SLAM (LEXIS): A Versatile System for Real-time Visual Scene Understanding
Christina Kassab, Matias Mattamala, Lintong Zhang, and Maurice Fallon · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
SoftGroup for 3D Instance Segmentation on Point Clouds
Thang Vu, Kookhoi Kim, Tung Minh Luu, Xuan Thanh Nguyen, and Chang-Dong Yoo · 2022
Cited alongside, same era.
Strategies for large scale elastic and semantic LiDAR reconstruction
Yiduo Wang, Milad Ramezani, Matias Mattamala, Sundara Tejaswi Digumarti, and Maurice Fallon · 2022
Cited alongside, same era.
Visual perception in the human brain: How the brain perceives and understands real-world scenes, 2023
Clemens G. Bartnik and Iris I. A. Groen · 2023
Cited alongside, same era.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick · 2023
Cited alongside, same era.
Mask3D: Mask Transformer for 3D Semantic Instance Segmentation
Jonas Schult, Francis Engelmann, Alexander Hermans, Or Litany, Siyu Tang, and Bastian Leibe · 2023
Cited alongside, same era.
What does CLIP know about a red circle? Visual prompt engineering for VLMs
Aleksandar Shtedritski, Christian Rupprecht, and Andrea Vedaldi · 2023
Cited alongside, same era.
OpenMask3D: Open-Vocabulary 3D Instance Segmentation
Ayça Takmaz, Elisabetta Fedele, Robert W. Sumner, Marc Pollefeys, Federico Tombari, and Francis Engelmann · 2023
Cited alongside, same era.
Closest in time.
Open3DSG: Open-Vocabulary 3D Scene Graphs from Point Clouds with Queryable Objects and Open-Set Relationships
Sebastian Koch, Narunas Vaskevicius, Mirco Colosi, Pedro Hermosilla, and Timo Ropinski · 2024
Closest in time.
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images
Han-Hung Lee, Yiming Zhang, and Angel X. Chang · 2024
Closest in time.
Clio: Real-time Task-Driven Open-Set 3D Scene Graphs
Dominic Maggio, Yun Chang, Nathan Hughes, Matthew Trang, Dan Griffith, Carlyn Dougherty, Eric Cristofalo, Lukas Schmid, and Luca Carlone · 2024
Closest in time.
GPT-4 Technical Report, 2024
OpenAI · 2024
Closest in time.
SAM 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloé Rolland, Laura Gustafson, Eric Mintun, Junting Pan, Kalyan Vasudev Alwala, Nicolas Carion, Chao-Yuan Wu, Ross B. Girshick, Piotr Dollár, and Christoph Feichtenhofer · 2024
Closest in time.
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
Dan Song, Xinwei Fu, Weizhi Nie, Wenhui Li, Lanjun Wang, You Yang, and Anan Liu · 2024
Closest in time.
LABELMAKER: Automatic Semantic Label Generation from RGB-D Trajectories
Silvan Weder, Hermann Blum, Francis Engelmann, and Marc Pollefeys · 2024
Closest in time.
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
Abdelrhman Werby, Chenguang Huang, Martin Büchner, Abhinav Valada, and Wolfram Burgard · 2024
Closest in time.
SAI3D: Segment Any Instance in 3D Scenes, 2024
Yingda Yin, Yuzheng Liu, Yang Xiao, Daniel Cohen-Or, Jingwei Huang, and Baoquan Chen · 2024
Closest in time.