Fetching the paper…
Reading the bibliography…
We introduce the task of open-vocabulary 3D instance segmentation.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
Martin A. Fischler and Robert C. Bolles · 1981
Earlier work this paper cites.
A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise
Martin Ester, Hans-Peter Kriegel, Jörg Sander, and Xiaowei Xu · 1996
Earlier work this paper cites.
Contextually Guided Semantic Labeling and Search for 3D Point Clouds
Abhishek Anand, Hema Swetha Koppula, Thorsten Joachims, and Ashutosh Saxena · 2011
Earlier work this paper cites.
Semantic Labeling of 3D Point Clouds for Indoor Scenes
Hema Koppula, Abhishek Anand, Thorsten Joachims, and Ashutosh Saxena · 2011
Earlier work this paper cites.
Simplified Markov Random Fields for Efficient Semantic Labeling of 3D Point Clouds
Yan Lu and Christopher Rasmussen · 2012
Earlier work this paper cites.
Fully Convolutional Networks for Semantic Segmentation
Jonathan Long, Evan Shelhamer, and Trevor Darrell · 2015
Earlier work this paper cites.
An Efficient Scene Semantic Labeling Approach for 3D Point Cloud
Tianyi Wang, Jian Li, and Xiangjing An · 2015
Earlier work this paper cites.
Fast Semantic Segmentation of 3D Point Clouds using a Dense CRF with Learned Parameters
Daniel Wolf, Johann Prankl, and Markus Vincze · 2015
Earlier work this paper cites.
3D Semantic Parsing of Large-Scale Indoor Spaces
Iro Armeni, Ozan Sener, Amir R. Zamir, Helen Jiang, Ioannis Brilakis, Martin Fischer, and Silvio Savarese · 2016
Earlier work this paper cites.
Point Cloud Labeling Using 3D Convolutional Neural Network
Jing Huang and Suya You · 2016
Earlier work this paper cites.
ScanNet: Richly-Annotated 3D Reconstructions of Indoor Scenes
Angela Dai, Angel X Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner · 2017
Earlier work this paper cites.
Exploring Spatial Context for 3D Semantic Segmentation of Point Clouds
Francis Engelmann, Theodora Kontogianni, Alexander Hermans, and Bastian Leibe · 2017
Earlier work this paper cites.
SEGCloud: Semantic Segmentation of 3D Point Clouds
Lyne P. Tchapmi, Christopher B. Choy, Iro Armeni, JunYoung Gwak, and Silvio Savarese · 2017
Earlier work this paper cites.
Point Convolutional Neural Networks by Extension Operators
Matan Atzmon, Haggai Maron, and Yaron Lipman · 2018
Earlier work this paper cites.
Know What Your Neighbors Do: 3D Semantic Segmentation of Point Clouds
Francis Engelmann, Theodora Kontogianni, Jonas Schult, and Bastian Leibe · 2018
Earlier work this paper cites.
3D Semantic Segmentation with Submanifold Sparse Convolutional Networks
Benjamin Graham, Martin Engelcke, and Laurens van der Maaten · 2018
Earlier work this paper cites.
Point-wise Convolutional Neural Network
Binh-Son Hua, Minh-Khoi Tran, and Sai-Kit Yeung · 2018
Earlier work this paper cites.
Large-scale Point Cloud Semantic Segmentation with Superpoint Graphs
Loic Landrieu and Martin Simonovsky · 2018
Earlier work this paper cites.
PointCNN: Convolution on X-transformed Points
Yangyan Li, Rui Bu, Mingchao Sun, Wei Wu, Xinhan Di, and Baoquan Chen · 2018
Earlier work this paper cites.
SGPN: Similarity Group Proposal Network for 3D Point Cloud Instance Segmentation
Weiyue Wang, Ronald Yu, Qiangui Huang, and Ulrich Neumann · 2018
Earlier work this paper cites.
SpiderCNN: Deep Learning on Point Sets with Parameterized Convolutional Filters
Yifan Xu, Tianqi Fan, Mingye Xu, Long Zeng, and Yu Qiao · 2018
Earlier work this paper cites.
4D Spatio-Temporal ConvNets: Minkowski Convolutional Neural Networks
Christopher Choy, JunYoung Gwak, and Silvio Savarese · 2019
Earlier work this paper cites.
3D-BEVIS: Birds-Eye-View Instance Segmentation
Cathrin Elich, Francis Engelmann, Theodora Kontogianni, and Bastian Leibe · 2019
Earlier work this paper cites.
3D-SIS: 3D Semantic Instance Segmentation of RGB-D Scans
Ji Hou, Angela Dai, and Matthias Nießner · 2019
Earlier work this paper cites.
3D Instance Segmentation via Multi-task Metric Learning
Jean Lahoud, Bernard Ghanem, Marc Pollefeys, and Martin R. Oswald · 2019
Earlier work this paper cites.
The Replica dataset: A digital replica of indoor spaces
Julian Straub, Thomas Whelan, Lingni Ma, Yufan Chen, Erik Wijmans, Simon Green, Jakob J. Engel, Raul Mur-Artal, Carl Ren, Shobhit Verma, Anton Clarkson, Mingfei Yan, Brian Budge, Yajie Yan, Xiaqing Pan, June Yon, Yuyang Zou, Kimberly Leon, Nigel Carter, Jesus Briales, Tyler Gillingham, Elias Mueggler, Luis Pesqueira, Manolis Savva, Dhruv Batra, Hauke M. Strasdat, Renzo De Nardi, Michael Goesele, Steven Lovegrove, and Richard Newcombe · 2019
Earlier work this paper cites.
KPConv: Flexible and Deformable Convolution for Point Clouds
Hugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui, François Goulette, and Leonidas J. Guibas · 2019
Cited alongside, same era.
Learning Object Bounding Boxes for 3D Instance Segmentation on Point Clouds
Bo Yang, Jianan Wang, Ronald Clark, Qingyong Hu, Sen Wang, Andrew Markham, and Niki Trigoni · 2019
Cited alongside, same era.
3D-MPA: Multi Proposal Aggregation for 3D Semantic Instance Segmentation
Francis Engelmann, Martin Bokeloh, Alireza Fathi, Bastian Leibe, and Matthias Nießner · 2020
Cited alongside, same era.
OccuSeg: Occupancy-aware 3D Instance Segmentation
Lei Han, Tian Zheng, Lan Xu, and Lu Fang · 2020
Cited alongside, same era.
PointGroup: Dual-Set Point Grouping for 3D Instance Segmentation
Li Jiang, Hengshuang Zhao, Shaoshuai Shi, Shu Liu, Chi-Wing Fu, and Jiaya Jia · 2020
Cited alongside, same era.
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis
Open-vocabulary Semantic Segmentation with Frozen Vision-Language Models
Chao Ma, Yu-Hao Yang, Yanfeng Wang, Ya Zhang, and Weidi Xie · 2022
Later among the works it cites.
DenseCLIP: Language-Guided Dense Prediction with Context-Aware Prompting
Yongming Rao, Wenliang Zhao, Guangyi Chen, Yansong Tang, Zheng Zhu, Guan Huang, Jie Zhou, and Jiwen Lu · 2022
Later among the works it cites.
Language-Grounded Indoor 3D Semantic Segmentation in the Wild
David Rozenberszki, Or Litany, and Angela Dai · 2022
Later among the works it cites.
Robotic Navigation with Large Pre-Trained Models of Language, Vision, and Action
Dhruv Shah, Blazej Osinski, Brian Ichter, and Sergey Levine · 2022
Later among the works it cites.
SoftGroup for 3D Instance Segmentation on 3D Point Clouds
Thang Vu, Kookhoi Kim, Tung M. Luu, Xuan Thanh Nguyen, and Chang D. Yoo · 2022
Later among the works it cites.
GroupViT: Semantic Segmentation Emerges from Text Supervision
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Cited alongside, same era.
FSS-1000: A 1000-Class Dataset for Few-Shot Segmentation
Tianhan Wei, Xiang Li, Yau Pun Chen, Yu-Wing Tai, and Chi-Keung Tang · 2020
Cited alongside, same era.
Emerging Properties in Self-Supervised Vision Transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Herv’e J’egou, Julien Mairal, Piotr Bojanowski, and Armand Joulin · 2021
Cited alongside, same era.
Scaling Open-Vocabulary Image Segmentation with Image-Level Labels
Golnaz Ghiasi, Xiuye Gu, Yin Cui, and Tsung-Yi Lin · 2021
Cited alongside, same era.
VMNet: Voxel-Mesh Network for Geodesic-Aware 3D Semantic Segmentation
Zeyu Hu, Xuyang Bai, Jiaxiang Shang, Runze Zhang, Jiayu Dong, Xin Wang, Guangyuan Sun, Hongbo Fu, and Chiew Lan Tai · 2021
Cited alongside, same era.
Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc V. Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig · 2021
Cited alongside, same era.
Learning Transferable Visual Models From Natural Language Supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Cited alongside, same era.
Jiarui Xu, Shalini De Mello, Sifei Liu, Wonmin Byeon, Thomas Breuel, Jan Kautz, and Xiaolong Wang · 2022
Later among the works it cites.
CoCa: Contrastive Captioners are Image-Text Foundation Models
Jiahui Yu, Zirui Wang, Vijay Vasudevan, Legg Yeung, Mojtaba Seyedhosseini, and Yonghui Wu · 2022
Later among the works it cites.
Extract Free Dense Labels from CLIP
Chong Zhou, Chen Change Loy, and Bo Dai · 2022
Later among the works it cites.
PLA: Language-Driven Open-Vocabulary 3D Scene Understanding
Runyu Ding, Jihan Yang, Chuhui Xue, Wenqing Zhang, Song Bai, and Xiaojuan Qi · 2023
Closest in time.
ImageBind: One Embedding Space To Bind Them All
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, and Ishan Misra · 2023
Closest in time.
Open-Vocabulary Multi-Label Classification via Multi-modal Knowledge Transfer
Sunan He, Taian Guo, Tao Dai, Ruizhi Qiao, Bo Ren, and Shu-Tao Xia · 2023
Closest in time.
Visual Language Maps for Robot Navigation
Chenguang Huang, Oier Mees, Andy Zeng, and Wolfram Burgard · 2023
Closest in time.
ConceptFusion: Open-Set Multimodal 3D Mapping
Krishna Murthy Jatavallabhula, Alihusein Kuwajerwala, Qiao Gu, Mohd Omama, Tao Chen, Shuang Li, Ganesh Iyer, Soroush Saryazdi, Nikhil Keetha, Ayush Tewari, Joshua B. Tenenbaum, Celso Miguel de Melo, Madhava Krishna, Liam Paull, Florian Shkurti, and Antonio Torralba · 2023
Closest in time.
LERF: Language Embedded Radiance Fields
Justin Kerr, Chung Min Kim, Ken Goldberg, Angjoo Kanazawa, and Matthew Tancik · 2023
Closest in time.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollar, and Ross Girshick · 2023
Closest in time.
Open-Vocabulary Object Detection upon Frozen Vision and Language Models
Weicheng Kuo, Yin Cui, Xiuye Gu, AJ Piergiovanni, and Anelia Angelova · 2023
Closest in time.
Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP
Feng Liang, Bichen Wu, Xiaoliang Dai, Kunpeng Li, Yinan Zhao, Hang Zhang, Peizhao Zhang, Peter Vajda, and Diana Marculescu · 2023
Closest in time.
Feature-Realistic Neural Fusion for Real-Time, Open Set Scene Understanding
Kirill Mazur, Edgar Sucar, and Andrew Davison · 2023
Closest in time.
DINOv2: Learning Robust Visual Features without Supervision
Maxime Oquab, Timothée Darcet, Theo Moutakanni, Huy V. Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, Russell Howes, Po-Yao Huang, Hu Xu, Vasu Sharma, Shang-Wen Li, Wojciech Galuba, Mike Rabbat, Mido Assran, Nicolas Ballas, Gabriel Synnaeve, Ishan Misra, Herve Jegou, Julien Mairal, Patrick Labatut, Armand Joulin, and Piotr Bojanowski · 2023
Closest in time.
OpenScene: 3D Scene Understanding with Open Vocabularies
Songyou Peng, Kyle Genova, Chiyu "Max" Jiang, Andrea Tagliasacchi, Marc Pollefeys, and Thomas Funkhouser · 2023
Closest in time.
Mask3D: Mask Transformer for 3D Semantic Instance Segmentation
Jonas Schult, Francis Engelmann, Alexander Hermans, Or Litany, Siyu Tang, and Bastian Leibe · 2023
Closest in time.
CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory
Nur Muhammad Shafiullah, Chris Paxton, Lerrel Pinto, Soumith Chintala, and Arthur D. Szlam · 2023
Closest in time.
3D Segmentation of Humans in Point Clouds with Synthetic Data
Ayça Takmaz, Jonas Schult, Irem Kaftan, Mertcan Akçay, Bastian Leibe, Robert Sumner, Francis Engelmann, and Siyu Tang · 2023
Closest in time.
ODISE: Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models
Jiarui Xu, Sifei Liu, Arash Vahdat, Wonmin Byeon, Xiaolong Wang, and Shalini De Mello · 2023
Closest in time.
LabelMaker: Automatic Semantic Label Generation from RGB-D Trajectories
Silvan Weder, Hermann Blum, Francis Engelmann, and Marc Pollefeys · 2024
Closest in time.