Fetching the paper…
Reading the bibliography…
Recently, groundbreaking results have been presented on open-vocabulary semantic image segmentation.
B. Zhou, H. Zhao, X. Puig, S. Fidler, A. Barriuso, and A. Torralba, “Scene Parsing through ADE20K Dataset,” in IEEE CVPR
2017
Earlier work this paper cites.
M. Grinvald, F. Furrer, T. Novkovic, J. J. Chung, C. Cadena, R. Siegwart, and J. Nieto, “Volumetric Instance-Aware Semantic Mapping and 3D Object Discovery,” IEEE Robotics and Automation Letters
2019
Earlier work this paper cites.
M. Strecke and J. Stuckler, “EM-Fusion: Dynamic Object-Level SLAM With Probabilistic Data Association,” in IEEE/CVF ICCV
2019
Earlier work this paper cites.
G. Narita, T. Seno, T. Ishikawa, and Y. Kaji, “PanopticFusion: Online Volumetric Semantic Mapping at the Level of Stuff and Things,” in IROS
2019
Earlier work this paper cites.
B. Xu, W. Li, D. Tzoumanikas, M. Bloesch, A. Davison, and S. Leutenegger, “MID-Fusion: Octree-based Object-Level Multi-Instance Dynamic SLAM,” in IEEE ICRA
2019
Earlier work this paper cites.
I. Armeni, Z.-Y. He, J. Gwak, A. R. Zamir, M. Fischer, J. Malik, and S. Savarese, “3D Scene Graph: A Structure for Unified Semantics, 3D Space, and Camera,” in IEEE/CVF ICCV
2019
Earlier work this paper cites.
A. Rosinol, M. Abate, Y. Chang, and L. Carlone, “Kimera: an Open-Source Library for Real-Time Metric-Semantic Localization and Mapping,” in IEEE ICRA
2020
Earlier work this paper cites.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis,” in ECCV
2020
Earlier work this paper cites.
A. Zareian, K. D. Rosa, D. H. Hu, and S.-F. Chang, “Open-Vocabulary Object Detection Using Captions,” in IEEE/CVF CVPR
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al
2021
Earlier work this paper cites.
V. Blukis, R. Knepper, and Y. Artzi, “Few-shot Object Grounding and Mapping for Natural Language Robot Instruction Following,” in Conference on Robot Learning
2021
Earlier work this paper cites.
M. Grinvald, F. Tombari, R. Siegwart, and J. Nieto, “TSDF++: A Multi-Object Formulation for Dynamic Object Tracking and Reconstruction,” in IEEE ICRA
2021
Earlier work this paper cites.
S.-C. Wu, J. Wald, K. Tateno, N. Navab, and F. Tombari, “SceneGraphFusion: Incremental 3D Scene Graph Prediction from RGB-D Sequences,” in IEEE/CVF CVPR
2021
Earlier work this paper cites.
S. Zhi, T. Laidlow, S. Leutenegger, and A. J. Davison, “In-Place Scene Labelling and Understanding with Implicit Scene Representation,” in IEEE/CVF International Conference on Computer Vision
2021
Earlier work this paper cites.
E. Sucar, S. Liu, J. Ortiz, and A. J. Davison, “iMAP: Implicit Mapping and Positioning in Real-Time,” in IEEE/CVF ICCV
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Peng, K. Genova, C. Jiang, A. Tagliasacchi, M. Pollefeys, T. Funkhouser, et al
2022
Cited alongside, same era.
S. Zhi, E. Sucar, A. Mouton, I. Haughton, T. Laidlow, and A. J. Davison, “iLabel: Revealing Objects in Neural Fields,” IEEE Robotics and Automation Letters
2022
Cited alongside, same era.
2022
Cited alongside, same era.
G. Ghiasi, X. Gu, Y. Cui, and T.-Y. Lin, “Scaling Open-Vocabulary Image Segmentation with Image-Level Labels,” in ECCV
2022
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
H. Zhang, P. Zhang, X. Hu, Y. Chen, L. Li, X. Dai, L. Wang, L. Yuan, J. Hwang, and J. Gao, “GLIPv2: Unifying Localization and Vision-Language Understanding,” in NeurIPS
2022
Cited alongside, same era.
X. Zou, Z.-Y. Dou, J. Yang, Z. Gan, L. Li, C. Li, X. Dai, H. Behl, J. Wang, L. Yuan, et al
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
L. Schmid, J. Delmerico, J. L. Schönberger, J. Nieto, M. Pollefeys, R. Siegwart, and C. Cadena, “Panoptic Multi-TSDFs: a Flexible Representation for Online Multi-resolution Volumetric Mapping and Long-term Dynamic Scene Consistency,” in IEEE ICRA
2022
Later among the works it cites.
N. Hughes, Y. Chang, and L. Carlone, “Hydra: A Real-time Spatial Perception System for 3D Scene Graph Construction and Optimization,” in Robotics: Science and Systems
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
A. Kundu, K. Genova, X. Yin, A. Fathi, C. Pantofaru, L. J. Guibas, A. Tagliasacchi, F. Dellaert, and T. Funkhouser, “Panoptic Neural Fields: A Semantic Object-Aware Neural Scene Representation,” in IEEE/CVF CVPR
2022
Later among the works it cites.
S. Kobayashi, E. Matsumoto, and V. Sitzmann, “Decomposing NeRF for Editing via Feature Field Distillation,” in NeurIPS
2022
Later among the works it cites.
V. Tschernezki, I. L. D. Larlus, and A. Vedaldi, “Neural Feature Fusion Fields: 3D Distillation of Self-Supervised 2D Image Representations,” in Conference on 3D Vision
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Zhu, S. Peng, V. Larsson, W. Xu, H. Bao, Z. Cui, M. R. Oswald, and M. Pollefeys, “NICE-SLAM: Neural Implicit Scalable Encoding for SLAM,” in IEEE/CVF CVPR
2022
Later among the works it cites.
2023
Closest in time.