Fetching the paper…
Reading the bibliography…
Mobile exploration is a longstanding challenge in robotics, yet current methods primarily focus on active perception instead of active interaction, limiting the robot's ability to interact with and fully explore its environment.
H. Choset et al. , “Sensor-based exploration: Incremental construction of the hierarchical generalized voronoi graph,” IJRR , 2000
2000
Earlier work this paper cites.
Y. Liu and G. Nejat, “Robotic urban search and rescue: A survey from the control perspective,” Journal of Intelligent & Robotic Systems , vol. 72, pp. 147–165, 2013
2013
Earlier work this paper cites.
D. Misra et al. , “Mapping instructions to actions in 3d environments with visual goal prediction,” in EMNLP , 2018
2018
Earlier work this paper cites.
Y. Jiang, N. Walker, J. Hart, and P. Stone, “Open-world reasoning for service robots,” in international conference on automated planning and scheduling , 2019
2019
Earlier work this paper cites.
F. Niroui et al. , “Deep reinforcement learning robot for search and rescue applications: Exploration in unknown cluttered environments,” RA-L , 2019
2019
Earlier work this paper cites.
T. Chen, S. Gupta, and A. Gupta, “Learning exploration policies for navigation,” in International Conference on Learning Representations , 2019
2019
Earlier work this paper cites.
M. Labbé and F. Michaud, “Rtab-map as an open-source lidar and visual simultaneous localization and mapping library for large-scale and long-term online operation,” Journal of field robotics , 2019
2019
Earlier work this paper cites.
C. R. Garrett, et al. , “Online replanning in belief space for partially observable task and motion problems,” in ICRA , 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
E. Rosen, et al. , “Building plannable representations with mixed reality,” in IROS , 2020
2020
Earlier work this paper cites.
C. Cao, H. Zhu, H. Choset, and J. Zhang, “Tare: A hierarchical framework for efficiently exploring complex 3d environments.” in RSS , 2021
2021
Earlier work this paper cites.
F. Xia et al. , “Relmogen: Integrating motion generation in reinforcement learning for mobile manipulation,” in ICRA , 2021
2021
Earlier work this paper cites.
K. Ehsani, et al. , “Manipulathor: A framework for visual object manipulation,” in CVPR , 2021
2021
Earlier work this paper cites.
A. Radford et al. , “Learning transferable visual models from natural language supervision,” in the 38th International Conference on Machine Learning , 2021
2021
Earlier work this paper cites.
K. Zheng, R. Chitnis, Y. Sung, G. Konidaris, and S. Tellex, “Towards optimal correlational object search,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 7313–7319
2022
Cited alongside, same era.
H. Ha and S. Song, “Semantic abstraction: Open-world 3D scene understanding from 2D vision-language models,” in Conference on Robot Learning , 2022
2022
Cited alongside, same era.
R. C. Quesada and Y. Demiris, “Proactive robot assistance: Affordance-aware augmented reality user interfaces,” IEEE Robotics & Automation Magazine , 2022
2022
Cited alongside, same era.
2023
Cited alongside, same era.
A. Kirillov et al. , “Segment anything,” in 2023 IEEE/CVF International Conference on Computer Vision (ICCV) , 2023
J. Achiam et al. , “Gpt-4 technical report,” arXiv preprint arXiv:2303.08774 , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Jiang et al. , “Roboexp: Action-conditioned scene graph via interactive exploration for robotic manipulation,” in 8th Conference on Robot Learning , 2024
2024
Later among the works it cites.
D. Honerkamp et al. , “Language-grounded dynamic scene graphs for interactive object search with mobile manipulation,” Robotics and Automation Letters , 2024
2024
Later among the works it cites.
A. Werby et al. , “Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation,” in Robotics: Science and Systems , 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
K. Zheng, A. Paul, and S. Tellex, “A system for generalized 3d multi-object search,” in IEEE International Conference on Robotics and Automation (ICRA) , 2023
2023
Cited alongside, same era.
F. Schmalstieg, D. Honerkamp, T. Welschehold, and A. Valada, “Learning hierarchical interactive multi-object search for mobile manipulation,” IEEE Robotics and Automation Letters , 2023
2023
Cited alongside, same era.
K. Jatavallabhula et al. , “Conceptfusion: Open-set multimodal 3d mapping,” Robotics: Science and Systems (RSS) , 2023
2023
Cited alongside, same era.
S. Peng et al. , “Openscene: 3d scene understanding with open vocabularies,” in IEEE/CVF conference on computer vision and pattern recognition , 2023
2023
Cited alongside, same era.
E. Rosen et al. , “Synthesizing navigation abstractions for planning with portable manipulation skills,” in Conference on Robot Learning , 2023
2023
Cited alongside, same era.
W. Huang, C. Wang, R. Zhang, Y. Li, J. Wu, and L. Fei-Fei, “Voxposer: Composable 3d value maps for robotic manipulation with language models,” in CoRL , 2023
2023
Cited alongside, same era.
D. Driess et al. , “PaLM-e: An embodied multimodal language model,” in ICML , 2023
2023
Cited alongside, same era.
2024
Later among the works it cites.
T. Cheng, L. Song, Y. Ge, W. Liu, X. Wang, and Y. Shan, “Yolo-world: Real-time open-vocabulary object detection,” in CVPR , 2024
2024
Later among the works it cites.
N. Yokoyama, S. Ha, D. Batra, J. Wang, and B. Bucher, “Vlfm: Vision-language frontier maps for zero-shot semantic navigation,” in ICRA , 2024
2024
Later among the works it cites.
D. Maggio, Y. Chang et al. , “Clio: Real-time task-driven open-set 3d scene graphs,” IEEE Robotics and Automation Letters , 2024
2024
Later among the works it cites.
K. Yamazaki, et al. , “Open-fusion: Real-time open-vocabulary 3d mapping and queryable scene representation,” in IEEE International Conference on Robotics and Automation (ICRA) , 2024
2024
Later among the works it cites.
M. Oquab et al. , “DINOv2: Learning robust visual features without supervision,” Transactions on Machine Learning Research , 2024
2024
Later among the works it cites.
Q. Gu et al. , “Conceptgraphs: Open-vocabulary 3d scene graphs for perception and planning,” in IEEE International Conference on Robotics and Automation , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
J. Yang et al. , “Llm-grounder: Open-vocabulary 3d visual grounding with large language model as an agent,” in ICRA , 2024
2024
Later among the works it cites.
H. Liu, C. Li, Y. Li, and Y. J. Lee, “Improved baselines with visual instruction tuning,” in CVPR , 2024
2024
Later among the works it cites.