Fetching the paper…
Reading the bibliography…
Navigating unfamiliar environments presents significant challenges for household robots, requiring the ability to recognize and reason about novel decoration and layout.
Heusel, Martin, et al. ”Gans trained by a two time-scale update rule converge to a local nash equilibrium.” Advances in neural information processing systems 30 (2017)
2017
Earlier work this paper cites.
Zhu, Yuke, et al. ”Target-driven visual navigation in indoor scenes using deep reinforcement learning.” 2017 IEEE international conference on robotics and automation (ICRA). IEEE, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Zhang, Richard, et al. ”The unreasonable effectiveness of deep features as a perceptual metric.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Mousavian, Arsalan, et al. ”Visual representations for semantic target driven navigation.” 2019 International Conference on Robotics and Automation (ICRA). IEEE, 2019
2019
Earlier work this paper cites.
Chaplot, Devendra Singh, et al. ”Neural topological slam for visual navigation.” Proceedings of the IEEE/CVF conference on computer vision and pattern recognition. 2020
2020
Earlier work this paper cites.
Chaplot, Devendra Singh, et al. ”Object goal navigation using goal-oriented semantic exploration.” Advances in Neural Information Processing Systems 33 (2020)
2020
Earlier work this paper cites.
Brown, Tom B. ”Language models are few-shot learners.” arXiv preprint arXiv:2005.14165 (2020)
2020
Earlier work this paper cites.
Maksymets, Oleksandr, et al. ”Thda: Treasure hunt data augmentation for semantic navigation.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2021
2021
Earlier work this paper cites.
Ye, Joel, et al. ”Auxiliary tasks and exploration enable objectgoal navigation.” Proceedings of the IEEE/CVF international conference on computer vision. 2021
2021
Earlier work this paper cites.
Dhariwal, Prafulla, and Alexander Nichol. ”Diffusion models beat gans on image synthesis.” Advances in neural information processing systems 34 (2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Hahn, Meera, et al. ”No rl, no simulation: Learning to navigate without navigating.” Advances in Neural Information Processing Systems 34 (2021): 26661-26673
2021
Earlier work this paper cites.
Mezghan, Lina, et al. ”Memory-augmented reinforcement learning for image-goal navigation.” 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). IEEE, 2022
2022
Earlier work this paper cites.
Al-Halah, Ziad, Santhosh Kumar Ramakrishnan, and Kristen Grauman. ”Zero experience required: Plug & play modular transfer learning for semantic visual navigation.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2022
2022
Earlier work this paper cites.
Ramakrishnan, Santhosh Kumar, et al. ”Poni: Potential functions for objectgoal navigation with interaction-free learning.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2022
2022
Earlier work this paper cites.
Ramrakhya, Ram, et al. ”Habitat-web: Learning embodied object-search strategies from human demonstrations at scale.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2022
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Majumdar, Arjun, et al. ”Zson: Zero-shot object-goal navigation using multimodal goal embeddings.” Advances in Neural Information Processing Systems 35 (2022): 32340-32352
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2024
Later among the works it cites.
Cai, Wenzhe, et al. ”Bridging zero-shot object navigation and foundation models through pixel-guided navigation skill.” 2024 IEEE International Conference on Robotics and Automation (ICRA). IEEE, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
Du, Yilun, et al. ”Learning universal policies via text-guided video generation.” Advances in Neural Information Processing Systems 36 (2024)
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Brooks, Tim, Aleksander Holynski, and Alexei A. Efros. ”Instructpix2pix: Learning to follow image editing instructions.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Zhou, Kaiwen, et al. ”Esc: Exploration with soft commonsense constraints for zero-shot object navigation.” International Conference on Machine Learning. PMLR, 2023
2023
Cited alongside, same era.
Yu, Bangguo, Hamidreza Kasaei, and Ming Cao. ”L3mvn: Leveraging large language models for visual target navigation.” 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). IEEE, 2023
2023
Cited alongside, same era.
Gadre, Samir Yitzhak, et al. ”Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2023
2023
Cited alongside, same era.
Shah, Dhruv, et al. ”Navigation with large language models: Semantic guesswork as a heuristic for planning.” Conference on Robot Learning. PMLR, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
Liang, Zhixuan, et al. ”Skilldiffuser: Interpretable hierarchical planning via skill abstractions in diffusion-based task execution.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2024
2024
Later among the works it cites.
Li, Xiaoqi, et al. ”Manipllm: Embodied multimodal large language model for object-centric robotic manipulation.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2024
2024
Later among the works it cites.
Liu, Haotian, et al. ”Visual instruction tuning.” Advances in neural information processing systems 36 (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Sun, Xinyu, et al. ”FGPrompt: fine-grained goal prompting for image-goal navigation.” Advances in Neural Information Processing Systems 36 (2024)
2024
Later among the works it cites.
Koh, Jing Yu, Daniel Fried, and Russ R. Salakhutdinov. ”Generating images with multimodal language models.” Advances in Neural Information Processing Systems 36 (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2025
Closest in time.