Fetching the paper…
Reading the bibliography…
Semantic reasoning and dynamic planning capabilities are crucial for an autonomous agent to perform complex navigation tasks in unknown environments.
Language models are few-shot learners
Brown, T.; et al. 2020 · 1901
Earlier work this paper cites.
Benchmarking classic and learned navigation in complex 3d environments
Mishkin, D.; et al. 2019 · 1901
Earlier work this paper cites.
Dd-ppo: Learning near-perfect pointgoal navigators from 2.5 billion frames
Wijmans, E.; et al. 2019 · 1911
Earlier work this paper cites.
Automatic evaluation of information ordering: Kendall’s tau
Lapata, M. 2006 · 2006
Earlier work this paper cites.
Allenact: A framework for embodied ai research
Weihs, L.; et al. 2020 · 2008
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S.; et al. 2011 · 2011
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, K.; et al. 2014 · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; et al. 2016 · 2016
Earlier work this paper cites.
Ai2-thor: An interactive 3d environment for visual ai
Kolve, E.; et al. 2017 · 2017
Earlier work this paper cites.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Zhu, Y.; et al. 2017 · 2017
Earlier work this paper cites.
On evaluation of embodied navigation agents
Anderson, P.; et al. 2018 · 2018
Earlier work this paper cites.
Group normalization
Wu, Y.; et al. 2018 · 2018
Earlier work this paper cites.
3d scene graph: A structure for unified semantics, 3d space, and camera
Armeni, I.; et al. 2019 · 2019
Earlier work this paper cites.
3-D scene graph: A sparse and semantic representation of physical environments for intelligent agents
Kim, U.; et al. 2019 · 2019
Earlier work this paper cites.
Habitat: A Platform for Embodied AI Research
Savva, M.; et al. 2019 · 2019
Earlier work this paper cites.
Object goal navigation using goal-oriented semantic exploration
Chaplot, D.; et al. 2020 · 2020
Cited alongside, same era.
Learning 3D semantic scene graphs from 3D indoor reconstructions
Wald, J.; et al. 2020 · 2020
Cited alongside, same era.
MultiON: Benchmarking Semantic Map Memory using Multi-Object Navigation
Wani, S.; et al. 2020 · 2020
Cited alongside, same era.
Kimera: From slam to spatial perception with 3d dynamic scene graphs
Rosinol, A.; et al. 2021 · 2021
Cited alongside, same era.
Habitat 2.0: Training home assistants to rearrange their habitat
Szot, A.; et al. 2021 · 2021
Cited alongside, same era.
SceneGraphFusion: Incremental 3D scene graph prediction from RGB-D sequences
Wu, S.; et al. 2021 · 2021
Cited alongside, same era.
Habitat-web: Learning embodied object-search strategies from human demonstrations at scale
Ramrakhya, R.; et al. 2022 · 2022
Later among the works it cites.
Llm-planner: Few-shot grounded planning for embodied agents with large language models
Song, C. H.; et al. 2022 · 2022
Later among the works it cites.
Palm-e: An embodied multimodal language model
Driess, D.; et al. 2023 · 2023
Closest in time.
Sequence-Agnostic Multi-Object Navigation
Gireesh, N.; et al. 2023 · 2023
Closest in time.
LLM+P: Empowering large language models with optimal planning proficiency
Liu, B.; et al. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Do as i can, not as i say: Grounding language in robotic affordances
Ahn, M.; et al. 2022 · 2022
Cited alongside, same era.
Learning Active Camera for Multi-Object Navigation
Chen, P.; et al. 2022 · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models
Chung, H. W.; et al. 2022 · 2022
Cited alongside, same era.
ProcTHOR: Large-Scale Embodied AI Using Procedural Generation
Deitke, M.; et al. 2022 · 2022
Cited alongside, same era.
Inner monologue: Embodied reasoning through planning with language models
Huang, W.; et al. 2022 · 2022
Cited alongside, same era.
Hydra: A real-time spatial perception engine for 3d scene graph construction and optimization
Hughes, N.; et al. 2022 · 2022
Cited alongside, same era.
Closest in time.
Multi-Object Navigation with dynamically learned neural implicit representations
Marza, P.; et al. 2023 · 2023
Closest in time.
Peng, B.; et al. 2023 · 2023
Closest in time.
Pirlnav: Pretraining with imitation and rl finetuning for objectnav
Ramrakhya, R.; et al. 2023 · 2023
Closest in time.
SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Task Planning
Rana, K.; et al. 2023 · 2023
Closest in time.
ViNT: A Foundation Model for Visual Navigation
Shah, D.; et al. 2023 · 2023
Closest in time.
Progprompt: Generating situated robot task plans using large language models
Singh, I.; et al. 2023 · 2023
Closest in time.
Leveraging Large Language Models for Visual Target Navigation
Yu, B.; et al. 2023 · 2023
Closest in time.
Multi-Object Navigation Using Potential Target Position Policy Function
Zeng, H.; et al. 2023 · 2023
Closest in time.
A survey of large language models
Zhao, W. X.; et al. 2023 · 2023
Closest in time.