Fetching the paper…
Reading the bibliography…
In this paper, we present MUVLA, a Map Understanding Vision-Language-Action model tailored for object navigation.
On the estimation of production frontiers: maximum likelihood estimation of the parameters of a discontinuous density function
Aigner, D. J.; Amemiya, T.; and Poirier, D. J. 1976 · 1976
Earlier work this paper cites.
Offline reinforcement learning: Tutorial, review, and perspectives on open problems
Levine, S.; Kumar, A.; Tucker, G.; and Fu, J. 2020 · 2005
Earlier work this paper cites.
Geoadditive expectile regression
Sobotka, F.; and Kneib, T. 2012 · 2012
Earlier work this paper cites.
Gibson env: Real-world perception for embodied agents
Xia, F.; Zamir, A. R.; He, Z.; Sax, A.; Malik, J.; and Savarese, S. 2018 · 2018
Earlier work this paper cites.
Visual semantic navigation using scene priors
Yang, W.; Wang, X.; Farhadi, A.; Gupta, A.; and Mottaghi, R. 2018 · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Kenton, J. D. M.-W. C.; and Toutanova, L. K. 2019 · 2019
Earlier work this paper cites.
Visual representations for semantic target driven navigation
Mousavian, A.; Toshev, A.; Fišer, M.; Košecká, J.; Wahid, A.; and Davidson, J. 2019 · 2019
Earlier work this paper cites.
Semantic visual navigation by watching youtube videos
Chang, M.; Gupta, A.; and Gupta, S. 2020 · 2020
Earlier work this paper cites.
Object goal navigation using goal-oriented semantic exploration
Chaplot, D. S.; Gandhi, D. P.; Gupta, A.; and Salakhutdinov, R. R. 2020 · 2020
Earlier work this paper cites.
Decision transformer: Reinforcement learning via sequence modeling
Chen, L.; Lu, K.; Rajeswaran, A.; Lee, K.; Grover, A.; Laskin, M.; Abbeel, P.; Srinivas, A.; and Mordatch, I. 2021 · 2021
Earlier work this paper cites.
Offline reinforcement learning as one big sequence modeling problem
Janner, M.; Li, Q.; and Levine, S. 2021 · 2021
Earlier work this paper cites.
Thda: Treasure hunt data augmentation for semantic navigation
Maksymets, O.; Cartillier, V.; Gokaslan, A.; Wijmans, E.; Galuba, W.; Lee, S.; and Batra, D. 2021 · 2021
Earlier work this paper cites.
Habitat-matterport 3d dataset (hm3d): 1000 large-scale 3d environments for embodied ai
Ramakrishnan, S. K.; Gokaslan, A.; Wijmans, E.; Maksymets, O.; Clegg, A.; Turner, J.; Undersander, E.; Galuba, W.; Westbury, A.; Chang, A. X.; et al. 2021 · 2021
Earlier work this paper cites.
Auxiliary tasks and exploration enable objectgoal navigation
Ye, J.; Batra, D.; Das, A.; and Wijmans, E. 2021 · 2021
Earlier work this paper cites.
Visual language maps for robot navigation
Huang, C.; Mees, O.; Zeng, A.; and Burgard, W. 2022 · 2022
Earlier work this paper cites.
Zson: Zero-shot object-goal navigation using multimodal goal embeddings
Majumdar, A.; Aggarwal, G.; Devnani, B.; Hoffman, J.; and Batra, D. 2022 · 2022
Earlier work this paper cites.
Behavior transformers: Cloning k k modes with one stone
Shafiullah, N. M.; Cui, Z.; Altanzaya, A. A.; and Pinto, L. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; Zhou, D.; et al. 2022 · 2022
Cited alongside, same era.
Achiam, J.; Adler, S.; Agarwal, S.; Ahmad, L.; Akkaya, I.; Aleman, F. L.; Almeida, D.; Altenschmidt, J.; Altman, S.; Anadkat, S.; et al. 2023 · 2023
Cited alongside, same era.
Bridging zero-shot object navigation and foundation models through pixel-guided navigation skill
Cai, W.; Huang, S.; Cheng, G.; Long, Y.; Gao, P.; Sun, C.; and Dong, H. 2023 · 2023
Cited alongside, same era.
Chang, M.; Gervet, T.; Khanna, M.; Yenamandra, S.; Shah, D.; Min, S. Y.; Shah, K.; Paxton, C.; Gupta, S.; Batra, D.; et al. 2023 · 2023
Li, Q.; Liang, Y.; Wang, Z.; Luo, L.; Chen, X.; Liao, M.; Wei, F.; Deng, Y.; Xu, S.; Zhang, Y.; et al. 2024 · 2024
Later among the works it cites.
InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment
Long, Y.; Cai, W.; Wang, H.; Zhan, G.; and Dong, H. 2024 · 2024
Later among the works it cites.
Policy agnostic rl: Offline rl and online rl fine-tuning of any class and backbone
Mark, M. S.; Gao, T.; Sampaio, G. G.; Srirama, M. K.; Sharma, A.; Finn, C.; and Kumar, A. 2024 · 2024
Later among the works it cites.
Group robust preference optimization in reward-free rlhf
Ramesh, S. S.; Hu, Y.; Chaimalas, I.; Mehta, V.; Sessa, P. G.; Bou Ammar, H.; and Bogunovic, I. 2024 · 2024
Later among the works it cites.
Deepseekmath: Pushing the limits of mathematical reasoning in open language models, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation
Gadre, S. Y.; Wortsman, M.; Ilharco, G.; Schmidt, L.; and Song, S. 2023 · 2023
Cited alongside, same era.
Navigating to objects in the real world
Gervet, T.; Chintala, S.; Batra, D.; Malik, J.; and Chaplot, D. S. 2023 · 2023
Cited alongside, same era.
Navigation with large language models: Semantic guesswork as a heuristic for planning
Shah, D.; Equi, M. R.; Osiński, B.; Xia, F.; Ichter, B.; and Levine, S. 2023 · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; Azhar, F.; et al. 2023 · 2023
Cited alongside, same era.
Q-learning decision transformer: Leveraging dynamic programming for conditional sequence modelling in offline rl
Yamagata, T.; Khalil, A.; and Santos-Rodriguez, R. 2023 · 2023
Cited alongside, same era.
L3mvn: Leveraging large language models for visual target navigation
Yu, B.; Kasaei, H.; and Cao, M. 2023 · 2023
Cited alongside, same era.
Esc: Exploration with soft commonsense constraints for zero-shot object navigation
Zhou, K.; Zheng, K.; Pryor, C.; Shen, Y.; Jin, H.; Getoor, L.; and Wang, X. E. 2023 · 2023
Cited alongside, same era.
Shao, Z.; Wang, P.; Zhu, Q.; Xu, R.; Song, J.; Bi, X.; Zhang, H.; Zhang, M.; Li, Y.; Wu, Y.; et al. 2024 · 2024
Later among the works it cites.
Voronav: Voronoi-based zero-shot object navigation with large language model
Wu, P.; Mu, Y.; Wu, B.; Hou, Y.; Ma, J.; Zhang, S.; and Liu, C. 2024 · 2024
Later among the works it cites.
Vlfm: Vision-language frontier maps for zero-shot semantic navigation
Yokoyama, N.; Ha, S.; Batra, D.; Wang, J.; and Bucher, B. 2024 · 2024
Later among the works it cites.
Fine-tuning large vision-language models as decision-making agents via reinforcement learning
Zhai, S.; Bai, H.; Lin, Z.; Pan, J.; Tong, P.; Zhou, Y.; Suhr, A.; Xie, S.; LeCun, Y.; Ma, Y.; et al. 2024 · 2024
Later among the works it cites.
Towards learning a generalist model for embodied navigation
Zheng, D.; Huang, S.; Zhao, L.; Zhong, Y.; and Wang, L. 2024 · 2024
Later among the works it cites.
Navgpt: Explicit reasoning in vision-and-language navigation with large language models
Zhou, G.; Hong, Y.; and Wu, Q. 2024 · 2024
Later among the works it cites.
Reinformer: Max-return sequence modeling for offline rl
Zhuang, Z.; Peng, D.; Liu, J.; Zhang, Z.; and Wang, D. 2024 · 2024
Later among the works it cites.
OctoNav: Towards Generalist Embodied Navigation
Gao, C.; Jin, L.; Peng, X.; Zhang, J.; Deng, Y.; Li, A.; Wang, H.; and Liu, S. 2025 · 2025
Closest in time.
Cppo: Accelerating the training of group relative policy optimization-based reasoning models
Lin, Z.; Lin, M.; Xie, Y.; and Ji, R. 2025 · 2025
Closest in time.
VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning
Qi, Z.; Zhang, Z.; Yu, Y.; Wang, J.; and Zhao, H. 2025 · 2025
Closest in time.
Does reinforcement learning really incentivize reasoning capacity in llms beyond the base model?
Yue, Y.; Chen, Z.; Lu, R.; Zhao, A.; Wang, Z.; Song, S.; and Huang, G. 2025 · 2025
Closest in time.