Fetching the paper…
Reading the bibliography…
There has been a significant research interest in employing large language models to empower intelligent robots with complex reasoning.
T. Winograd, “Procedures as a representation for data in a computer program for understanding natural language,” Ph.D. dissertation, Massachusetts Institute of Technology, 1971
1971
Earlier work this paper cites.
S. Harnad, “The symbol grounding problem,” Physica D , vol. 42, pp. 335–346, 1990
1990
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction . Cambridge, MA: MIT Press, 1998
1998
Earlier work this paper cites.
M. MacMahon, B. Stankiewicz, and B. Kuipers, “Walk the talk: Connecting language, knowledge, and action in route instructions,” in Proceedings of the National Conference on Artificial Intelligence (AAAI) , 2006
2006
Earlier work this paper cites.
D. Koller and N. Friedman, Probabilistic graphical models: principles and techniques , 2009
2009
Earlier work this paper cites.
T. Kollar, S. Tellex, D. Roy, and N. Roy, “Toward understanding natural language directions,” in Proceedings of the ACM/IEEE International Conference on Human-Robot Interaction (HRI) , 2010
2010
Earlier work this paper cites.
C. Matuszek, D. Fox, and K. Koscher, “Following directions using statistical machine translation,” in Proceedings of the ACM/IEEE International Conference on Human-Robot Interaction (HRI) , 2010
2010
Earlier work this paper cites.
D. L. Chen and R. J. Mooney, “Learning to interpret natural language navigation instructions from observations,” in Proceedings of the National Conference on Artificial Intelligence (AAAI) , 2011
2011
Earlier work this paper cites.
S. Tellex, T. Kollar, S. Dickerson, M. R. Walter, A. G. Banerjee, S. Teller, and N. Roy, “Understanding natural language commands for robotic navigation and mobile manipulation,” in Proceedings of the National Conference on Artificial Intelligence (AAAI) , 2011
2011
Earlier work this paper cites.
C. Matuszek, E. Herbst, L. Zettlemoyer, and D. Fox, “Learning to parse natural language commands to a robot control system,” in Proceedings of the International Symposium on Experimental Robotics (ISER) , 2012
2012
Earlier work this paper cites.
A. Nordmann, N. Hochgeschwender, and S. B. Wrede, “A survey on domain-specific languages in robotics,” in Simulation, Modeling, and Programming for Autonomous Robots , 2014
2014
Earlier work this paper cites.
T. M. Howard, S. Tellex, and N. Roy, “A natural language planner interface for mobile manipulators,” in Proceedings of the IEEE International Conference on Robotics and Automation (ICRA) , 2014
2014
Earlier work this paper cites.
J. Thomason, S. Zhang, R. J. Mooney, and P. Stone, “Learning to interpret natural language commands through human-robot dialog,” in Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI) , 2015
2015
Earlier work this paper cites.
D. K. Misra, J. Sung, K. Lee, and A. Saxena, “Tell me Dave: Context-sensitive grounding of natural language to manipulation instructions,” International Journal of Robotics Research , vol. 35, no. 1-3, pp. 281–300, January 2016
2016
Earlier work this paper cites.
J. Thomason, J. Sinapov, M. Svetlik, P. Stone, and R. J. Mooney, “Learning multi-modal grounded linguistic semantics by playing “I spy”,” in Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI) , 2016
2016
Earlier work this paper cites.
H. Mei, M. Bansal, and M. Walter, “Listen, attend, and walk: Neural mapping of navigational instructions to action sequences,” in Proceedings of the National Conference on Artificial Intelligence (AAAI) , 2016
2016
Earlier work this paper cites.
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. D. Reid, S. Gould, and A. van den Hengel, “Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2017
2017
Earlier work this paper cites.
J. Thomason, J. Sinapov, R. J. Mooney, and P. Stone, “Guiding exploratory behaviors for multi-modal grounding of linguistic descriptions,” in Proceedings of the National Conference on Artificial Intelligence (AAAI) , 2018
2018
Earlier work this paper cites.
M. Shridhar and D. Hsu, “Interactive visual grounding of referring expressions for human-robot interaction,” in Proceedings of Robotics: Science and Systems (RSS) , 2018
2018
Earlier work this paper cites.
R. Paul, J. Arkin, D. Aksaray, N. Roy, and T. M. Howard, “Efficient grounding of abstract spatial concepts for natural language interaction with robot platforms,” International Journal of Robotics Research , vol. 37, no. 10, pp. 1269–1299, June 2018
2018
Earlier work this paper cites.
D. Fried, R. Hu, V. Cirik, A. Rohrbach, J. Andreas, L.-P. Morency, T. Berg-Kirkpatrick, K. Saenko, D. Klein, and T. Darrell, “Speaker-follower models for vision-and-language navigation,” in Advances in Neural Information Processing Systems (NeurIPS) , Dec. 2018
2018
Earlier work this paper cites.
F. Zhu, Y. Zhu, X. Chang, and X. Liang, “Vision-language navigation with self-supervised auxiliary reasoning tasks,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , Jun. 2020
2020
Earlier work this paper cites.
A. Majumdar, A. Shrivastava, S. Lee, P. Anderson, D. Parikh, and D. Batra, “Improving vision-and-language navigation with image-text pairs from the Web,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
2022
Later among the works it cites.
2022
Later among the works it cites.
OpenAI, “GPT-4 technical report,” arXiv preprint arXiv:2303.08774 , 2023
2023
Closest in time.
2023
Closest in time.
“Anthropic introducing 100k Context windows,” https://www.anthropic.com/index/100k-context-windows , accessed: 2023-05-11
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2021
Cited alongside, same era.
A. Kamath, M. Singh, Y. LeCun, G. Synnaeve, I. Misra, and N. Carion, “MDETR - Modulated detection for end-to-end multi-modal understanding,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2021
2021
Cited alongside, same era.
Z. Zhao, E. Wallace, S. Feng, D. Klein, and S. Singh, “Calibrate before use: Improving few-shot performance of language models,” in International Conference on Machine Learning . PMLR, 2021, pp. 12 697–12 706
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
R. Wang, J. Mao, J. Hsu, H. Zhao, J. Wu, and Y. Gao, “Programmatically grounded, compositionally generalizable robotic manipulation,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2023
2023
Closest in time.
A. Z. Ren, B. Govil, T.-Y. Yang, K. R. Narasimhan, and A. Majumdar, “Leveraging language for accelerated learning of tool manipulation,” in Proceedings of the Conference on Robot Learning (CoRL) , 2023
2023
Closest in time.
S. Y. Gadre, M. Wortsman, G. Ilharco, L. Schmidt, and S. Song, “Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Closest in time.
D. Shah, B. Osiński, S. Levine et al. , “LM-Nav: Robotic navigation with large pre-trained models of language, vision, and action,” in Proceedings of the Conference on Robot Learning (CoRL) , 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.