Fetching the paper…
Reading the bibliography…
Language-guided Embodied AI benchmarks requiring an agent to navigate an environment and manipulate objects typically allow one-way communication: the human user gives a natural language command to the agent, and the agent can only follow the command passively.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
K. Jokinen and M. McTear, “Spoken dialogue systems,” Synthesis Lectures on Human Language Technologies , vol. 2, no. 1, pp. 1–151, 2009
2009
Earlier work this paper cites.
V. Rus, B. Wyse, P. Piwek, M. Lintean, S. Stoyanchev, and C. Moldovan, “The first question generation shared task evaluation challenge,” in Proceedings of the 6th International Natural Language Generation Conference , 2010
2010
Earlier work this paper cites.
X. Lu, “The relationship of lexical richness to the quality of esl learners’ oral narratives,” The Modern Language Journal , vol. 96, no. 2, pp. 190–208, 2012
2012
Earlier work this paper cites.
S. Tellex, R. Knepper, A. Li, D. Rus, and N. Roy, “Asking for help using inverse semantics,” in Proceedings of Robotics: Science and Systems , Berkeley, USA, July 2014
2014
Earlier work this paper cites.
M.-T. Luong, H. Pham, and C. D. Manning, “Effective approaches to attention-based neural machine translation,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing , 2015, pp. 1412–1421
2015
Earlier work this paper cites.
P.-H. Su, M. Gašić, N. Mrkšić, L. M. Rojas-Barahona, S. Ultes, D. Vandyke, T.-H. Wen, and S. Young, “On-line active reward learning for policy optimisation in spoken dialogue systems,” in Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Berlin, Germany, Aug. 2016, pp. 2431–2441
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
N. Mrkšić, D. Ó Séaghdha, T.-H. Wen, B. Thomson, and S. Young, “Neural belief tracker: Data-driven dialogue state tracking,” in Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Vancouver, Canada, July 2017, pp. 1777–1788
2017
Earlier work this paper cites.
B. Peng, X. Li, L. Li, J. Gao, A. Celikyilmaz, S. Lee, and K.-F. Wong, “Composite task-completion dialogue policy learning via hierarchical deep reinforcement learning,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2017
2017
Earlier work this paper cites.
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. Van Den Hengel, “Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 3674–3683
2018
Earlier work this paper cites.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied Question Answering,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018
2018
Cited alongside, same era.
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi, “Iqa: Visual question answering in interactive environments,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 4089–4098
2018
Cited alongside, same era.
J. Gao, M. Galley, and L. Li, “Neural approaches to conversational ai,” in The 41st International ACM SIGIR Conference on Research & Development in Information Retrieval , 2018, pp. 1371–1374
2018
Cited alongside, same era.
J. Y. Chai, Q. Gao, L. She, S. Yang, S. Saba-Sadiya, and G. Xu, “Language to action: Towards interactive task learning with physical agents.” in IJCAI , 2018, pp. 2–9
2018
Cited alongside, same era.
J. Thomason, A. Padmakumar, J. Sinapov, N. Walker, Y. Jiang, H. Yedidsion, J. Hart, P. Stone, and R. Mooney, “Jointly improving parsing and perception for natural language commands through human-robot dialog,” Journal of Artificial Intelligence Research , vol. 67, pp. 327–374, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
X. Hu, Z. Wen, Y. Wang, X. Li, and G. de Melo, “Interactive question clarification in dialogue via reinforcement learning,” in Proceedings of the 28th International Conference on Computational Linguistics: Industry Track , 2020, pp. 78–89
2020
Later among the works it cites.
H. R. Roman, Y. Bisk, J. Thomason, A. Celikyilmaz, and J. Gao, “Rmm: A recursive mental model for dialog navigation,” in Findings of the 2020 Conference on Empirical Methods in Natural Language Processing , 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
K. Nguyen, D. Dey, C. Brockett, and B. Dolan, “Vision-based navigation with language-based assistance via imitation learning with indirect intervention,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 12 527–12 537
2019
Cited alongside, same era.
K. Nguyen and H. Daumé III, “Help, anna! visual navigation with natural multimodal assistance via retrospective curiosity-encouraging imitation learning,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP) , November 2019
2019
Cited alongside, same era.
M. Shridhar, J. Thomason, D. Gordon, Y. Bisk, W. Han, R. Mottaghi, L. Zettlemoyer, and D. Fox, “Alfred: A benchmark for interpreting grounded instructions for everyday tasks,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 10 740–10 749
2020
Cited alongside, same era.
F. Xia, W. B. Shen, C. Li, P. Kasimbeg, M. E. Tchapmi, A. Toshev, R. Martín-Martín, and S. Savarese, “Interactive gibson benchmark: A benchmark for interactive navigation in cluttered environments,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 713–720, 2020
2020
Cited alongside, same era.
J. Thomason, M. Murray, M. Cakmak, and L. Zettlemoyer, “Vision-and-dialog navigation,” in Conference on Robot Learning , 2020, pp. 394–406
2020
Cited alongside, same era.
2020
Later among the works it cites.
T.-C. Chi, M. Shen, M. Eric, S. Kim, and D. Hakkani-tur, “Just ask: An interactive learning framework for vision and language navigation,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 03, 2020, pp. 2459–2466
2020
Later among the works it cites.
V. Blukis, C. Paxton, D. Fox, A. Garg, and Y. Artzi, “A persistent spatial semantic representation for high-level natural language instruction execution,” in Conference on Robot Learning , 2021, pp. 706–717
2021
Later among the works it cites.
A. Pashevich, C. Schmid, and C. Sun, “Episodic transformer for vision-and-language navigation,” 2021 IEEE/CVF International Conference on Computer Vision (ICCV) , pp. 15 922–15 932, 2021
2021
Later among the works it cites.
S. Y. Min, D. S. Chaplot, P. K. Ravikumar, Y. Bisk, and R. Salakhutdinov, “FILM: Following instructions in language with modular methods,” in International Conference on Learning Representations , 2022
2022
Closest in time.
A. Padmakumar, J. Thomason, A. Shrivastava, P. Lange, A. Narayan-Chen, S. Gella, R. Piramithu, G. Tur, and D. Hakkani-Tur, “TEACh: Task-driven Embodied Agents that Chat,” in Conference on Artificial Intelligence (AAAI) , 2022
2022
Closest in time.