Fetching the paper…
Reading the bibliography…
Robots navigating in human environments should use language to ask for assistance and be able to understand human responses.
The HCRC Map Task Corpus
A. Anderson, M. Bader, E. Bard, Boyle, G. M. E., Doherty, S. Garrod, S. Isard, J. Kowtko, J. McAllister, J. Miller, C. Sotillo, H. S. Thompson, and R. Weinert · 1991
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
M. MacMahon, B. Stankiewicz, and B. Kuipers · 2006
Earlier work this paper cites.
SCARE: a Situated Corpus with Annotated Referring Expressions
L. Stoia, D. M. Shockley, D. K. Byron, and E. Fosler-Lussier · 2008
Earlier work this paper cites.
The Indiana “Cooperative Remote Search Task” (CReST) Corpus
K. Eberhard, H. Nicholson, S. Kübler, S. Gunderson, and M. Scheutz · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
D. L. Chen and R. J. Mooney · 2011
Earlier work this paper cites.
Integrating reinforcement learning with human demonstrations of varying ability
M. E. Taylor, H. B. Suay, and S. Chernova · 2011
Earlier work this paper cites.
Asking for help using inverse semantics
S. Tellex, R. Knepper, A. Li, D. Rus, and N. Roy · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Matterport3D: Learning from RGB-D data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niessner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Earlier work this paper cites.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. Lawrence Zitnick, and R. Girshick · 2017
Earlier work this paper cites.
Visual Dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra · 2017
Earlier work this paper cites.
Improving reinforcement learning with confidence-based demonstrations
Z. Wang and M. E. Taylor · 2017
Cited alongside, same era.
AI2-THOR: An Interactive 3D Environment for Visual AI
E. Kolve, R. Mottaghi, W. Han, E. VanderBilt, L. Weihs, A. Herrasti, D. Gordon, Y. Zhu, A. Gupta, and A. Farhadi · 2017
Cited alongside, same era.
Language to action: Towards interactive task learning with physical agents
J. Y. Chai, Q. Gao, L. She, S. Yang, S. Saba-Sadiya, and G. Xu · 2018
Cited alongside, same era.
Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Cited alongside, same era.
Mapping navigation instructions to continuous control actions with position visitation prediction
V. Blukis, D. Misra, R. A. Knepper, and Y. Artzi · 2018
Cited alongside, same era.
Dempster-shafer theoretic resolution of referential ambiguity
T. Williams, F. Yazdani, P. Suresh, M. Scheutz, and M. Beetz · 2019
Closest in time.
Touchdown: Natural language navigation and spatial reasoning in visual street environments
H. Chen, A. Suhr, D. Misra, N. Snavely, and Y. Artzi · 2019
Closest in time.
GQA: A new dataset for real-world visual reasoning and compositional question answering
D. A. Hudson and C. D. Manning · 2019
Closest in time.
From recognition to cognition: Visual commonsense reasoning
R. Zellers, Y. Bisk, A. Farhadi, and Y. Choi · 2019
Closest in time.
CLEVR-Dialog: A diagnostic dataset for multi-round reasoning in visual dialog
S. Kottur, J. M. Moura, D. Parikh, D. Batra, and M. Rohrbach · 2019
Closest in time.
CoQa: A conversational question answering challenge
S. Reddy, D. Chen, and C. D. Manning · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra · 2018
Cited alongside, same era.
IQA: Visual question answering in interactive environments
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi · 2018
Cited alongside, same era.
QuAC: Question answering in context
E. Choi, H. He, M. Iyyer, M. Yatskar, S. Yih, Y. Choi, P. Liang, and L. Zettlemoyer · 2018
Cited alongside, same era.
Interpretation of natural language rules in conversational machine reading
M. Saeidi, M. Bartolo, P. Lewis, S. Singh, T. Rocktäschel, M. Sheldon, G. Bouchard, and S. Riedel · 2018
Cited alongside, same era.
Talk the walk: Navigating new york city through grounded dialogue
H. de Vries, K. Shuster, D. Batra, D. Parikh, J. Weston, and D. Kiela · 2018
Cited alongside, same era.
Improving grounded natural language understanding through human-robot dialog
J. Thomason, A. Padmakumar, J. Sinapov, N. Walker, Y. Jiang, H. Yedidsion, J. Hart, P. Stone, and R. J. Mooney · 2019
Cited alongside, same era.
Learning from human-robot interactions in modeled scenes
M. Murnane, M. Breitmeyer, F. Ferraro, C. Matuszek, and D. Engel · 2019
Cited alongside, same era.
Closest in time.
Vision-based navigation with language-based assistance via imitation learning with indirect intervention
K. Nguyen, D. Dey, C. Brockett, and B. Dolan · 2019
Closest in time.
Help, Anna! Visual Navigation with Natural Multimodal Assistance via Retrospective Curiosity-Encouraging Imitation Learning
K. Nguyen and H. Daumé III · 2019
Closest in time.
A research platform for multi-robot dialogue with humans
M. Marge, S. Nogar, C. Hayes, S. Lukin, J. Bloecker, E. Holder, and C. Voss · 2019
Closest in time.
Stay on the path: Instruction fidelity in vision-and-language navigation
V. Jain, G. Magalhaes, A. Ku, A. Vaswani, E. Ie, and J. Baldridge · 2019
Closest in time.
Shifting the Baseline: Single Modality Performance on Visual Navigation & QA
J. Thomason, D. Gordon, and Y. Bisk · 2019
Closest in time.
Learning to navigate unseen environments: Back translation with environmental dropout
H. Tan, L. Yu, and M. Bansal · 2019
Closest in time.