Fetching the paper…
Reading the bibliography…
We study the challenging problem of releasing a robot in a previously unseen environment, and having it follow unconstrained natural language navigation instructions.
Walk the talk: Connecting language, knowledge, and action in route instructions
M. MacMahon, B. Stankiewicz, and B. Kuipers · 2006
Earlier work this paper cites.
Improved techniques for grid mapping with rao-blackwellized particle filters
G. Grisetti, C. Stachniss, and W. Burgard · 2007
Earlier work this paper cites.
Ros: an open-source robot operating system
M. Quigley, K. Conley, B. Gerkey, J. Faust, T. Foote, J. Leibs, R. Wheeler, and A. Y. Ng · 2009
Earlier work this paper cites.
Toward understanding natural language directions
T. Kollar, S. Tellex, D. Roy, and N. Roy · 2010
Earlier work this paper cites.
Learning to follow navigational directions
A. Vogel and D. Jurafsky · 2010
Earlier work this paper cites.
Natural language command of an autonomous micro-air vehicle
A. S. Huang, S. Tellex, A. Bachrach, T. Kollar, D. Roy, and N. Roy · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
D. L. Chen and R. J. Mooney · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. R. Walter, A. G. Banerjee, S. J. Teller, and N. Roy · 2011
Earlier work this paper cites.
Grounding spatial relations for human-robot interaction
S. Guadarrama, L. Riano, D. Golland, D. Go, Y. Jia, D. Klein, P. Abbeel, T. Darrell, et al · 2013
Earlier work this paper cites.
Learning to parse natural language commands to a robot control system
C. Matuszek, E. Herbst, L. Zettlemoyer, and D. Fox · 2013
Earlier work this paper cites.
Sinkhorn distances: Lightspeed computation of optimal transport
M. Cuturi · 2013
Earlier work this paper cites.
Learning models for following natural language directions in unknown environments
S. Hemachandra, F. Duvallet, T. M. Howard, N. Roy, A. Stentz, and M. R. Walter · 2015
Earlier work this paper cites.
U-Net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Tell me dave: Context-sensitive grounding of natural language to manipulation instructions
D. K. Misra, J. Sung, K. Lee, and A. Saxena · 2016
Earlier work this paper cites.
Efficient grounding of abstract spatial concepts for natural language interaction with robot manipulators
R. Paul, J. Arkin, N. Roy, and T. M Howard · 2016
Earlier work this paper cites.
Natural language communication with robots
Y. Bisk, D. Yuret, and D. Marcu · 2016
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
H. Mei, M. Bansal, and M. R. Walter · 2016
Earlier work this paper cites.
Integrated intelligence for human-robot teams
J. Oh, T. M. Howard, M. R. Walter, D. Barber, M. Zhu, S. Park, A. Suppe, L. Navarro-Serment, F. Duvallet, A. Boularias, et al · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Gated-attention architectures for task-oriented language grounding
D. S. Chaplot, K. M. Sathyendra, R. K. Pasumarthi, D. Rajagopal, and R. Salakhutdinov · 2017
Cited alongside, same era.
Matterport3d: Learning from rgb-d data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niessner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Cited alongside, same era.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Cited alongside, same era.
Automatic differentiation in PyTorch
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer · 2017
Cited alongside, same era.
Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Cited alongside, same era.
E. Fahnestock, S. Patki, and T. M. Howard · 2019
Later among the works it cites.
Vision-and-dialog navigation
J. Thomason, M. Murray, M. Cakmak, and L. Zettlemoyer · 2019
Later among the works it cites.
Help, Anna! visual navigation with natural multimodal assistance via retrospective curiosity-encouraging imitation learning
K. Nguyen and H. Daumé III · 2019
Later among the works it cites.
Cross-lingual vision-language navigation
A. Yan, X. Wang, J. Feng, L. Li, and W. Y. Wang · 2019
Later among the works it cites.
Navigation agents for the visually impaired: A sidewalk simulator and experiments
M. Weiss, S. Chamorro, R. Girgis, M. Luck, S. E. Kahou, J. P. Cohen, D. Nowrouzezahrai, D. Precup, F. Golemo, and C. Pal · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation
X. Wang, W. Xiong, H. Wang, and W. Y. Wang · 2018
Cited alongside, same era.
Speaker-follower models for vision-and-language navigation
D. Fried, R. Hu, V. Cirik, A. Rohrbach, J. Andreas, L.-P. Morency, T. Berg-Kirkpatrick, K. Saenko, D. Klein, and T. Darrell · 2018
Cited alongside, same era.
End-to-end driving via conditional imitation learning
F. Codevilla, M. Miiller, A. López, V. Koltun, and A. Dosovitskiy · 2018
Cited alongside, same era.
Temporal grounding graphs for language understanding with accrued visual-linguistic context
R. Paul, A. Barbu, S. Felshin, B. Katz, and N. Roy · 2018
Cited alongside, same era.
Learning to navigate in cities without a map
P. Mirowski, M. Grimes, M. Malinowski, K. M. Hermann, K. Anderson, D. Teplyashin, K. Simonyan, A. Zisserman, R. Hadsell, et al · 2018
Cited alongside, same era.
On evaluation of embodied navigation agents
P. Anderson, A. Chang, D. S. Chaplot, A. Dosovitskiy, S. Gupta, V. Koltun, J. Kosecka, J. Malik, R. Mottaghi, M. Savva, and A. R. Zamir · 2018
Cited alongside, same era.
Gibson env: real-world perception for embodied agents
F. Xia, A. R. Zamir, Z.-Y. He, A. Sax, J. Malik, and S. Savarese · 2018
Cited alongside, same era.
Later among the works it cites.
Touchdown: Natural language navigation and spatial reasoning in visual street environments
H. Chen, A. Suhr, D. Misra, and Y. Artzi · 2019
Later among the works it cites.
Talk2nav: Long-range vision-and-language navigation in cities
A. B. Vasudevan, D. Dai, and L. Van Gool · 2019
Later among the works it cites.
Run through the streets: A new dataset and baseline models for realistic urban navigation
T. Paz-Argaman and R. Tsarfaty · 2019
Later among the works it cites.
Effective and general evaluation for instruction conditioned navigation using dynamic time warping
G. Magalhaes, V. Jain, A. Ku, E. Ie, and J. Baldridge · 2019
Later among the works it cites.
Off-policy deep reinforcement learning without exploration
S. Fujimoto, D. Meger, and D. Precup · 2019
Later among the works it cites.
Habitat: A Platform for Embodied AI Research
Manolis Savva*, Abhishek Kadian*, Oleksandr Maksymets*, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra · 2019
Later among the works it cites.
Learning unknown groundings for natural language interaction with mobile robots
M. Tucker, D. Aksaray, R. Paul, G. J. Stein, and N. Roy · 2020
Closest in time.
The RobotSlang Benchmark: Dialog-guided robot localization and navigation
S. Banerjee, J. Thomason, and J. J. Corso · 2020
Closest in time.
Reverie: Remote embodied visual referring expression in real indoor environments
Y. Qi, Q. Wu, P. Anderson, X. Wang, W. Y. Wang, C. Shen, and A. van den Hengel · 2020
Closest in time.
Room-Across-Room: Multilingual vision-and-language navigation with dense spatiotemporal grounding
A. Ku, P. Anderson, R. Patel, E. Ie, and J. Baldridge · 2020
Closest in time.
Retouchdown: Adding touchdown to streetlearn as a shareable resource for language grounding tasks in street view
H. Mehta, Y. Artzi, J. Baldridge, E. Ie, and P. Mirowski · 2020
Closest in time.
Beyond the nav-graph: Vision-and-language navigation in continuous environments
J. Krantz, E. Wijmans, A. Majumdar, D. Batra, and S. Lee · 2020
Closest in time.
Are we making real progress in simulated environments? Measuring the sim2real gap in embodied visual navigation
A. Kadian, J. Truong, A. Gokaslan, A. Clegg, E. Wijmans, S. Lee, M. Savva, S. Chernova, and D. Batra · 2020
Closest in time.