Fetching the paper…
Reading the bibliography…
Natural language provides an accessible and expressive interface to specify long-term tasks for robotic agents.
Logic and conversation
H. P. Grice · 1975
Earlier work this paper cites.
Geometric modeling using octree encoding
D. Meagher · 1982
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
M. MacMahon, B. Stankiewicz, and B. Kuipers · 2006
Earlier work this paper cites.
Incremental natural language processing for hri
T. Brick and M. Scheutz · 2007
Earlier work this paper cites.
Probabilistic semantic mapping with a virtual sensor for building/nature detection
M. Persson, T. Duckett, C. Valgren, and A. Lilienthal · 2007
Earlier work this paper cites.
Conceptual spatial representations for indoor mobile robots
H. Zender, O. M. Mozos, P. Jensfelt, G.-J. Kruijff, and W. Burgard · 2008
Earlier work this paper cites.
Toward understanding natural language directions
T. Kollar, S. Tellex, D. Roy, and N. Roy · 2010
Earlier work this paper cites.
Reading between the lines: Learning to map high-level instructions to commands
S. R. K. Branavan, L. S. Zettlemoyer, and R. Barzilay · 2010
Earlier work this paper cites.
Following directions using statistical machine translation
C. Matuszek, D. Fox, and K. Koscher · 2010
Earlier work this paper cites.
Multi-modal semantic place classification
A. Pronobis, O. Martinez Mozos, B. Caputo, and P. Jensfelt · 2010
Earlier work this paper cites.
”Approaching the Symbol Grounding Problem with Probabilistic Graphical Models
S. Tellex, T. Kollar, S. Dickerson, M. R. Walter, A. Gopal Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. Walter, A. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Semantic mapping with mobile robots
A. Pronobis · 2011
Earlier work this paper cites.
Learning to parse natural language commands to a robot control system
C. Matuszek, E. Herbst, L. Zettlemoyer, and D. Fox · 2012
Earlier work this paper cites.
A Joint Model of Language and Perception for Grounded Attribute Learning
C. Matuszek, N. FitzGerald, L. Zettlemoyer, L. Bo, and D. Fox · 2012
Earlier work this paper cites.
Learning Semantic Maps from Natural Language Descriptions
M. R. Walter, S. Hemachandra, B. Homberg, S. Tellex, and S. Teller · 2013
Earlier work this paper cites.
Toward information theoretic human-robot dialog
S. Tellex, P. Thaker, R. Deits, D. Simeonov, T. Kollar, and N. Roy · 2013
Earlier work this paper cites.
Imitation learning for natural language direction following through unknown environments
F. Duvallet, T. Kollar, and A. Stentz · 2013
Earlier work this paper cites.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Y. Artzi and L. Zettlemoyer · 2013
Earlier work this paper cites.
Tell me dave: Context-sensitive grounding of natural language to mobile manipulation instructions
D. K. Misra, J. Sung, K. Lee, and A. Saxena · 2014
Earlier work this paper cites.
Asking for help using inverse semantics
S. Tellex, R. Knepper, A. Li, D. Rus, and N. Roy · 2014
Earlier work this paper cites.
Learning spatial-semantic representations from natural language descriptions and scene classifications
S. Hemachandra, M. R. Walter, S. Tellex, and S. Teller · 2014
Earlier work this paper cites.
Learning to interpret natural language commands through human-robot dialog
J. Thomason, S. Zhang, R. J. Mooney, and P. Stone · 2015
Earlier work this paper cites.
Learning models for following natural language directions in unknown environments
S. Hemachandra, F. Duvallet, T. M. Howard, N. Roy, A. Stentz, and M. R. Walter · 2015
Earlier work this paper cites.
Semantic mapping for mobile robotics tasks: A survey
I. Kostavelis and A. Gasteratos · 2015
Earlier work this paper cites.
Recovering from Failure by Asking for Help
R. A. Knepper, S. Tellex, A. Li, N. Roy, and D. Rus · 2015
Earlier work this paper cites.
Environment-driven lexicon induction for high-level instructions
D. K. Misra, K. Tao, P. Liang, and A. Saxena · 2015
Cited alongside, same era.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Cited alongside, same era.
Value iteration networks
A. Tamar, Y. Wu, G. Thomas, S. Levine, and P. Abbeel · 2016
Cited alongside, same era.
Mapping instructions and visual observations to actions with reinforcement learning
D. Misra, J. Langford, and Y. Artzi · 2017
Cited alongside, same era.
Grounded language learning in a simulated 3d world
K. M. Hermann, F. Hill, S. Green, F. Wang, R. Faulkner, H. Soyer, D. Szepesvari, W. Czarnecki, M. Jaderberg, D. Teplyashin, et al · 2017
Cited alongside, same era.
Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation?
Touchdown: Natural language navigation and spatial reasoning in visual street environments
H. Chen, A. Suhr, D. Misra, and Y. Artzi · 2019
Later among the works it cites.
Chasing ghosts: Instruction following as bayesian state tracking
P. Anderson, A. Shrivastava, D. Parikh, D. Batra, and S. Lee · 2019
Later among the works it cites.
Detectron2
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Later among the works it cites.
Learning navigation behaviors end-to-end with autorl
H.-T. L. Chiang, A. Faust, M. Fiser, and A. Francis · 2019
Later among the works it cites.
Stay on the path: Instruction fidelity in vision-and-language navigation
V. Jain, G. Magalhaes, A. Ku, A. Vaswani, E. Ie, and J. Baldridge · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. McCormac, A. Handa, S. Leutenegger, and A. J. Davison · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Densely connected convolutional networks
G. Huang, Z. Liu, L. van der Maaten, and K. Q. Weinberger · 2017
Cited alongside, same era.
Grounding robot plans from natural language instructions with incomplete world knowledge
D. Nyga, S. Roy, R. Paul, D. Park, M. Pomarlan, M. Beetz, and N. Roy · 2018
Cited alongside, same era.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Cited alongside, same era.
Temporal spatial inverse semantics for robots communicating with humans
Z. Gong and Y. Zhang · 2018
Cited alongside, same era.
Gated-attention architectures for task-oriented language grounding
D. S. Chaplot, K. M. Sathyendra, R. K. Pasumarthi, D. Rajagopal, and R. Salakhutdinov · 2018
Cited alongside, same era.
Later among the works it cites.
Few-shot object grounding and mapping for natural language robot instruction following
V. Blukis, R. A. Knepper, and Y. Artzi · 2020
Later among the works it cites.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
M. Shridhar, J. Thomason, D. Gordon, Y. Bisk, W. Han, R. Mottaghi, L. Zettlemoyer, and D. Fox · 2020
Later among the works it cites.
Sim-to-real transfer for vision-and-language navigation
P. Anderson, A. Shrivastava, J. Truong, A. Majumdar, D. Parikh, D. Batra, and S. Lee · 2020
Later among the works it cites.
Moca: A modular object-centric approach for interactive instruction following
K. P. Singh, S. Bhambri, B. Kim, R. Mottaghi, and J. Choi · 2020
Later among the works it cites.
Conditional driving from natural language instructions
J. Roh, C. Paxton, A. Pronobis, A. Farhadi, and D. Fox · 2020
Later among the works it cites.
Pixl2r: Guiding reinforcement learning using natural language by mapping pixels to rewards
P. Goyal, S. Niekum, and R. J. Mooney · 2020
Later among the works it cites.
Beyond the nav-graph: Vision-and-language navigation in continuous environments
J. Krantz, E. Wijmans, A. Majumdar, D. Batra, and S. Lee · 2020
Later among the works it cites.
Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding
A. Ku, P. Anderson, R. Patel, E. Ie, and J. Baldridge · 2020
Later among the works it cites.
Abp, alfred leaderboard, may 10th 2021
B. Kim, S. Bhambri, K. P. Singh, R. Mottaghi, and J. Choi · 2021
Closest in time.
Episodic Transformer for Vision-and-Language Navigation, 2021
A. Pashevich, C. Schmid, and C. Sun · 2021
Closest in time.
Lwit, alfred leaderboard, may 10th 2021
Anonymous · 2021
Closest in time.
H. Saha, F. Fotouhi, Q. Liu, and S. Sarkar · 2021
Closest in time.
Hierarchical task learning from language instructions with unified transformers and self-monitoring
Y. Zhang and J. Chai · 2021
Closest in time.
Alfred leaderboard
M. Shridhar, J. Thomason, D. Gordon, Y. Bisk, W. Han, R. Mottaghi, L. Zettlemoyer, and D. Fox · 2021
Closest in time.
Modular framework for visuomotor language grounding
K. Nottingham, L. Liang, D. Shin, C. C. Fowlkes, R. Fox, and S. Singh · 2021
Closest in time.
Agent with the big picture: Perceiving surroundings for interactive instruction following
B. Kim, S. Bhambri, K. P. Singh, R. Mottaghi, and J. Choi · 2021
Closest in time.
Look wide and interpret twice: Improving performance on interactive instruction-following tasks
V.-Q. Nguyen, M. Suganuma, and T. Okatani · 2021
Closest in time.
Contact-graspnet: Efficient 6-dof grasp generation in cluttered scenes
M. Sundermeyer, A. Mousavian, R. Triebel, and D. Fox · 2021
Closest in time.
Neural geometric level of detail: Real-time rendering with implicit 3d shapes
T. Takikawa, J. Litalien, K. Yin, K. Kreis, C. Loop, D. Nowrouzezahrai, A. Jacobson, M. McGuire, and S. Fidler · 2021
Closest in time.