Fetching the paper…
Reading the bibliography…
Recently there has been a rising interest in training agents, embodied in virtual environments, to perform language-directed tasks by deep reinforcement learning.
Shortest connection networks and some generalizations
R. C. Prim · 1957
Earlier work this paper cites.
The symbol grounding problem
S. Harnad · 1990
Earlier work this paper cites.
Grounding language in perception
J. M. Siskind · 1994
Earlier work this paper cites.
Reinforcement Learning : An Introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Artificial Intelligence , 101(1), 1998
Planning and acting in partially observable stochastic domains · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour · 1999
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
L. Smith and M. Gasser · 2005
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
D. L. Chen and R. J. Mooney · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. R. Walter, A. G. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Lecture 6.5 - RMSprop
T. Tieleman and G. Hinton · 2012
Earlier work this paper cites.
Grounded language learning from video described with sentences
H. Yu and J. M. Siskind · 2013
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, Ç. Gülçehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
VQA: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. Lawrence Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Virtual embodiment: A scalable long-term strategy for artificial intelligence research
D. Kiela, L. Bulat, A. L. Vero, and S. Clark · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Hierarchical question-image co-attention for visual question answering
J. Lu, J. Yang, D. Batra, and D. Parikh · 2016
Cited alongside, same era.
Stacked attention networks for image question answering
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola · 2016
Cited alongside, same era.
Physical causality of action verbs in grounded language understanding
Q. Gao, M. Doering, S. Yang, and J. Y. Chai · 2016
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Later among the works it cites.
MINOS: Multimodal indoor simulator for navigation in complex environments
M. Savva, A. X. Chang, A. Dosovitskiy, T. Funkhouser, and V. Koltun · 2017
Later among the works it cites.
Unifying map and landmark based representations for visual navigation
S. Gupta, D. F. Fouhey, S. Levine, and J. Malik · 2017
Later among the works it cites.
Driving under the influence (of language)
D. P. Barrett, S. A. Bronikowski, H. Yu, and J. M. Siskind · 2017
Later among the works it cites.
Gated-attention architectures for task-oriented language grounding
D. S. Chaplot, K. M. Sathyendra, R. K. Pasumarthi, D. Rajagopal, and R. Salakhutdinov · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Grounding of textual phrases in images by reconstruction
A. Rohrbach, M. Rohrbach, R. Hu, T. Darrell, and B. Schiele · 2016
Cited alongside, same era.
Zero-shot task generalization with multi-task deep reinforcement learning
J. Oh, S. P. Singh, H. Lee, and P. Kohli · 2017
Cited alongside, same era.
Grounded language learning in a simulated 3d world
K. M. Hermann, F. Hill, S. Green, F. Wang, R. Faulkner, H. Soyer, D. Szepesvari, W. M. Czarnecki, M. Jaderberg, D. Teplyashin, M. Wainwright, C. Apps, D. Hassabis, and P. Blunsom · 2017
Cited alongside, same era.
Efficient Parallel Methods for Deep Reinforcement Learning
A. V. Clemente, H. N. Castejón, and A. Chandra · 2017
Cited alongside, same era.
Reinforcement learning with unsupervised auxiliary tasks
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2017
Cited alongside, same era.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, D. Kumaran, and R. Hadsell · 2017
Cited alongside, same era.
Embodied Question Answering
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra · 2018
Closest in time.
Interactive grounded language acquisition and generalization in a 2d world
H. Yu, H. Zhang, and W. Xu · 2018
Closest in time.
Hierarchical and interpretable skill acquisition in multi-task reinforcement learning
T. Shu, C. Xiong, and R. Socher · 2018
Closest in time.
Building generalizable agents with a realistic and rich 3d environment
Y. Wu, Y. Wu, G. Gkioxari, and Y. Tian · 2018
Closest in time.
FiLM: Visual Reasoning with a General Conditioning Layer
E. Perez, F. Strub, H. De Vries, V. Dumoulin, and A. Courville · 2018
Closest in time.
Learning to navigate in cities without a map
P. Mirowski, M. K. Grimes, M. Malinowski, K. M. Hermann, K. Anderson, D. Teplyashin, K. Simonyan, K. Kavukcuoglu, A. Zisserman, and R. Hadsell · 2018
Closest in time.
Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Closest in time.
IQA: visual question answering in interactive environments
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi · 2018
Closest in time.