Fetching the paper…
Reading the bibliography…
We tackle a task where an agent learns to navigate in a 2D maze-like environment called XWORLD.
Bidirectional recurrent neural networks
M. Schuster and K. K. Paliwal · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
D. L. Chen and R. J. Mooney · 2011
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
J. Duchi, E. Hazan, and Y. Singer · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. R. Walter, A. G. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Grounded language learning from video described with sentences
H. Yu and J. M. Siskind · 2013
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, Ç. Gülçehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. Lawrence Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Robot language learning, generation, and comprehension
D. P. Barrett, S. A. Bronikowski, H. Yu, and J. M. Siskind · 2015
Earlier work this paper cites.
Are you talking to a machine? dataset and methods for multilingual image question
H. Gao, J. Mao, J. Zhou, Z. Huang, L. Wang, and W. Xu · 2015
Cited alongside, same era.
Learning like a child: Fast novel visual concept learning from sentence descriptions of images
J. Mao, X. Wei, Y. Yang, J. Wang, Z. Huang, and A. L. Yuille · 2015
Cited alongside, same era.
A roadmap towards machine intelligence
T. Mikolov, A. Joulin, and M. Baroni · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Cited alongside, same era.
Exploring models and data for image question answering
M. Ren, R. Kiros, and R. Zemel · 2015
Cited alongside, same era.
Building machines that learn and think like people
B. M. Lake, T. D. Ullman, J. B. Tenenbaum, and S. J. Gershman · 2016
Later among the works it cites.
Hierarchical question-image co-attention for visual question answering
J. Lu, J. Yang, D. Batra, and D. Parikh · 2016
Later among the works it cites.
Neural programmer: Inducing latent programs with gradient descent
A. Neelakantan, Q. V. Le, and I. Sutskever · 2016
Later among the works it cites.
Grounding of textual phrases in images by reconstruction
A. Rohrbach, M. Rohrbach, R. Hu, T. Darrell, and B. Schiele · 2016
Later among the works it cites.
Prioritized experience replay
T. Schaul, J. Quan, I. Antonoglou, and D. Silver · 2016
Later among the works it cites.
Mazebase: A sandbox for learning from games
S. Sukhbaatar, A. Szlam, G. Synnaeve, S. Chintala, and R. Fergus · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Beattie, J. Z. Leibo, D. Teplyashin, T. Ward, M. Wainwright, H. Küttler, A. Lefrancq, S. Green, V. Valdés, A. Sadik, J. Schrittwieser, K. Anderson, S. York, M. Cant, A. Cain, A. Bolton, S. Gaffney, H. King, D. Hassabis, S. Legg, and S. Petersen · 2016
Cited alongside, same era.
Physical causality of action verbs in grounded language understanding
Q. Gao, M. Doering, S. Yang, and J. Y. Chai · 2016
Cited alongside, same era.
The malmo platform for artificial intelligence experimentation
M. Johnson, K. Hofmann, T. Hutton, and D. Bignell · 2016
Cited alongside, same era.
ViZDoom: A Doom-based AI research platform for visual reinforcement learning
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski · 2016
Cited alongside, same era.
Virtual embodiment: A scalable long-term strategy for artificial intelligence research
D. Kiela, L. Bulat, A. L. Vero, and S. Clark · 2016
Cited alongside, same era.
Neural module networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein
Cited in the paper.
Learning to compose neural networks for question answering
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein
Cited in the paper.
Later among the works it cites.
Value iteration networks
A. Tamar, S. Levine, P. Abbeel, Y. WU, and G. Thomas · 2016
Later among the works it cites.
Zero-shot visual question answering
D. Teney and A. v. d. Hengel · 2016
Later among the works it cites.
Stacked attention networks for image question answering
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola · 2016
Later among the works it cites.
Learning to reason: End-to-end module networks for visual question answering
R. Hu, J. Andreas, M. Rohrbach, T. Darrell, and K. Saenko · 2017
Closest in time.