Fetching the paper…
Reading the bibliography…
We build a virtual agent for learning language in a 2D maze-like world.
Verbal behavior
Burrhus Frederic Skinner · 1957
Earlier work this paper cites.
Understanding natural language
Terry Winograd · 1972
Earlier work this paper cites.
Child’s talk: Learning to use language
Jerome Bruner · 1985
Earlier work this paper cites.
Principal Component Analysis
I.T. Jolliffe · 1986
Earlier work this paper cites.
The symbol grounding problem
Stevan Harnad · 1990
Earlier work this paper cites.
Linguistics and cognitive science: Problems and mysteries
Noam Chomsky · 1991
Earlier work this paper cites.
Grounding language in perception
Jeffrey Mark Siskind · 1994
Earlier work this paper cites.
A computational study of cross-situational techniques for learning word-to-meaning mappings
Jeffrey Mark Siskind · 1996
Earlier work this paper cites.
A solution to plato’s problem: The latent semantic analysis theory of acquisition, induction, and representation of knowledge
Thomas Landauer and Susan Dumais · 1997
Earlier work this paper cites.
Bidirectional recurrent neural networks
M. Schuster and K. K. Paliwal · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Grounding the lexical semantics of verbs in visual perception using force dynamics and event logic
Jeffrey Mark Siskind · 1999
Earlier work this paper cites.
Constructing a Language: A Usage-Based Theory of Language Acquisition
Michael Tomasello · 2003
Earlier work this paper cites.
Meaning and links
William A. Woods · 2007
Earlier work this paper cites.
Visualizing high-dimensional data using t-SNE
L.J.P van der Maaten and G.E. Hinton · 2008
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston · 2009
Earlier work this paper cites.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
David L Chen and Raymond J Mooney · 2011
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer · 2011
Cited alongside, same era.
Understanding natural language commands for robotic navigation and mobile manipulation
Stefanie Tellex, Thomas Kollar, Steven Dickerson, Matthew R Walter, Ashis Gopal Banerjee, Seth Teller, and Nicholas Roy · 2011
Cited alongside, same era.
Grounded language learning from video described with sentences
Haonan Yu and Jeffrey Mark Siskind · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
VQA: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Cited alongside, same era.
Hierarchical question-image co-attention for visual question answering
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh · 2016
Later among the works it cites.
Grounding of textual phrases in images by reconstruction
Anna Rohrbach, Marcus Rohrbach, Ronghang Hu, Trevor Darrell, and Bernt Schiele · 2016
Later among the works it cites.
Prioritized experience replay
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2016
Later among the works it cites.
Mazebase: A sandbox for learning from games
Sainbayar Sukhbaatar, Arthur Szlam, Gabriel Synnaeve, Soumith Chintala, and Rob Fergus · 2016
Later among the works it cites.
Multicore-TSNE
Dmitry Ulyanov · 2016
Later among the works it cites.
Stacked attention networks for image question answering
Zichao Yang, Xiaodong He, Jianfeng Gao, Li Deng, and Alex Smola · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Cited alongside, same era.
Are you talking to a machine? dataset and methods for multilingual image question
Haoyuan Gao, Junhua Mao, Jie Zhou, Zhiheng Huang, Lei Wang, and Wei Xu · 2015
Cited alongside, same era.
Learning like a child: Fast novel visual concept learning from sentence descriptions of images
Junhua Mao, Xu Wei, Yi Yang, Jiang Wang, Zhiheng Huang, and Alan L Yuille · 2015
Cited alongside, same era.
A roadmap towards machine intelligence
Tomas Mikolov, Armand Joulin, and Marco Baroni · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Exploring models and data for image question answering
Mengye Ren, Ryan Kiros, and Richard Zemel · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Cited alongside, same era.
Later among the works it cites.
Bottom-up and top-down attention for image captioning and VQA
Peter Anderson, Xiaodong He, Chris Buehler, Damien Teney, Mark Johnson, Stephen Gould, and Lei Zhang · 2017
Later among the works it cites.
Modular multitask reinforcement learning with policy sketches
Jacob Andreas, Dan Klein, and Sergey Levine · 2017
Later among the works it cites.
Driving under the influence (of language)
Daniel Paul Barrett, Scott Alan Bronikowski, Haonan Yu, and Jeffrey Mark Siskind · 2017
Later among the works it cites.
Modulating early visual processing by language
Harm de Vries, Florian Strub, Jérémie Mary, Hugo Larochelle, Olivier Pietquin, and Aaron C. Courville · 2017
Later among the works it cites.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojciech Marian Czarnecki, Max Jaderberg, Denis Teplyashin, Marcus Wainwright, Chris Apps, Demis Hassabis, and Phil Blunsom · 2017
Later among the works it cites.
Learning to reason: End-to-end module networks for visual question answering
Ronghang Hu, Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Kate Saenko · 2017
Later among the works it cites.
Where is misty? interpreting spatial descriptors by modeling regions in space
Nikita Kitaev and Dan Klein · 2017
Later among the works it cites.
Natural language does not emerge ’naturally’ in multi-agent dialog
Satwik Kottur, José M. F. Moura, Stefan Lee, and Dhruv Batra · 2017
Later among the works it cites.
Mapping instructions and visual observations to actions with reinforcement learning
Dipendra Misra, John Langford, and Yoav Artzi · 2017
Later among the works it cites.
Zero-shot task generalization with multi-task deep reinforcement learning
Junhyuk Oh, Satinder P. Singh, Honglak Lee, and Pushmeet Kohli · 2017
Later among the works it cites.
Gated-attention architectures for task-oriented language grounding
Devendra Singh Chaplot, Kanthashree Mysore Sathyendra, Rama Kumar Pasumarthi, Dheeraj Rajagopal, and Ruslan Salakhutdinov · 2018
Closest in time.