Fetching the paper…
Reading the bibliography…
To be successful in real-world tasks, Reinforcement Learning (RL) needs to exploit the compositional, relational, and hierarchical structure of the world, and learn to transfer it to the task at hand.
Human behavior and the principle of least effort
George Kingsley Zipf · 1949
Earlier work this paper cites.
A synopsis of linguistic theory, 1957
John R Firth · 1957
Earlier work this paper cites.
Zork I, 1980
Infocom · 1980
Earlier work this paper cites.
The development of categorization in the second year and its relation to other cognitive and linguistic developments
Alison Gopnik and Andrew Meltzoff · 1987
Earlier work this paper cites.
Indexing by latent semantic analysis
Scott Deerwester, Susan T Dumais, George W Furnas, Thomas K Landauer, and Richard Harshman · 1990
Earlier work this paper cites.
Handbook of Intelligent Control: Neural, Fuzzy, and Adaptative Approaches
David Ashley White and Donald A Sofge · 1992
Earlier work this paper cites.
Pragmatics: An Introduction
J. Mey · 1993
Earlier work this paper cites.
Temporal Difference Learning and TD-Gammon
Gerald Tesauro · 1995
Earlier work this paper cites.
being-in-the-world
Mark A DePristo and Robert Zubek · 2001
Earlier work this paper cites.
Optimizing dialogue management with reinforcement learning: Experiments with the njfun system
Satinder Singh, Diane Litman, Michael Kearns, and Marilyn Walker · 2002
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
Andrew G Barto and Sridhar Mahadevan · 2003
Earlier work this paper cites.
Guiding a Reinforcement Learner with Natural Language Advice: Initial Results in RoboCup Soccer
Gregory Kuhlmann, Peter Stone, Raymond Mooney, and Jude Shavlik · 2004
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
Matt MacMahon, Brian Stankiewicz, and Benjamin Kuipers · 2006
Earlier work this paper cites.
Open information extraction from the web
Michele Banko, Michael J Cafarella, Stephen Soderland, Matthew Broadhead, and Oren Etzioni · 2007
Earlier work this paper cites.
Core knowledge
Elizabeth Spelke and Katherine D Kinzler · 2007
Earlier work this paper cites.
Maximum Entropy Inverse Reinforcement Learning
Brian D Ziebart, Andrew Maas, J Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
Reading to learn: constructing features from semantic abstracts
Jacob Eisenstein, James Clarke, Dan Goldwasser, and Dan Roth · 2009
Earlier work this paper cites.
Reading between the lines: Learning to map high-level instructions to commands
S. R. K. Branavan, Luke S Zettlemoyer, and Regina Barzilay · 2010
Earlier work this paper cites.
Toward understanding natural language directions
Thomas Kollar, Stefanie Tellex, Deb Roy, and Nicholas Roy · 2010
Earlier work this paper cites.
Apprenticeship learning about multiple intentions
Monica Babes, Vukosi Marivate, Kaushik Subramanian, and Michael L Littman · 2011
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
David L Chen and Raymond J Mooney · 2011
Earlier work this paper cites.
Cognitive effects of language on human navigation
Anna Shusterman, Sang Ah Lee, and Elizabeth Spelke · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
Stefanie Tellex, Thomas Kollar, Steven Dickerson, Matthew R Walter, Ashis Gopal Banerjee, Seth Teller, and Nicholas Roy · 2011
Earlier work this paper cites.
Learning to Win by Reading Manuals in a Monte-Carlo Framework
S. R. K. Branavan, David Silver, and Regina Barzilay · 2012
Earlier work this paper cites.
Weakly Supervised Learning of Semantic Parsers for Mapping Instructions to Actions
Yoav Artzi and Luke Zettlemoyer · 2013
Earlier work this paper cites.
DeViSE: A Deep Visual-Semantic Embedding Model
Andrea Frome, Greg S Corrado, Jon Shlens, Samy Bengio, Jeff Dean, Marc Aurelio Ranzato, and Tomas Mikolov · 2013
Earlier work this paper cites.
Distributed Representations of Words and Phrases and their Compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Zero-shot Learning Through Cross-modal Transfer
Richard Socher, Milind Ganjoo, Christopher D. Manning, and Andrew Y. Ng · 2013
Earlier work this paper cites.
Alignment-based compositional semantics for instruction following
Jacob Andreas and Dan Klein · 2015
Earlier work this paper cites.
VQA: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Cited alongside, same era.
Question answering systems: survey and trends
Abdelghani Bouziane, Djelloul Bouchiha, Noureddine Doumi, and Mimoun Malki · 2015
Cited alongside, same era.
Grounding english commands to reward functions
James MacGlashan, Monica Babes-Vroman, Marie desJardins, Michael L. Littman, Smaranda Muresan, Shawn Squire, Stefanie Tellex, Dilip Arumugam, and Lei Yang · 2015
Cited alongside, same era.
Language understanding for text-based games using deep reinforcement learning
Karthik Narasimhan, Tejas D. Kulkarni, and Regina Barzilay · 2015
Cited alongside, same era.
Neural module networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Cited alongside, same era.
Natural language communication with robots
Yonatan Bisk, Deniz Yuret, and Daniel Marcu · 2016
Neural Modular Control for Embodied Question Answering
Abhishek Das, Georgia Gkioxari, Stefan Lee, Devi Parikh, and Dhruv Batra · 2018
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Later among the works it cites.
Wizard of wikipedia: Knowledge-powered conversational agents
Emily Dinan, Stephen Roller, Kurt Shuster, Angela Fan, Michael Auli, and Jason Weston · 2018
Later among the works it cites.
Iqa: Visual question answering in interactive environments
Daniel Gordon, Aniruddha Kembhavi, Mohammad Rastegari, Joseph Redmon, Dieter Fox, and Ali Farhadi · 2018
Later among the works it cites.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Generative Adversarial Imitation Learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
The malmo platform for artificial intelligence experimentation
Matthew Johnson, Katja Hofmann, Tim Hutton, and David Bignell · 2016
Cited alongside, same era.
Listen, Attend, and Walk: Neural Mapping of Navigational Instructions to Action Sequences
Hongyuan Mei, Mohit Bansal, and Matthew R. Walter · 2016
Cited alongside, same era.
Learning Language Games through Interaction
Sida I Wang, Percy Liang, and Christopher D Manning · 2016
Cited alongside, same era.
Modular Multitask Reinforcement Learning with Policy Sketches
Jacob Andreas, Dan Klein, and Sergey Levine · 2017
Cited alongside, same era.
A Brief Survey of Deep Reinforcement Learning
Kai Arulkumaran, Marc Peter Deisenroth, Miles Brundage, and Anil Anthony Bharath · 2017
Cited alongside, same era.
Representation learning for grounded spatial reasoning
Michael Janner, Karthik Narasimhan, and Regina Barzilay · 2018
Later among the works it cites.
Flipdial: A generative model for two-way visual dialogue
Daniela Massiceti, N Siddharth, Puneet K Dokania, and Philip HS Torr · 2018
Later among the works it cites.
Grounding Language for Transfer in Deep Reinforcement Learning
Karthik Narasimhan, Regina Barzilay, and Tommi Jaakkola · 2018
Later among the works it cites.
An algorithmic perspective on imitation learning
Takayuki Osa, Joni Pajarinen, Gerhard Neumann, J Andrew Bagnell, Pieter Abbeel, Jan Peters, et al · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer · 2018
Later among the works it cites.
Dissecting contextual word embeddings: Architecture and representation
Matthew E. Peters, Mark Neumann, Luke Zettlemoyer, and Wen-tau Yih · 2018
Later among the works it cites.
Hierarchical and Interpretable Skill Acquisition in Multi-task Reinforcement Learning
Tianmin Shu, Caiming Xiong, and Richard Socher · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Deep reinforcement learning for general video game ai
Ruben Rodriguez Torrado, Philip Bontrager, Julian Togelius, Jialin Liu, and Diego Perez-Liebana · 2018
Later among the works it cites.
Chalet: Cornell house agent learning environment
Claudia Yan, Dipendra Misra, Andrew Bennnett, Aaron Walsman, Yonatan Bisk, and Yoav Artzi · 2018
Later among the works it cites.
Interactive Grounded Language Acquisition and Generalization in a 2d World
Haonan Yu, Haichao Zhang, and Wei Xu · 2018
Later among the works it cites.
Counting to Explore and Generalize in Text-based Games
Xingdi Yuan, Marc-Alexandre Côté, Alessandro Sordoni, Romain Laroche, Remi Tachet des Combes, Matthew Hausknecht, and Adam Trischler · 2018
Later among the works it cites.
SWAG: A large-scale adversarial dataset for grounded commonsense inference
Rowan Zellers, Yonatan Bisk, Roy Schwartz, and Yejin Choi · 2018
Later among the works it cites.
Learning to Generalize from Sparse and Underspecified Rewards
Rishabh Agarwal, Chen Liang, Dale Schuurmans, and Mohammad Norouzi · 2019
Closest in time.
Learning to Understand Goal Specifications by Modelling Reward
Dzmitry Bahdanau, Felix Hill, Jan Leike, Edward Hughes, Arian Hosseini, Pushmeet Kohli, and Edward Grefenstette · 2019
Closest in time.
BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, and Yoshua Bengio · 2019
Closest in time.
From Language to Goals: Inverse Reinforcement Learning for Vision-Based Instruction Following
Justin Fu, Anoop Korattikara, Sergey Levine, and Sergio Guadarrama · 2019
Closest in time.
Assessing BERT’s Syntactic Abilities
Yoav Goldberg · 2019
Closest in time.
Using Natural Language for Reward Shaping in Reinforcement Learning
Prasoon Goyal, Scott Niekum, and Raymond J. Mooney · 2019
Closest in time.
Hierarchical decision making by generating and following natural language instructions
Hengyuan Hu, Denis Yarats, Qucheng Gong, Yuandong Tian, and Mike Lewis · 2019
Closest in time.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Closest in time.
What do you learn from context? probing for sentence structure in contextualized word representations
Ian Tenney, Patrick Xia, Berlin Chen, Alex Wang, Adam Poliak, R Thomas McCoy, Najoung Kim, Benjamin Van Durme, Sam Bowman, Dipanjan Das, and Ellie Pavlick · 2019
Closest in time.
Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
Xin Wang, Qiuyuan Huang, Asli Çelikyilmaz, Jianfeng Gao, Dinghan Shen, Yuan-Fang Wang, William Yang Wang, and Lei Zhang · 2019
Closest in time.