Fetching the paper…
Reading the bibliography…
In this paper, we explore the utilization of natural language to drive transfer for reinforcement learning (RL).
Improving generalization for temporal difference learning: The successor representation
Dayan, P. (1993) · 1993
Earlier work this paper cites.
Multitask learning
Caruana, R. (1997) · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Introduction to reinforcement learning
Sutton, R. S., and Barto, A. G. (1998) · 1998
Earlier work this paper cites.
Value-function-based transfer for reinforcement learning using structure mapping
Liu, Y., and Stone, P. (2006) · 1999
Earlier work this paper cites.
A framework for transfer in reinforcement learning
Konidaris, G. D. (2006) · 2006
Earlier work this paper cites.
Building portable options: Skill transfer in reinforcement learning.
Konidaris, G., and Barto, A. G. (2007) · 2007
Earlier work this paper cites.
Cross-domain transfer for reinforcement learning
Taylor, M. E., and Stone, P. (2007) · 2007
Earlier work this paper cites.
Transfer learning via inter-task mappings for temporal difference learning
Taylor, M. E., Stone, P., and Liu, Y. (2007) · 2007
Earlier work this paper cites.
Transferring instances for model-based reinforcement learning
Taylor, M. E., Jong, N. K., and Stone, P. (2008) · 2008
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Taylor, M. E., and Stone, P. (2009) · 2009
Earlier work this paper cites.
Reading between the lines: Learning to map high-level instructions to commands
Branavan, S., Zettlemoyer, L. S., and Barzilay, R. (2010) · 2010
Earlier work this paper cites.
Toward understanding natural language directions
Kollar, T., Tellex, S., Roy, D., and Roy, N. (2010) · 2010
Earlier work this paper cites.
Learning to follow navigational directions
Vogel, A., and Jurafsky, D. (2010) · 2010
Earlier work this paper cites.
Non-linear monte-carlo search in Civilization II.
Branavan, S., Silver, D., and Barzilay, R. (2011) · 2011
Earlier work this paper cites.
Amazon’s mechanical turk
Buhrmester, M., Kwang, T., and Gosling, S. D. (2011) · 2011
Earlier work this paper cites.
Learning to win by reading manuals in a monte-carlo framework
Branavan, S., Silver, D., and Barzilay, R. (2012) · 2012
Earlier work this paper cites.
Transfer in reinforcement learning via shared features
Konidaris, G., Scheidwasser, I., and Barto, A. (2012) · 2012
Earlier work this paper cites.
Transferring expectations in model-based reinforcement learning
Nguyen, T., Silander, T., and Leong, T. Y. (2012) · 2012
Earlier work this paper cites.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Artzi, Y., and Zettlemoyer, L. (2013) · 2013
Earlier work this paper cites.
Pomdp-based dialogue manager adaptation to extended domains
Gašic, M., et al. (2013) · 2013
Earlier work this paper cites.
Learning to parse natural language commands to a robot control system
Matuszek, C., Herbst, E., Zettlemoyer, L., and Fox, D. (2013) · 2013
Cited alongside, same era.
Efficient estimation of word representations in vector space
Mikolov, T., Chen, K., Corrado, G., and Dean, J. (2013) · 2013
Cited alongside, same era.
Ella: An efficient lifelong learning algorithm
Ruvolo, P., and Eaton, E. (2013) · 2013
Cited alongside, same era.
A video game description language for model-based or interactive learning
Schaul, T. (2013) · 2013
Cited alongside, same era.
Zero-shot learning through cross-modal transfer
Socher, R., Ganjoo, M., Manning, C. D., and Ng, A. (2013) · 2013
Cited alongside, same era.
Learning semantic maps from natural language descriptions.
Walter, M. R., Hemachandra, S., Homberg, B., Tellex, S., and Teller, S. (2013) · 2013
End-to-end training of deep visuomotor policies
Levine, S., Finn, C., Darrell, T., and Abbeel, P. (2016) · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Mnih, V., Badia, A. P., Mirza, M., Graves, A., Lillicrap, T. P., Harley, T., Silver, D., and Kavukcuoglu, K. (2016) · 2016
Later among the works it cites.
Actor-mimic: Deep multitask and transfer reinforcement learning
Parisotto, E., Ba, J. L., and Salakhutdinov, R. (2016) · 2016
Later among the works it cites.
General video game ai: Competition, challenges and opportunities
Perez-Liebana, D., Samothrakis, S., Togelius, J., Schaul, T., and Lucas, S. M. (2016) · 2016
Later among the works it cites.
Rusu, A. A., Rabinowitz, N. C., Desjardins, G., Soyer, H., Kirkpatrick, J., Kavukcuoglu, K., Pascanu, R., and Hadsell, R. (2016) · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Online multi-task learning for policy gradient methods
Ammar, H. B., Eaton, E., Ruvolo, P., and Taylor, M. (2014) · 2014
Cited alongside, same era.
Learning spatial-semantic representations from natural language descriptions and scene classifications
Hemachandra, S., Walter, M. R., Tellex, S., and Teller, S. (2014) · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D., and Ba, J. (2014) · 2014
Cited alongside, same era.
Alignment-based compositional semantics for instruction following
Andreas, J., and Klein, D. (2015) · 2015
Cited alongside, same era.
Unsupervised domain adaptation for zero-shot learning
Kodirov, E., Xiang, T., Fu, Z., and Gong, S. (2015) · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., and Hassabis, D. (2015) · 2015
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al. (2016) · 2016
Later among the works it cites.
Value iteration networks
Tamar, A., Wu, Y., Thomas, G., Levine, S., and Abbeel, P. (2016) · 2016
Later among the works it cites.
Modular multitask reinforcement learning with policy sketches
Andreas, J., Klein, D., and Levine, S. (2017) · 2017
Closest in time.
Successor features for transfer in reinforcement learning
Barreto, A., Dabney, W., Munos, R., Hunt, J. J., Schaul, T., van Hasselt, H. P., and Silver, D. (2017) · 2017
Closest in time.
Policy reuse in deep reinforcement learning.
Glatt, R., and Costa, A. H. R. (2017) · 2017
Closest in time.
Learning invariant feature spaces to transfer skills with reinforcement learning
Gupta, A., Devin, C., Liu, Y., Abbeel, P., and Levine, S. (2017) · 2017
Closest in time.
Guiding reinforcement learning exploration using natural language
Harrison, B., Ehsan, U., and Riedl, M. O. (2017) · 2017
Closest in time.
Schema networks: Zero-shot transfer with a generative causal model of intuitive physics
Kansky, K., Silver, T., Mély, D. A., Eldawy, M., Lázaro-Gredilla, M., Lou, X., Dorfman, N., Sidor, S., Phoenix, S., and George, D. (2017) · 2017
Closest in time.
Zero-shot task generalization with multi-task deep reinforcement learning
Oh, J., Singh, S., Lee, H., and Kohli, P. (2017) · 2017
Closest in time.
Rajendran, J., Lakshminarayanan, A., Khapra, M. M., Prasanna, P., and Ravindran, B. (2017)
2017
Closest in time.
Domain randomization for transferring deep neural networks from simulation to the real world
Tobin, J., et al. (2017) · 2017
Closest in time.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I. (2017) · 2017
Closest in time.
Knowledge transfer for deep reinforcement learning with hierarchical experience replay
Yin, H., and Pan, S. J. (2017) · 2017
Closest in time.
Cross-domain transfer in reinforcement learning using target apprentice
Joshi, G., and Chowdhary, G. (2018) · 2018
Closest in time.
Joint dictionaries for zero-shot learning
Kolouri, S., Rostami, M., Owechko, Y., and Kim, K. (2018) · 2018
Closest in time.