Fetching the paper…
Reading the bibliography…
We investigate a paradigm in multi-task reinforcement learning (MT-RL) in which an agent is placed in an environment and needs to learn to perform a series of tasks, within this space.
Reinforcement learning: An introduction , volume 1
Sutton, Richard S and Barto, Andrew G · 1998
Earlier work this paper cites.
Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning
Sutton, Richard S, Precup, Doina, and Singh, Satinder · 1999
Earlier work this paper cites.
Hierarchical reinforcement learning with the maxq value function decomposition
Dietterich, Thomas G · 2000
Earlier work this paper cites.
Discovering hierarchy in reinforcement learning with hexq
Hengst, Bernhard · 2002
Earlier work this paper cites.
Learning options in reinforcement learning
Stolle, Martin and Precup, Doina · 2002
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
Barto, Andrew G and Mahadevan, Sridhar · 2003
Earlier work this paper cites.
A framework for learning predictive structures from multiple tasks and unlabeled data
Ando, Rie Kubota and Zhang, Tong · 2005
Earlier work this paper cites.
Tree-based batch mode reinforcement learning
Ernst, Damien, Geurts, Pierre, and Wehenkel, Louis · 2005
Earlier work this paper cites.
Value-iteration based fitted policy iteration: learning with a single trajectory
Antos, András, Szepesvári, Csaba, and Munos, Rémi · 2007
Cited alongside, same era.
Building portable options: Skill transfer in reinforcement learning
Konidaris, George and Barto, Andrew G · 2007
Cited alongside, same era.
Convex multi-task feature learning
Argyriou, Andreas, Evgeniou, Theodoros, and Pontil, Massimiliano · 2008
Cited alongside, same era.
Learning deep architectures for ai
Bengio, Yoshua · 2009
Cited alongside, same era.
A dirty model for multi-task learning
Jalali, Ali, Sanghavi, Sujay, Ruan, Chao, and Ravikumar, Pradeep K · 2010
Cited alongside, same era.
Bayesian multi-task reinforcement learning
Lazaric, Alessandro and Ghavamzadeh, Mohammad · 2010
Cited alongside, same era.
Clustered multi-task learning via alternating structure optimization
Zhou, Jiayu, Chen, Jianhui, and Ye, Jieping · 2011
Later among the works it cites.
Learning incoherent sparse and low-rank patterns from multiple tasks
Chen, Jianhui, Liu, Ji, and Ye, Jieping · 2012
Later among the works it cites.
Transfer in reinforcement learning via shared features
Konidaris, George, Scheidwasser, Ilya, and Barto, Andrew G · 2012
Later among the works it cites.
Transfer in reinforcement learning: a framework and a survey
Lazaric, Alessandro · 2012
Later among the works it cites.
Sparse multi-task reinforcement learning
Calandriello, Daniele, Lazaric, Alessandro, and Restelli, Marcello · 2014
Later among the works it cites.
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Rusu, Andrei A, Veness, Joel, Bellemare, Marc G, Graves, Alex, Riedmiller, Martin, Fidjeland, Andreas K, Ostrovski, Georg, et al · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bayesian multitask inverse reinforcement learning
Dimitrakakis, Christos and Rothkopf, Constantin A · 2011
Cited alongside, same era.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
Sutton, Richard S, Modayil, Joseph, Delp, Michael, Degris, Thomas, Pilarski, Patrick M, White, Adam, and Precup, Doina · 2011
Cited alongside, same era.
Transfer learning for reinforcement learning domains: A survey
Taylor, Matthew E and Stone, Peter
Cited in the paper.
Transfer learning for reinforcement learning domains: A survey
Taylor, Matthew E and Stone, Peter
Cited in the paper.
Later among the works it cites.
Universal value function approximators
Schaul, Tom, Horgan, Daniel, Gregor, Karol, and Silver, David · 2015
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J, Guez, Arthur, Sifre, Laurent, van den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, et al · 2016
Closest in time.