Fetching the paper…
Reading the bibliography…
Control applications often feature tasks with similar, but not identical, dynamics.
Reinforcement Learning: An Introduction
R.S. Sutton and A.G. Barto · 1998
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
L.P. Kaelbling, M.L. Littman, and A.R. Cassandra · 1998
Earlier work this paper cites.
Gaussian process latent variable models for visualisation of high dimensional data
N.D. Lawrence · 2004
Earlier work this paper cites.
Gaussian Processes for Machine Learning
C.E. Rasmussen and C.K.I Williams · 2005
Earlier work this paper cites.
Semiparametric latent factor models
Y. W. Teh, M. Seeger, and M. I. Jordan · 2005
Earlier work this paper cites.
Sparse Gaussian processes using pseudo-inputs
E. Snelson and Z. Ghahramani · 2005
Earlier work this paper cites.
Scaling up POMDPs for dialogue management: The “summary POMDP” method
J. Williams and S. Young · 2005
Earlier work this paper cites.
An analytic solution to discrete Bayesian reinforcement learning
P. Poupart, N. Vlassis, J. Hoey, and K. Regan · 2006
Earlier work this paper cites.
Proto-transfer learning in Markov decision processes using spectral methods
K. Ferguson and S. Mahadevan · 2006
Earlier work this paper cites.
Representation transfer for reinforcement learning
M.E. Taylor and P. Stone · 2007
Cited alongside, same era.
Bayesian reinforcement learning in continuous POMDPs with application to robot navigation
S. Ross, B. Chaib-draa, and J. Pineau · 2008
Cited alongside, same era.
Sarsop: Efficient point-based POMDP planning by approximating optimally reachable belief spaces
H. Kurniawati, D. Hsu, and W.S. Lee · 2008
Cited alongside, same era.
Transfer of task representation in reinforcement learning using policy-based proto-value functions
E. Ferrante, A. Lazaric, and M. Restelli · 2008
Cited alongside, same era.
Variational Learning of Inducing Variables in Sparse Gaussian Processes
M. Titsias · 2009
Cited alongside, same era.
Monte-Carlo planning in large POMDPs
D. Silver and J. Veness · 2010
Value function approximation in reinforcement learning using the Fourier basis
G.D. Konidaris, S. Osentoski, and P.S. Thomas · 2011
Later among the works it cites.
Computationally efficient convolved multiple output Gaussian processes
M.A. Álvarez and N.D. Lawrence · 2011
Later among the works it cites.
Variational Gaussian process dynamical systems
A.C. Damianou, M.K. Titsias, and N.D. Lawrence · 2011
Later among the works it cites.
Reinforcement learning to adjust parametrized motor primitives to new situations
J. Kober, A. Wilhelm, E. Oztop, and J. Peters · 2012
Later among the works it cites.
Learning parameterized skills
B.C. da Silva, G.D. Konidaris, and A.G. Barto · 2012
Later among the works it cites.
Efficient Bayes-adaptive reinforcement learning using sample-based search
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Planning under uncertainty for robotic tasks with mixed observability
S C W Ong and D Hsu · 2010
Cited alongside, same era.
A computational decision theory for interactive assistants
Alan Fern and Prasad Tadepalli · 2010
Cited alongside, same era.
The Indian buffet process: An introduction and review
T.L. Griffiths and Z. Ghahramani · 2011
Cited alongside, same era.
A. Guez, D. Silver, and P. Dayan · 2012
Later among the works it cites.
Transfer learning in sequential decision problems: A hierarchical bayesian approach
A. Wilson, A. Fern, and P. Tadepalli · 2012
Later among the works it cites.
Gaussian process regression networks
A.G. Wilson, D.A. Knowles, and Z. Ghahramani · 2012
Later among the works it cites.
Planning how to learn
H. Y. Bai, D. Hsu, and W. S. Lee · 2013
Closest in time.