Fetching the paper…
Reading the bibliography…
We present algorithms to effectively represent a set of Markov decision processes (MDPs), whose optimal policies have already been learned, by a smaller source subset for lifelong, policy-reuse-based transfer learning in reinforcement learning.
Introductory Real Analysis
Andrey Kolmogorov and Sergei V. Fomin · 1970
Earlier work this paper cites.
Reducibility among combinatorial problems
Richard Karp · 1972
Earlier work this paper cites.
Optimization by simulated annealing
S. Kirkpatrick, C. D. Gelatt, and M. P. Vecchi · 1983
Earlier work this paper cites.
Aggregating strategies
Vladimir Vovk · 1990
Earlier work this paper cites.
Explanation-based neural network learning for robot control
Tom M. Mitchell and Sebastian Thrun · 1993
Earlier work this paper cites.
The weighted majority algorithm
Nick Littlestone and Manfred K. Warmuth · 1994
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Martin L. Puterman · 1994
Earlier work this paper cites.
Learning bayesian networks: The combination of knowledge and statistical data
David Heckerman and David M. Chickering · 1995
Earlier work this paper cites.
Lifelong robot learning
Sebastian Thrun and Tom Mitchell · 1995
Earlier work this paper cites.
Explanation-Based Neural Network Learning: A Lifelong Learning Approach
Sebastian Thrun · 1996
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
A model of inductive bias learning
Jonathan Baxter · 2000
Earlier work this paper cites.
Simulated annealing, algorithms for continuous global optimization: convergence conditions
M. Locatelli · 2000
Earlier work this paper cites.
Learning equivalence classes of bayesian-network structures
David Maxwell Chickering and Craig Boutilier · 2002
Earlier work this paper cites.
Optimal structure identification with greedy search
David Chickering · 2003
Earlier work this paper cites.
SMDP homomorphisms: An algebraic approach to astraction in semi-markov decision processes
Balaraman Ravindran and Andrew G. Barto · 2003
Cited alongside, same era.
Metrics for finite markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2004
Cited alongside, same era.
Task similarity measures for transfer in reinforcement learning task libraries
James L. Carroll and Kevin Seppi · 2005
Cited alongside, same era.
Proto-value functions: Developmental reinforcement learning
Sridhar Mahadevan · 2005
Cited alongside, same era.
Robust control of markov decision processes with uncertain transition matrices
Arnab Nilim and Laurent El Ghaoui · 2005
Cited alongside, same era.
Monte Carlo Statistical Methods
Christian P. Robert and George Casella · 2005
Cited alongside, same era.
Concentration of Measure Inequalities for the Analysis of Randomized Algorithms
Devdatt P. Dubhashi and Alessandro Panconesi · 2009
Later among the works it cites.
Website morphing
John R. Hauser, Glen L. Urban, Guilherme Liberali, , and Michael Braun · 2009
Later among the works it cites.
Markov Chains and Mixing Times
David A. Levin, Yuval Peres, and Elizabeth L. Wilmer · 2009
Later among the works it cites.
Transfer via soft homomorphisms
Jonathan Sorg and Satinder Singh · 2009
Later among the works it cites.
Transfer learning for reinforcement learning domains: A survey
Matthew Taylor and Peter Stone · 2009
Later among the works it cites.
Arbitrarily modulated markov decision processes
Jia Yuan Yu and Shie Mannor · 2009
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Prediction Learning and Games
Nicolo Cesa-Bianchi and Gabor Lugosi · 2006
Cited alongside, same era.
Probabilistic policy reuse in a reinforcement learning agent
Fernando Fernandez, Javier Garcia, and Manuela Veloso · 2006
Cited alongside, same era.
Building portable options: skill transfer in reinforcement learning
George Konidaris and Andrew G. Barto · 2007
Cited alongside, same era.
An experts algorithm for transfer learning
Erik Talvitie and Satinder Singh · 2007
Cited alongside, same era.
Multi-task reinforcement learning: A hierarchical bayesian approach
Aaron Wilson, Alan Fern, Soumya Ray, and Prasad Tadepalli · 2007
Cited alongside, same era.
Transfer of task representation in reinforcement learning using policy-based protovalue functions
Eliseo Ferrante, Alessandro Lazaric, and Marcello Restelli · 2008
Cited alongside, same era.
Jia Yuan Yu, Shie Mannor, and Nahum Shumkin · 2009
Later among the works it cites.
Using bisimulation for policy transfer in mdps
Pablo Samuel Castro and Doina Precup · 2010
Later among the works it cites.
Probabilistic policy reuse for inter-task transfer learning
Fernando Fernandez, Javier Garcia, and Manuela Veloso · 2010
Later among the works it cites.
Transferring from multiple mdps
Allesandro Lazaric and Marcello Restilli · 2011
Later among the works it cites.
Reinforcement learning transfer via sparse coding
Haitham Bou Ammar, Karl Tuyls, Matthew E. Taylor, Kurt Driessen, and Gerhard Weiss · 2012
Later among the works it cites.
Security games with limited surveillance
Bo An, David Kempe, Christopher Kiekintveld, Eric Shieh, Satinder Singh, Milind Tambe, and Yevgeniy Vorobeychik · 2012
Later among the works it cites.
Regret bounds for reinforcement learning with policy advice
Mohammad Gheshlaghi Azar, Alessandro Lazaric, and Emma Brunskill · 2013
Closest in time.
Relativized hierarchical decomposition of markov decision processes
Balaraman Ravindran · 2013
Closest in time.