Fetching the paper…
Reading the bibliography…
Learning from prior tasks and transferring that experience to improve future performance is critical for building lifelong learning agents.
Some aspects of the sequential design of experiments
Robbins, H. (1952) · 1952
Earlier work this paper cites.
Perturbation bounds in connection with singular value decomposition
Wedin, P. (1972) · 1972
Earlier work this paper cites.
Matrix perturbation theory
Stewart, G. W. and Sun, J.-g. (1990) · 1990
Earlier work this paper cites.
Finite-time analysis of the multi-armed bandit problem
Auer, P., Cesa-Bianchi, N., and Fischer, P. (2002) · 2002
Earlier work this paper cites.
Prediction, Learning, and Games
Cesa-Bianchi, N. and Lugosi, G. (2006) · 2006
Earlier work this paper cites.
Online multitask learning
Dekel, O., Long, P. M., and Singer, Y. (2006) · 2006
Earlier work this paper cites.
The epoch-greedy algorithm for multi-armed bandits with side information
Langford, J. and Zhang, T. (2007) · 2007
Cited alongside, same era.
Online multi-task learning with hard constraints
Lugosi, G., Papaspiliopoulos, O., and Stoltz, G. (2009) · 2009
Cited alongside, same era.
Linear algorithms for online multitask classification
Cavallanti, G., Cesa-Bianchi, N., and Gentile, C. (2010) · 2010
Cited alongside, same era.
A survey on transfer learning
Pan, S. J. and Yang, Q. (2010) · 2010
Cited alongside, same era.
On upper-confidence bound policies for switching bandit problems
Garivier, A. and Moulines, E. (2011) · 2011
Cited alongside, same era.
A spectral algorithm for latent dirichlet allocation
Anandkumar, A., Foster, D. P., Hsu, D., Kakade, S., and Liu, Y.-K. (2012a)
Cited in the paper.
Transfer in reinforcement learning: a framework and a survey
Lazaric, A. (2011) · 2011
Later among the works it cites.
Online learning of multiple tasks and their relationships
Saha, A., Rai, P., Daumé III, H., and Venkatasubramanian, S. (2011) · 2011
Later among the works it cites.
Contextual bandit learning with predictable rewards
Agarwal, A., Dudík, M., Kale, S., Langford, J., and Schapire, R. E. (2012) · 2012
Later among the works it cites.
Optimization for Machine Learning
Audibert, J.-Y., Bubeck, S., and Munos, R. (2012) · 2012
Later among the works it cites.
Directed exploration in reinforcement learning with transferred knowledge
Mann, T. A. and Choe, Y. (2012) · 2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Anandkumar, A., Ge, R., Hsu, D., Kakade, S. M., and Telgarsky, M. (2012b)
Cited in the paper.
A method of moments for mixture models and hidden markov models
Anandkumar, A., Hsu, D., and Kakade, S. M. (2012c)
Cited in the paper.