Fetching the paper…
Reading the bibliography…
Many control tasks exhibit similar dynamics that can be modeled as having common latent structure.
Dynamic Programming
Richard Bellman · 1957
Earlier work this paper cites.
Markov decision processes: Discrete stochastic dynamic programming
Martin L Puterman · 1995
Earlier work this paper cites.
Neuro-Dynamic Programming
Dimitri P. Bertsekas and John N. Tsitsiklis · 1996
Earlier work this paper cites.
Integral probability metrics and their generating classes of functions
Alfred Müller · 1997
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L. Littman, and Anthony R. Cassandra · 1998
Earlier work this paper cites.
Generalized Hidden Parameter MDPs Transferable Model-based RL in a Handful of Trials
Christian F. Perez, Felipe Petroski Such, and Theofanis Karaletsos · 2002
Earlier work this paper cites.
Equivalence notions and model minimization in markov decision processes
Robert Givan, Thomas Dean, and Matthew Greig · 2003
Earlier work this paper cites.
Metrics for finite markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2004
Earlier work this paper cites.
Error bounds for approximate value iteration
Rémi Munos · 2005
Earlier work this paper cites.
Towards a unified theory of state abstraction for mdps
Lihong Li, Thomas J. Walsh, and Michael L. Littman · 2006
Earlier work this paper cites.
Learning invariant representations for reinforcement learning without reconstruction
Amy Zhang, Rowan McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine · 2006
Earlier work this paper cites.
Using bisimulation for policy transfer in mdps
Pablo Samuel Castro and Doina Precup · 2010
Earlier work this paper cites.
Bisimulation metrics for continuous markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2011
Earlier work this paper cites.
Sample complexity of multi-task reinforcement learning
Emma Brunskill and Lihong Li · 2013
Earlier work this paper cites.
Finale Doshi-Velez and George Konidaris · 2013
Cited alongside, same era.
Online multi-task learning for policy gradient methods
Haitham Bou Ammar, Eric Eaton, Paul Ruvolo, and Matthew Taylor · 2014
Cited alongside, same era.
Sparse multi-task reinforcement learning
Daniele Calandriello, Alessandro Lazaric, and Marcello Restelli · 2014
Cited alongside, same era.
Abstraction selection in model-based reinforcement learning
Nan Jiang, Alex Kulesza, and Satinder Singh · 2015
Cited alongside, same era.
The benefit of multitask representation learning
Andreas Maurer, Massimiliano Pontil, and Bernardino Romera-Paredes · 2016
Cited alongside, same era.
Actor-mimic: Deep multitask and transfer reinforcement learning
Provably efficient RL with rich observations via latent state decoding
Simon S. Du, Akshay Krishnamurthy, Nan Jiang, Alekh Agarwal, Miroslav Dudík, and John Langford · 2019
Later among the works it cites.
Deepmdp: Learning continuous latent space models for representation learning
Carles Gelada, Saurabh Kumar, Jacob Buckman, Ofir Nachum, and Marc G Bellemare · 2019
Later among the works it cites.
A model-based approach for sample-efficient multi-task reinforcement learning, 2019
Nicholas C. Landolfi, Garrett Thomas, and Tengyu Ma · 2019
Later among the works it cites.
Algorithmic framework for model-based deep reinforcement learning with theoretical guarantees
Yuping Luo, Huazhe Xu, Yuanzhi Li, Yuandong Tian, Trevor Darrell, and Tengyu Ma · 2019
Later among the works it cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
Kate Rakelly, Aurick Zhou, Chelsea Finn, Sergey Levine, and Deirdre Quillen · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Emilio Parisotto, Jimmy Ba, and Ruslan Salakhutdinov · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Distral: Robust multitask reinforcement learning
Yee Teh, Victor Bapst, Wojciech M. Czarnecki, John Quan, James Kirkpatrick, Raia Hadsell, Nicolas Heess, and Razvan Pascanu · 2017
Cited alongside, same era.
State abstractions for lifelong reinforcement learning
David Abel, Dilip Arumugam, Lucas Lehnert, and Michael Littman · 2018
Cited alongside, same era.
Meta-learning by adjusting priors based on extended PAC-Bayes theory
Ron Amit and Ron Meir · 2018
Cited alongside, same era.
Gradnorm: Gradient normalization for adaptive loss balancing in deep multitask networks
Zhao Chen, Vijay Badrinarayanan, Chen-Yu Lee, and Andrew Rabinovich · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Later among the works it cites.
ProMP: Proximal meta-policy search
Jonas Rothfuss, Dennis Lee, Ignasi Clavera, Tamim Asfour, and Pieter Abbeel · 2019
Later among the works it cites.
Improving sample efficiency in model-free reinforcement learning from images
Denis Yarats, Amy Zhang, Ilya Kostrikov, Brandon Amos, Joelle Pineau, and Rob Fergus · 2019
Later among the works it cites.
Learning causal state representations of partially observable environments
Amy Zhang, Zachary C. Lipton, Luis Pineda, Kamyar Azizzadenesheli, Anima Anandkumar, Laurent Itti, Joelle Pineau, and Tommaso Furlanello · 2019
Later among the works it cites.
Sharing knowledge in multi-task deep reinforcement learning
Carlo D’Eramo, Davide Tateo, Andrea Bonarini, Marcello Restelli, and Jan Peters · 2020
Closest in time.
Temple: Learning template of transitions for sample efficient multi-task rl, 2020
Yanchao Sun, Xiangyu Yin, and Furong Huang · 2020
Closest in time.
Sequential transfer in reinforcement learning with a generative model
Andrea Tirinzoni, Riccardo Poiani, and Marcello Restelli · 2020
Closest in time.
Meta-learning without memorization
Mingzhang Yin, George Tucker, Mingyuan Zhou, Sergey Levine, and Chelsea Finn · 2020
Closest in time.
Gradient surgery for multi-task learning, 2020
Tianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine, Karol Hausman, and Chelsea Finn · 2020
Closest in time.