Fetching the paper…
Reading the bibliography…
We study the problem of representation learning in goal-conditioned hierarchical reinforcement learning.
Approximations of dynamic programs, i
Ward Whitt · 1978
Earlier work this paper cites.
Adaptive aggregation methods for infinite horizon dynamic programming
Dimitri P Bertsekas and David Alfred Castanon · 1989
Earlier work this paper cites.
Aggregation and disaggregation techniques and methodology in optimization
David F Rogers, Robert D Plante, Richard T Wong, and James R Evans · 1991
Earlier work this paper cites.
Feudal reinforcement learning
Peter Dayan and Geoffrey E Hinton · 1993
Earlier work this paper cites.
Planning simple trajectories using neural subgoal generators
Jürgen Schmidhuber and Reiner Wahnsiedler · 1993
Earlier work this paper cites.
Model minimization in markov decision processes
Thomas Dean and Robert Givan · 1997
Earlier work this paper cites.
Reinforcement learning with hierarchies of machines
Ronald Parr and Stuart J Russell · 1998
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Earlier work this paper cites.
Hierarchical reinforcement learning with the maxq value function decomposition
Thomas G Dietterich · 2000
Earlier work this paper cites.
Model minimization in hierarchical reinforcement learning
Balaraman Ravindran and Andrew G Barto · 2002
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
Andrew G Barto and Sridhar Mahadevan · 2003
Earlier work this paper cites.
Towards a unified theory of state abstraction for mdps
Lihong Li, Thomas J Walsh, and Michael L Littman · 2006
Cited alongside, same era.
Performance loss bounds for approximate value iteration with state aggregation
Benjamin Van Roy · 2006
Cited alongside, same era.
Transfer via soft homomorphisms
Jonathan Sorg and Satinder Singh · 2009
Cited alongside, same era.
An information-theoretic approach to curiosity-driven reinforcement learning
Susanne Still and Doina Precup · 2012
Cited alongside, same era.
Motor primitive discovery
Philip S Thomas and Andrew G Barto · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Cited alongside, same era.
Early visual concept learning with unsupervised deep learning
Irina Higgins, Loic Matthey, Xavier Glorot, Arka Pal, Benigno Uria, Charles Blundell, Shakir Mohamed, and Alexander Lerchner · 2016
Later among the works it cites.
Near optimal behavior via approximate state abstraction
David Abel, D Ellis Hershkowitz, and Michael L Littman · 2017
Later among the works it cites.
Constrained policy optimization
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel · 2017
Later among the works it cites.
The option-critic architecture
Pierre-Luc Bacon, Jean Harb, and Doina Precup · 2017
Later among the works it cites.
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
Shixiang Gu, Ethan Holly, Timothy Lillicrap, and Sergey Levine · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Diederik P Kingma and Max Welling · 2013
Cited alongside, same era.
Pac-inspired option discovery in lifelong reinforcement learning
Emma Brunskill and Lihong Li · 2014
Cited alongside, same era.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Cited alongside, same era.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Cited alongside, same era.
Embed to control: A locally linear latent dynamics model for control from raw images
Manuel Watter, Jost Springenberg, Joschka Boedecker, and Martin Riedmiller · 2015
Cited alongside, same era.
Later among the works it cites.
Andrew Levy, Robert Platt, and Kate Saenko · 2017
Later among the works it cites.
Feudal networks for hierarchical reinforcement learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul, Nicolas Heess, Max Jaderberg, David Silver, and Koray Kavukcuoglu · 2017
Later among the works it cites.
Learning deep representations by mutual information estimation and maximization
R Devon Hjelm, Alex Fedorov, Samuel Lavoie-Marchildon, Karan Grewal, Adam Trischler, and Yoshua Bengio · 2018
Closest in time.
Mine: Mutual information neural estimation
Mohamed Ishmael Belghazi, Aristide Baratin, Sai Rajeswar, Sherjil Ozair, Yoshua Bengio, Aaron Courville, and R Devon Hjelm · 2018
Closest in time.
Data-efficient hierarchical reinforcement learning
Ofir Nachum, Shane Gu, Honglak Lee, and Sergey Levine · 2018
Closest in time.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Closest in time.