Fetching the paper…
Reading the bibliography…
Building generalizable goal-conditioned agents from rich observations is a key to reinforcement learning (RL) solving real world problems.
Dynamic Programming
Richard Bellman · 1957
Earlier work this paper cites.
Learning by analogy: Formulating and generalizing plans from past experience
Jaime Carbonell · 1983
Earlier work this paper cites.
Bisimulation through probabilistic testing (preliminary report)
K. G. Larsen and A. Skou · 1989
Earlier work this paper cites.
Learning to achieve goals
L P Kaelbling · 1993
Earlier work this paper cites.
Equivalence notions and model minimization in Markov decision processes
Robert Givan, Thomas L. Dean, and Matthew Greig · 2003
Earlier work this paper cites.
Metrics for finite Markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2004
Earlier work this paper cites.
Task similarity measures for transfer in reinforcement learning task libraries
James Carroll and Kevin Seppi · 2005
Earlier work this paper cites.
Probabilistic policy reuse in a reinforcement learning agent
Fernando Fernández and Manuela Veloso · 2006
Earlier work this paper cites.
Towards a unified theory of state abstraction for MDPs
Lihong Li, Thomas J Walsh, and Michael L Littman · 2006
Earlier work this paper cites.
Deep auto-encoder neural networks in reinforcement learning
Sascha Lange and Martin A. Riedmiller · 2010
Earlier work this paper cites.
Bisimulation metrics for continuous Markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2011
Earlier work this paper cites.
Transfer in reinforcement learning via shared features
George Konidaris, Ilya Scheidwasser, and Andrew G. Barto · 2012
Earlier work this paper cites.
Autonomous reinforcement learning on raw visual input data in a real world application
Sascha Lange, Martin Riedmiller, Arne Voigtlander, and Arne Voigtländer · 2012
Earlier work this paper cites.
Distributed Representations of Words and Phrases and their Compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean · 2013
Earlier work this paper cites.
Bisimulation metrics are optimal value functions
Norman Ferns and Doina Precup · 2014
Earlier work this paper cites.
Learning image representations tied to egomotion
Dinesh Jayaraman and Kristen Grauman · 2015
Earlier work this paper cites.
Learning state representations with robotic priors
Rico Jonschkowski and Oliver Brock · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Deep visual analogy-making
Scott E. Reed, Yi Zhang, Yuting Zhang, and Honglak Lee · 2015
Earlier work this paper cites.
Universal Value Function Approximators
Tom Schaul, Daniel Horgan, Karol Gregor, and David Silver · 2015
Cited alongside, same era.
Embed to Control: A Locally Linear Latent Dynamics Model for Control from Raw Images
Manuel Watter, Jost Tobias Springenberg, Joschka Boedecker, and Martin Riedmiller · 2015
Cited alongside, same era.
Measuring the distance between finite markov decision processes
Jinhua Song, Yang Gao, Hao Wang, and Bo An · 2016
Cited alongside, same era.
Hindsight Experience Replay
Marcin Andrychowicz, Filip Wolski, Alex Ray, Jonas Schneider, Rachel Fong, Peter Welinder, Bob Mcgrew, Josh Tobin, Pieter Abbeel, and Wojciech Zaremba · 2017
Cited alongside, same era.
Successor Features for Transfer in Reinforcement Learning
Andre Barreto, Will Dabney, Remi Munos, Jonathan J Hunt, Tom Schaul, Hado P van Hasselt, and David Silver · 2017
Cited alongside, same era.
Pves: Position-velocity encoders for unsupervised learning of structured state representations
Mid-Level Visual Representations Improve Generalization and Sample Efficiency for Learning Visuomotor Policies
Alexander Sax, Bradley Emi, Amir Zamir, Leonidas Guibas, Silvio Savarese, and Jitendra Malik · 2019
Later among the works it cites.
Learning robotic manipulation through visual planning and acting
Angelina Wang, Thanard Kurutach, Pieter Abbeel, and Aviv Tamar · 2019
Later among the works it cites.
Unsupervised Control Through Non-Parametric Discriminative Rewards
David Warde-Farley, Tom Van De Wiele, Tejas Kulkarni, Catalin Ionescu, Steven Hansen, and Mnih Volodymyr · 2019
Later among the works it cites.
Scalable methods for computing state similarity in deterministic Markov decision processes
Pablo Samuel Castro · 2020
Later among the works it cites.
Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement
Ben Eysenbach, Xinyang Geng, Sergey Levine, and Russ R Salakhutdinov · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rico Jonschkowski, Roland Hafner, Jonathan Scholz, and Martin Riedmiller · 2017
Cited alongside, same era.
The kinetics human action video dataset
Will Kay, João Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, Mustafa Suleyman, and Andrew Zisserman · 2017
Cited alongside, same era.
Grasp2vec: Learning object representations from self-supervised grasping
Eric Jang, Coline Devin, Vincent Vanhoucke, and Sergey Levine · 2018
Cited alongside, same era.
Data-Efficient Hierarchical Reinforcement Learning
Ofir Nachum, Shane Shixiang Gu, Honglak Lee, and Sergey Levine · 2018
Cited alongside, same era.
Visual Reinforcement Learning with Imagined Goals
Ashvin Nair, Vitchyr Pong, Murtaza Dalal, Shikhar Bahl, Steven Lin, and Sergey Levine · 2018
Cited alongside, same era.
Representation Learning with Contrastive Predictive Coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Cited alongside, same era.
Natural environment benchmarks for reinforcement learning
Amy Zhang, Yuxin Wu, and Joelle Pineau · 2018
Cited alongside, same era.
Hallucinative topological memory for zero-shot visual planning
Kara Liu, Thanard Kurutach, Christine Tung, Pieter Abbeel, and Aviv Tamar · 2020
Later among the works it cites.
Plan2vec: Unsupervised representation learning by latent plans
Ge Yang, Amy Zhang, Ari Morcos, Joelle Pineau, Pieter Abbeel, and Roberto Calandra · 2020
Later among the works it cites.
Contrastive behavioral similarity embeddings for generalization in reinforcement learning
Rishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro, and Marc G Bellemare · 2021
Later among the works it cites.
MICo: Improved representations via sampling-based state similarity for markov decision processes
Pablo Samuel Castro, Tyler Kastner, Prakash Panangaden, and Mark Rowland · 2021
Later among the works it cites.
Actionable models: Unsupervised offline reinforcement learning of robotic skills
Yevgen Chebotar, Karol Hausman, Yao Lu, Ted Xiao, Dmitry Kalashnikov, Jacob Varley, Alex Irpan, Benjamin Eysenbach, Ryan Julian, Chelsea Finn, and Sergey Levine · 2021
Later among the works it cites.
Variational empowerment as representation learning for goal-conditioned reinforcement learning
Jongwook Choi, Archit Sharma, Honglak Lee, Sergey Levine, and Shixiang Shane Gu · 2021
Later among the works it cites.
Pybullet, a python module for physics simulation for games, robotics and machine learning
Erwin Coumans and Yunfei Bai · 2021
Later among the works it cites.
Learning domain invariant representations in goal-conditioned block MDPs
Beining Han, Chongyi Zheng, Harris Chan, Keiran Paster, Michael R. Zhang, and Jimmy Ba · 2021
Later among the works it cites.
What can I do here? learning new skills by imagining visual affordances
Alexander Khazatsky, Ashvin Nair, Daniel Jing, and Sergey Levine · 2021
Later among the works it cites.
Offline reinforcement learning with implicit q-learning
Ilya Kostrikov, Ashvin Nair, and Sergey Levine · 2021
Later among the works it cites.
Rapid exploration for open-world navigation with latent goal models
Dhruv Shah, Benjamin Eysenbach, Nicholas Rhinehart, and Sergey Levine · 2021
Later among the works it cites.
Model-based visual planning with self-supervised functional distances
Stephen Tian, Suraj Nair, Frederik Ebert, Sudeep Dasari, Benjamin Eysenbach, Chelsea Finn, and Sergey Levine · 2021
Later among the works it cites.
Learning invariant representations for reinforcement learning without reconstruction
Amy Zhang, Rowan Thomas McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine · 2021
Later among the works it cites.