Bisimulation through probabilistic testing
Kim G Larsen and Arne Skou · 1991
Earlier work this paper cites.
Extrapolation limitations of multilayer feedforward neural networks
Pamela J Haley and DONALD Soloway · 1992
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Martin L Puterman · 1994
Earlier work this paper cites.
Equivalence notions and model minimization in markov decision processes
Robert Givan, Thomas Dean, and Matthew Greig · 2003
Earlier work this paper cites.
Curl: Contrastive unsupervised representations for reinforcement learning
Original
Michael Laskin, Aravind Srinivas, and Pieter Abbeel · 2003
Earlier work this paper cites.
Metrics for finite Markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2004
Earlier work this paper cites.
Reinforcement learning with augmented data
Original
Michael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto, Pieter Abbeel, and Aravind Srinivas · 2004
Earlier work this paper cites.
Methods for computing state similarity in Markov decision processes
Norm Ferns, Pablo Samuel Castro, Doina Precup, and Prakash Panangaden · 2006
Earlier work this paper cites.
Dimensionality reduction by learning an invariant mapping
Raia Hadsell, Sumit Chopra, and Yann LeCun · 2006
Earlier work this paper cites.
Predictive information accelerates learning in rl
Original
Kuang-Huei Lee, Ian Fischer, Anthony Liu, Yijie Guo, Honglak Lee, John Canny, and Sergio Guadarrama · 2007
Earlier work this paper cites.
Optimal transport: old and new
Cédric Villani · 2008
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Matthew E. Taylor and Peter Stone · 2009
Earlier work this paper cites.
Using bisimulation for policy transfer in MDPs
Pablo Samuel Castro and Doina Precup · 2010
Earlier work this paper cites.
Bisimulation metrics for continuous markov decision processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2011
Earlier work this paper cites.
Bisimulation metrics are optimal value functions
Norman Ferns and Doina Precup · 2014
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Google vizier: A service for black-box optimization
Daniel Golovin, Benjamin Solnik, Subhodeep Moitra, Greg Kochanski, John Karro, and D Sculley · 2017
Earlier work this paper cites.
Implicit regularization in matrix factorization
Suriya Gunasekar, Blake E Woodworth, Srinadh Bhojanapalli, Behnam Neyshabur, and Nati Srebro · 2017
Earlier work this paper cites.
Robust and efficient transfer learning with hidden parameter Markov decision processes
Taylor W. Killian, George Dimitri Konidaris, and Finale Doshi-Velez · 2017
Earlier work this paper cites.
Zero-shot task generalization with multi-task deep reinforcement learning
Junhyuk Oh, Satinder P. Singh, Honglak Lee, and Pushmeet Kohli · 2017
Earlier work this paper cites.
The 2017 davis challenge on video object segmentation
Original
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alex Sorkine-Hornung, and Luc Van Gool · 2017
Earlier work this paper cites.