Fetching the paper…
Reading the bibliography…
We study goal-conditioned RL through the lens of generalization, but not in the traditional sense of random augmentations and domain randomization.
Principia Mathematica , volume 2
Alfred North Whitehead and Bertrand Russell · 1927
Earlier work this paper cites.
Report on a General Problem-Solving Program
Allen Newell · 1959
Earlier work this paper cites.
A Note on Two Problems in Connexion With Graphs
Ew Dijkstra · 1959
Earlier work this paper cites.
Principles of Neurodynamics: Perceptrons and the Theory of Brain Mechanisms
F. Rosenblatt · 1961
Earlier work this paper cites.
A Formal Basis for the Heuristic Determination of Minimum Cost Paths
Peter E. Hart, Nils J. Nilsson, and Bertram Raphael · 1968
Earlier work this paper cites.
A Neural Network Digit Recognizer
David J. Burr · 1986
Earlier work this paper cites.
Soar: An Architecture for General Intelligence
John E Laird, Allen Newell, and Paul S Rosenbloom · 1987
Earlier work this paper cites.
Dyna, an Integrated Architecture for Learning, Planning, and Reacting
Richard S Sutton · 1991
Earlier work this paper cites.
Q-Learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Improving Generalization for Temporal Difference Learning: The Successor Representation
Peter Dayan · 1993
Earlier work this paper cites.
A Global Geometric Framework for Nonlinear Dimensionality Reduction
Joshua B Tenenbaum, Vin de Silva, and John C Langford · 2000
Earlier work this paper cites.
Randomized Kinodynamic Planning
Steven M. LaValle and James J. Kuffner · 2001
Earlier work this paper cites.
On the Sample Complexity of Reinforcement Learning
Sham Machandranath Kakade · 2003
Earlier work this paper cites.
Learning Methods for Generic Object Recognition With Invariance to Pose and Lighting
Y. LeCun, Fu Jie Huang, and L. Bottou · 2004
Earlier work this paper cites.
Maximum Entropy Inverse Reinforcement Learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
An Elementary Proof of the Triangle Inequality for the Wasserstein Metric
Philippe Clement and Wolfgang Desch · 2008
Earlier work this paper cites.
Using Bisimulation for Policy Transfer in MDPs
Pablo Castro and Doina Precup · 2010
Earlier work this paper cites.
Bisimulation Metrics for Continuous Markov Decision Processes
Norm Ferns, Prakash Panangaden, and Doina Precup · 2011
Earlier work this paper cites.
Group Equivariant Convolutional Networks
Taco S. Cohen and Max Welling · 2016
Earlier work this paper cites.
Improved Deep Metric Learning With Multi-Class N-Pair Loss Objective
Kihyuk Sohn · 2016
Earlier work this paper cites.
Value Iteration Networks
Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, and Pieter Abbeel · 2016
Earlier work this paper cites.
Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation
Tejas D Kulkarni, Karthik Narasimhan, Ardavan Saeedi, and Josh Tenenbaum · 2016
Earlier work this paper cites.
Successor Features for Transfer in Reinforcement Learning
André Barreto, Will Dabney, Rémi Munos, Jonathan J. Hunt, Tom Schaul, Hado van Hasselt, and David Silver · 2017
Earlier work this paper cites.
Information Theoretic Mpc for Model-Based Reinforcement Learning
Grady Williams, Nolan Wagener, Brian Goldfain, Paul Drews, James M Rehg, Byron Boots, and Evangelos A Theodorou · 2017
Earlier work this paper cites.
Gotta Learn Fast: A New Benchmark for Generalization in RL
Alex Nichol, Vicki Pfau, Christopher Hesse, Oleg Klimov, and John Schulman · 2018
Earlier work this paper cites.
Generalization and Regularization in DQN
Jesse Farebrother, Marlos C Machado, and Michael Bowling · 2018
Earlier work this paper cites.
Illuminating Generalization in Deep Reinforcement Learning Through Procedural Level Generation
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius, and Sebastian Risi · 2018
Earlier work this paper cites.
A Study on Overfitting in Deep Reinforcement Learning
Chiyuan Zhang, Oriol Vinyals, Remi Munos, and Samy Bengio · 2018
Earlier work this paper cites.
Assessing Generalization in Deep Reinforcement Learning
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Song · 2018
Cited alongside, same era.
Semi-Parametric Topological Memory for Navigation
Nikolay Savinov, Alexey Dosovitskiy, and Vladlen Koltun · 2018
Cited alongside, same era.
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning With a Stochastic Actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Gated Path Planning Networks
Lisa Lee, Emilio Parisotto, Devendra Singh Chaplot, Eric Xing, and Ruslan Salakhutdinov · 2018
Cited alongside, same era.
Deep Reinforcement Learning in a Handful of Trials Using Probabilistic Dynamics Models
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine · 2018
Cited alongside, same era.
Learning Transferable Visual Models From Natural Language Supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Later among the works it cites.
Model-Based Reinforcement Learning via Latent-Space Collocation
Oleh Rybkin, Chuning Zhu, Anusha Nagabandi, Kostas Daniilidis, Igor Mordatch, and Sergey Levine · 2021
Later among the works it cites.
Intrinsically Motivated Goal-Conditioned Reinforcement Learning: A Short Survey
Cédric Colas, Tristan Karch, Olivier Sigaud, and Pierre-Yves Oudeyer · 2022
Later among the works it cites.
Rethinking Goal-Conditioned Supervised Learning and Its Connection to Offline RL
Rui Yang, Yiming Lu, Wenzhe Li, Hao Sun, Meng Fang, Yali Du, Xiu Li, Lei Han, and Chongjie Zhang · 2022
Later among the works it cites.
How Far I’ll Go: Offline Goal-Conditioned Reinforcement Learning via f f -Advantage Regression
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural Network Dynamics for Model-Based Deep Reinforcement Learning With Model-Free Fine-Tuning
Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine · 2018
Cited alongside, same era.
Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Control
Kendall Lowrey, Aravind Rajeswaran, Sham Kakade, Emanuel Todorov, and Igor Mordatch · 2018
Cited alongside, same era.
Geoffrey Irving, Paul Christiano, and Dario Amodei · 2018
Cited alongside, same era.
Quantifying Generalization in Reinforcement Learning
Karl Cobbe, Oleg Klimov, Chris Hesse, Taehoon Kim, and John Schulman · 2019
Cited alongside, same era.
Action Robust Reinforcement Learning and Applications in Continuous Control
Chen Tessler, Yonathan Efroni, and Shie Mannor · 2019
Cited alongside, same era.
Generalization in Reinforcement Learning With Selective Noise Injection and Information Bottleneck
Maximilian Igl, Kamil Ciosek, Yingzhen Li, Sebastian Tschiatschek, Cheng Zhang, Sam Devlin, and Katja Hofmann · 2019
Cited alongside, same era.
Unsupervised State Representation Learning in Atari
Ankesh Anand, Evan Racah, Sherjil Ozair, Yoshua Bengio, Marc-Alexandre Côté, and R. Devon Hjelm · 2019
Cited alongside, same era.
Yecheng Jason Ma, Jason Yan, Dinesh Jayaraman, and Osbert Bastani · 2022
Later among the works it cites.
High-Resolution Image Synthesis With Latent Diffusion Models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Exploring Length Generalization in Large Language Models
Cem Anil, Yuhuai Wu, Anders Andreassen, Aitor Lewkowycz, Vedant Misra, Vinay Ramasesh, Ambrose Slone, Guy Gur-Ari, Ethan Dyer, and Behnam Neyshabur · 2022
Later among the works it cites.
Maximum Entropy RL (Provably) Solves Some Robust RL Problems
Benjamin Eysenbach and Sergey Levine · 2022
Later among the works it cites.
Robust Reinforcement Learning: A Review of Foundations and Recent Advances
Janosch Moos, Kay Hansel, Hany Abdulsamad, Svenja Stark, Debora Clever, and Jan Peters · 2022
Later among the works it cites.
Bisimulation Makes Analogies in Goal-Conditioned Reinforcement Learning
Philippe Hansen-Estruch, Amy Zhang, Ashvin Nair, Patrick Yin, and Sergey Levine · 2022
Later among the works it cites.
ViKiNG: Vision-Based Kilometer-Scale Navigation With Geographic Hints
Dhruv Shah and Sergey Levine · 2022
Later among the works it cites.
Improved Representation of Asymmetrical Distances With Interval Quasimetric Embeddings
Tongzhou Wang and Phillip Isola · 2022
Later among the works it cites.
Introduction to Algorithms
Thomas H Cormen, Charles E Leiserson, Ronald L Rivest, and Clifford Stein · 2022
Later among the works it cites.
Contrastive Learning as Goal-Conditioned Reinforcement Learning
Benjamin Eysenbach, Tianjun Zhang, Ruslan Salakhutdinov, and Sergey Levine · 2022
Later among the works it cites.
PALMER: Perception-Action Loop With Memory for Long-Horizon Planning
Onur Beker, Mohammad Mohammadi, and Amir Zamir · 2022
Later among the works it cites.
Large Language Models Can Self-Improve
Jiaxin Huang, Shixiang Shane Gu, Le Hou, Yuexin Wu, Xuezhi Wang, Hongkun Yu, and Jiawei Han · 2022
Later among the works it cites.
Optimal Goal-Reaching Reinforcement Learning via Quasimetric Learning
Tongzhou Wang, Antonio Torralba, Phillip Isola, and Amy Zhang · 2023
Later among the works it cites.
RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Anthony Brohan, Noah Brown, Justice Carbajal, Yevgen Chebotar, Xi Chen, Krzysztof Choromanski, Tianli Ding, Danny Driess, et al · 2023
Later among the works it cites.
Maximum State Entropy Exploration Using Predecessor and Successor Representations
Arnav Kumar Jain, Lucas Lehnert, Irina Rish, and Glen Berseth · 2023
Later among the works it cites.
ProgPrompt: Generating Situated Robot Task Plans Using Large Language Models
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg · 2023
Later among the works it cites.
Metric Residual Network for Sample Efficient Goal-Conditioned Reinforcement Learning
Bo Liu, Yihao Feng, Qiang Liu, and Peter Stone · 2023
Later among the works it cites.
OpenAI, Josh Achiam, Steven Adler, et al · 2024
Later among the works it cites.
Hiql: Offline Goal-Conditioned RL With Latent States as Actions
Seohong Park, Dibya Ghosh, Benjamin Eysenbach, and Sergey Levine · 2024
Later among the works it cites.
Accelerating Goal-Conditioned RL Algorithms and Research
Michał Bortkiewicz, Władek Pałucki, Vivek Myers, Tadeusz Dziarmaga, Tomasz Arczewski, Łukasz Kuciński, and Benjamin Eysenbach · 2024
Later among the works it cites.
Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference
Benjamin Eysenbach, Vivek Myers, Ruslan Salakhutdinov, and Sergey Levine · 2024
Later among the works it cites.
Closing the Gap Between TD Learning and Supervised Learning - a Generalisation Point of View
Raj Ghugare, Matthieu Geist, Glen Berseth, and Benjamin Eysenbach · 2024
Later among the works it cites.
Supervised Pretraining Can Learn In-Context Reinforcement Learning
Jonathan Lee, Annie Xie, Aldo Pacchiano, Yash Chandak, Chelsea Finn, Ofir Nachum, and Emma Brunskill · 2024
Later among the works it cites.