Generalized hindsight for reinforcement learning
Alexander Li, Lerrel Pinto, and Pieter Abbeel · 2020
Later among the works it cites.
Learning latent plans from play
Corey Lynch, Mohi Khansari, Ted Xiao, Vikash Kumar, Jonathan Tompson, Sergey Levine, and Pierre Sermanet · 2020
Later among the works it cites.
Accelerating online reinforcement learning with offline datasets
Original
Ashvin Nair, Murtaza Dalal, Abhishek Gupta, and Sergey Levine · 2020
Later among the works it cites.
Maximum entropy gain exploration for long horizon multi-goal reinforcement learning
Silviu Pitis, Harris Chan, Stephen Zhao, Bradly Stadie, and Jimmy Ba · 2020
Later among the works it cites.
Evolutionary stochastic policy distillation
Original
Hao Sun, Xinyu Pan, Bo Dai, Dahua Lin, and Bolei Zhou · 2020
Later among the works it cites.
Mopo: Model-based offline policy optimization
Tianhe Yu, Garrett Thomas, Lantao Yu, Stefano Ermon, James Y Zou, Sergey Levine, Chelsea Finn, and Tengyu Ma · 2020
Later among the works it cites.
Deep reinforcement learning at the edge of the statistical precipice
Rishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron Courville, and Marc G Bellemare · 2021
Later among the works it cites.
Actionable models: Unsupervised offline reinforcement learning of robotic skills
Yevgen Chebotar, Karol Hausman, Yao Lu, Ted Xiao, Dmitry Kalashnikov, Jake Varley, Alex Irpan, Benjamin Eysenbach, Ryan Julian, Chelsea Finn, et al · 2021
Later among the works it cites.
Learning to reach goals via iterated supervised learning
Dibya Ghosh, Abhishek Gupta, Ashwin Reddy, Justin Fu, Coline Manon Devin, Benjamin Eysenbach, and Sergey Levine · 2021
Later among the works it cites.
FOCAL: efficient fully-offline meta-reinforcement learning via distance metric learning and behavior regularization
Lanqing Li, Rui Yang, and Dijun Luo · 2021
Later among the works it cites.
Offline reinforcement learning with value-based episodic memory
Original
Xiaoteng Ma, Yiqin Yang, Hao Hu, Qihan Liu, Jun Yang, Chongjie Zhang, Qianchuan Zhao, and Bin Liang · 2021
Later among the works it cites.
S4rl: Surprisingly simple self-supervision for offline reinforcement learning
Original
Samarth Sinha and Animesh Garg · 2021
Later among the works it cites.
Offline reinforcement learning with reverse model-based imagination
Jianhao Wang, Wenzhe Li, Haozhe Jiang, Guangxiang Zhu, Siyuan Li, and Chongjie Zhang · 2021
Later among the works it cites.
Believe what you see: Implicit constraint approach for offline multi-agent reinforcement learning
Original
Yiqin Yang, Xiaoteng Ma, Chenghao Li, Zewu Zheng, Qiyuan Zhang, Gao Huang, Jun Yang, and Qianchuan Zhao · 2021
Later among the works it cites.
Mutual information-based state-control for intrinsically motivated reinforcement learning
Rui Zhao, Yang Gao, Pieter Abbeel, Volker Tresp, and Wei Xu · 2021
Later among the works it cites.