Fetching the paper…
Reading the bibliography…
We present RLLTE: a long-term evolution, extremely modular, and open-source framework for reinforcement learning (RL) research and application.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Earlier work this paper cites.
Brockman, G.; Cheung, V.; Pettersson, L.; Schneider, J.; Schulman, J.; Tang, J.; and Zaremba, W. 2016 · 2016
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
Pathak, D.; Agrawal, P.; Efros, A. A.; and Darrell, T. 2017 · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
Silver, D.; Schrittwieser, J.; Simonyan, K.; Antonoglou, I.; Huang, A.; Guez, A.; Hubert, T.; Baker, L.; Lai, M.; Bolton, A.; et al. 2017 · 2017
Earlier work this paper cites.
Reinforcement learning with augmented data
Laskin, M.; Lee, K.; Stooke, A.; Pinto, L.; Abbeel, P.; and Srinivas, A. 2020 · 2020
Cited alongside, same era.
Curl: Contrastive unsupervised representations for reinforcement learning
Laskin, M.; Srinivas, A.; and Abbeel, P. 2020 · 2020
Cited alongside, same era.
RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Raileanu, R.; Rocktäschel, T.; and Raileanu, R. 2020 · 2020
Cited alongside, same era.
Stable-baselines3: Reliable reinforcement learning implementations
Raffin, A.; Hill, A.; Gleave, A.; Kanervisto, A.; Ernestus, M.; and Dormann, N. 2021 · 2021
Cited alongside, same era.
Cleanrl: High-quality single-file implementations of deep reinforcement learning algorithms
Huang, S.; Dossa, R. F. J.; Ye, C.; Braga, J.; Chakraborty, D.; Mehta, K.; and Araújo, J. G. 2022 · 2022
Later among the works it cites.
Tianshou: A highly modularized deep reinforcement learning library
Weng, J.; Chen, H.; Yan, D.; You, K.; Duburcq, A.; Zhang, M.; Su, Y.; Su, H.; and Zhu, J. 2022 · 2022
Later among the works it cites.
Faster sorting algorithms discovered using deep reinforcement learning
Mankowitz, D. J.; Michi, A.; Zhernov, A.; Gelmi, M.; Selvi, M.; Paduraru, C.; Leurent, E.; Iqbal, S.; Lespiau, J.-B.; Ahern, A.; et al. 2023 · 2023
Closest in time.
Gymnasium: A Standard Interface for Reinforcement Learning Environments
Towers, M.; Kwiatkowski, A.; Terry, J.; Balis, J. U.; De Cola, G.; Deleu, T.; Goulão, M.; Kallinteris, A.; Krimmel, M.; KG, A.; et al. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…