Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) has revolutionized learning and actuation in applications such as game playing and robotic control.
The nist definition of cloud computing
P. Mell and T. Grance · 2011
Earlier work this paper cites.
MuJoCo: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Elasticity in cloud computing: What it is, and what it is not
N. Herbst, Samuel Kounev, and Ralf H. Reussner · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller · 2013
Earlier work this paper cites.
Containers and cloud: From lxc to docker to kubernetes
David Bernstein · 2014
Earlier work this paper cites.
Standard & Poor’s compustat, 2015
Wharton Research Data Service · 2015
Earlier work this paper cites.
Microservices architecture enables devops: Migration to a cloud-native architecture
Armin Balalaie, A. Heydarnoori, and Pooyan Jamshidi · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2016
Earlier work this paper cites.
Deep reinforcement learning with double Q-learning
Hado Van Hasselt, Arthur Guez, and David Silver · 2016
Earlier work this paper cites.
OpenAI baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov · 2017
Earlier work this paper cites.
Cloud-native applications
Dennis Gannon, R. Barga, and Neel Sundaresan · 2017
Cited alongside, same era.
Evolution strategies as a scalable alternative to reinforcement learning
Tim Salimans, Jonathan Ho, Xi Chen, Szymon Sidor, and Ilya Sutskever · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Mastering the game of Go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Dopamine: A research framework for deep reinforcement learning
Pablo Samuel Castro, Subhodeep Moitra, Carles Gelada, Saurabh Kumar, and Marc G Bellemare · 2018
rlpyt: A research code base for deep reinforcement learning in pytorch
Adam Stooke and Pieter Abbeel · 2019
Later among the works it cites.
FinRL: A deep reinforcement learning library for automated stock trading in quantitative finance
Xiao-Yang Liu, Hongyang Yang, Qian Chen, Runjia Zhang, Liuqing Yang, Bowen Xiao, and Christina Dan Wang · 2020
Later among the works it cites.
NVIDIA DGX SuperPOD: Scalable infrastructure for AI leadership
NVIDIA DGX A100 system reference architecture · 2020
Later among the works it cites.
DD-PPO: Learning near-perfect pointgoal navigators from 2.5 billion frames
Erik Wijmans, Abhishek Kadian, Ari S. Morcos, Stefan Lee, Irfan Essa, Devi Parikh, M. Savva, and Dhruv Batra · 2020
Later among the works it cites.
NVIDIA A100 tensor core GPU: Performance and innovation
Jack Choquette, Wishwesh Gandhi, Olivier Giroux, Nick Stam, and Ronny Krashinsky · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke Hoof, and David Meger · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, P. Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
RLlib: Abstractions for distributed reinforcement learning
Eric Liang, Richard Liaw, Robert Nishihara, Philipp Moritz, Roy Fox, Ken Goldberg, Joseph Gonzalez, Michael Jordan, and Ion Stoica · 2018
Cited alongside, same era.
OpenAI spinning up
OpenAI · 2018
Cited alongside, same era.
Stable baselines3
Antonin Raffin, Ashley Hill, Maximilian Ernestus, Adam Gleave, Anssi Kanervisto, and Noah Dormann · 2019
Cited alongside, same era.
Matteo Hessel, Manuel Kroiss, Aidan Clark, Iurii Kemaev, John Quan, Thomas Keck, Fabio Viola, and Hado van Hasselt · 2021
Closest in time.
FinRL-Podracer: High performance and scalable deep reinforcement learning for quantitative finance
Zechu Li, Xiao-Yang Liu, Jiahao Zheng, Zhaoran Wang, Anwar Walid, and Jian Guo · 2021
Closest in time.
ElegantRL: A lightweight and stable deep reinforcement learning library
Xiao-Yang Liu, Zechu Li, Zhaoran Wang, and Jiahao Zheng · 2021
Closest in time.
Isaac Gym: High performance GPU-based physics simulation for robot learning
Viktor Makoviychuk, Lukasz Wawrzyniak, Yunrong Guo, Michelle Lu, Kier Storey, Miles Macklin, David Hoeller, Nikita Rudin, Arthur Allshire, Ankur Handa, et al · 2021
Closest in time.
Reinforcement learning for robot research: A comprehensive review and open issues
Tengteng Zhang and Hongwei Mo · 2021
Closest in time.