2016

OpenAI Gym

Brockman, Greg, Cheung, Vicki, Pettersson, Ludwig et al.

Understand

OpenAI Gym is a toolkit for reinforcement learning research.

  • It includes a growing collection of benchmark problems that expose a common interface, and a website where people can share their results and compare the performance of algorithms.
  • This whitepaper discusses the components of OpenAI Gym and the design decisions that went into the software.

Built on

  • Dynamic programming and optimal control

    Dimitri P Bertsekas, Dimitri P Bertsekas, Dimitri P Bertsekas, and Dimitri P Bertsekas · 1995

    Earlier work this paper cites.

  • Reinforcement Learning: An Introduction

    R. S. Sutton and A. G. Barto · 1998

    Earlier work this paper cites.

  • RL-Glue: Language-independent software for reinforcement-learning experiments

    B. Tanner and A. White · 2009

    Earlier work this paper cites.

  • Pachi: State of the art open source go program

    Petr Baudiš and Jean-loup Gailly · 2011

    Earlier work this paper cites.

  • Mujoco: A physics engine for model-based control

    Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012

    Earlier work this paper cites.

Similar

  • The Arcade Learning Environment: An evaluation platform for general agents

    M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013

    Cited alongside, same era.

  • RLLib: Lightweight standard and on/off policy reinforcement learning library (C++)

    S. Abeyruwan · 2013

    Cited alongside, same era.

  • The reinforcement learning competition 2014

    Christos Dimitrakakis, Guangliang Li, and Nikoalos Tziortziotis · 2014

    Cited alongside, same era.

  • Human-level control through deep reinforcement learning

    V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, Sadik Beattie, C., Antonoglou A., H. I., King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015

    Cited alongside, same era.

  • Trust region policy optimization

    J. Schulman, S. Levine, P. Abbeel, M. I. Jordan, and P. Moritz · 2015

    Cited alongside, same era.

Then

Beyond the bibliography

alphaXiv searches the wider corpus for related work and actual follow-ups.

Open on alphaXiv

alphaXiv is searching for related work…