Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (RL) is a powerful framework to train decision-making models in complex environments.
Markov games as a framework for multi-agent reinforcement learning
Littman, M. L · 1994
Earlier work this paper cites.
Openai gym, 2016
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V., Badia, A. P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
Gu, S., Holly, E., Lillicrap, T., and Levine, S · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al · 2017
Earlier work this paper cites.
Elf: An extensive, lightweight and flexible research platform for real-time strategy games, 2017
Tian, Y., Gong, Q., Shang, W., Wu, Y., and Zitnick, C. L · 2017
Earlier work this paper cites.
JAX: composable transformations of Python+NumPy programs, 2018
Bradbury, J., Frostig, R., Hawkins, P., Johnson, M. J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., and Zhang, Q · 2018
Cited alongside, same era.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures, 2018
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., Legg, S., and Kavukcuoglu, K · 2018
Cited alongside, same era.
Openai five
OpenAI · 2018
Cited alongside, same era.
Reinforcement Learning: An Introduction
Sutton, R. S. and Barto, A. G · 2018
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Cited alongside, same era.
Accelerating reinforcement learning through gpu atari emulation, 2020
Acme: A research framework for distributed reinforcement learning, 2020
Hoffman, M., Shahriari, B., Aslanides, J., Barth-Maron, G., Behbahani, F., Norman, T., Abdolmaleki, A., Cassirer, A., Yang, F., Baumli, K., Henderson, S., Novikov, A., Colmenarejo, S. G., Cabi, S., Gulcehre, C., Paine, T. L., Cowie, A., Wang, Z., Piot, B., and de Freitas, N · 2020
Later among the works it cites.
Brax – a differentiable physics engine for large scale rigid body simulation, 2021
Freeman, C. D., Frey, E., Raichuk, A., Girgin, S., Mordatch, I., and Bachem, O · 2021
Closest in time.
Isaac gym: High performance gpu-based physics simulation for robot learning, 2021
Makoviychuk, V., Wawrzyniak, L., Guo, Y., Lu, M., Storey, K., Macklin, M., Hoeller, D., Rudin, N., Allshire, A., Handa, A., and State, G · 2021
Closest in time.
Megaverse: Simulating embodied agents at one million experiences per second, 2021
Petrenko, A., Wijmans, E., Shacklett, B., and Koltun, V · 2021
Closest in time.
Mava: a research framework for distributed multi-agent reinforcement learning, 2021
Pretorius, A., Tessera, K.-a., Smit, A. P., Formanek, C., Grimbly, S. J., Eloff, K., Danisa, S., Francis, L., Shock, J., Kamper, H., Brink, W., Engelbrecht, H., Laterre, A., and Beguir, K · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dalton, S., Frosio, I., and Garland, M · 2020
Cited alongside, same era.
Seed rl: Scalable and efficient deep-rl with accelerated central inference, 2020
Espeholt, L., Marinier, R., Stanczyk, P., Wang, K., and Michalski, M · 2020
Cited alongside, same era.
Closest in time.
Trott, A., Srinivasa, S., van der Wal, D., Haneuse, S., and Zheng, S · 2021
Closest in time.
The ai economist: Optimal economic policy design via two-level deep reinforcement learning, 2021
Zheng, S., Trott, A., Srinivasa, S., Parkes, D. C., and Socher, R · 2021
Closest in time.