Fetching the paper…
Reading the bibliography…
We present Toribash Learning Environment (ToriLLE), a learning environment for machine learning agents based on the video game Toribash.
F. A. Gers, J. Schmidhuber, and F. Cummins, “Learning to forget: Continual prediction with lstm,” 1999
1999
Earlier work this paper cites.
Nabistudios, “Toribash.” http://www.toribash.com/
2006
Earlier work this paper cites.
J. Byrne, M. O’Neill, and A. Brabazon, “Optimising offensive moves in toribash using a genetic algorithm,” in Proceedings of the Sixteenth International Conference on Soft Computing (MENDEL)
2010
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” arXiv preprint arXiv:1412.6980
2014
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. I. Jordan, and P. Moritz, “Trust region policy optimization,” in International Conference on Machine Learning
2015
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski, “Vizdoom: A doom-based ai research platform for visual reinforcement learning,” in Computational Intelligence and Games (CIG), 2016 IEEE Conference on
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al
2016
Earlier work this paper cites.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell, “Curiosity-driven exploration by self-supervised prediction,” in International Conference on Machine Learning
2017
Cited alongside, same era.
Accessed 1-April-2018
S. Põder, “Toribash move evolver.” https://github.com/windo/toribash-evolver · 2018
Closest in time.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning 2018
2018
Closest in time.
A. Hill, A. Raffin, M. Ernestus, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines.” https://github.com/hill-a/stable-baselines
2018
Closest in time.
A. Tavakoli, F. Pardo, and P. Kormushev, “Action branching architectures for deep reinforcement learning,” in AAAI
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Accessed 23-October-2018
OpenAI, “Openai five.” https://blog.openai.com/openai-five-benchmark-results/ · 2018
Cited alongside, same era.
Accessed 22-May-2018
ToribashUsers, “Toribash forums: Violence evolved: a script that uses a genetic algorithm to evolve openers.” http://forum.toribash.com/showthread.php?t=167355 · 2018
Cited alongside, same era.
Accessed 22-May-2018
ToribashUsers, “Toribash forums: Suggestions and help for neat a.i. script.” http://forum.toribash.com/showthread.php?t=25263 · 2018
Cited alongside, same era.
2018
Closest in time.
2018
Closest in time.
T. Bansal, J. Pachocki, S. Sidor, I. Sutskever, and I. Mordatch, “Emergent complexity via multi-agent competition,” in International Conference on Learning Representations 2018
2018
Closest in time.
2018
Closest in time.
O. Vinyals, I. Babuschkin, J. Chung, M. Mathieu, M. Jaderberg, W. Czarnecki, A. Dudzik, A. Huang, P. Georgiev, R. Powell, et al
2019
Closest in time.