Fetching the paper…
Reading the bibliography…
The partially observable card game Hanabi has recently been proposed as a new AI challenge problem due to its dependence on implicit communication conventions and apparent necessity of theory of mind reasoning for efficient play.
Mindblindness: An Essay on Autism and Theory of Mind
Simon Baron-Cohen, Leda Cosmides, and John Tooby · 1995
Earlier work this paper cites.
A framework for sequential planning in multi-agent settings
Prashant Doshi and Piotr J. Gmytrasiewicz · 2004
Earlier work this paper cites.
Lecture 6.5—RmsProp: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Earlier work this paper cites.
Theory of mind: mechanisms, methods, and new directions
Lindsey J. Byom and Bilge Mutlu · 2013
Earlier work this paper cites.
How to make the perfect fireworks display: Two strategies for hanabi
Christopher Cox, Jessica De Silva, Philip Deorsey, Franklin HJ Kenter, Troy Retter, and Josh Tobin · 2015
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Qmdp-net: Deep learning for planning under partial observability
Péter Karkus, David Hsu, and Wee Sun Lee · 2017
Cited alongside, same era.
Bayesian action decoder for deep multi-agent reinforcement learning
Jakob N Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling · 2018
Cited alongside, same era.
Rainbow: Combining improvements in deep reinforcement learning
Matteo Hessel, Joseph Modayil, Hado Van Hasselt, Tom Schaul, Georg Ostrovski, Will Dabney, Dan Horgan, Bilal Piot, Mohammad Azar, and David Silver · 2018
Cited alongside, same era.
Github - quuxplusone/hanabi: Framework for writing bots that play hanabi, 2018
A. O’Dwyer · 2018
Later among the works it cites.
Zheng Tian, Shihao Zou, Tim Warr, Lisheng Wu, and Jun Wang · 2018
Later among the works it cites.
Github - wuthefwasthat/hanabi.rs: Hanabi simulation in rust
J. Wu · 2018
Later among the works it cites.
The hanabi challenge: A new frontier for ai research, 2019
Nolan Bard, Jakob N. Foerster, Sarath Chandar, Neil Burch, Marc Lanctot, H. Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, Iain Dunning, Shibl Mourad, Hugo Larochelle, Marc G. Bellemare, and Michael Bowling · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…