Fetching the paper…
Reading the bibliography…
Agents that interact with other agents often do not know a priori what the other agents' strategies are, but have to maximise their own online return while interacting with and learning about others.
Stochastic games
Lloyd S Shapley · 1953
Earlier work this paper cites.
Stochastic dynamic programming, mathematical centre tracts, vol. 139, 1981
J Van Der Wal · 1981
Earlier work this paper cites.
Exploration strategies for model-based learning in multi-agent systems
David Carmel and Shaul Markovitch · 1990
Earlier work this paper cites.
Machine learning, volume 1 of 1, 1997
Tom M Mitchell · 1997
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative learning
Ming Tan · 1997
Earlier work this paper cites.
On the evolution of behavioral heterogeneity in individuals and populations
Carl T Bergstrom and Peter Godfrey-Smith · 1998
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L Littman, and Anthony R Cassandra · 1998
Earlier work this paper cites.
Optimal Learning: Computational procedures for Bayes-adaptive Markov decision processes
Michael O’Gordon Duff and Andrew Barto · 2002
Earlier work this paper cites.
Coordination in multiagent reinforcement learning: A bayesian approach
Georgios Chalkiadakis and Craig Boutilier · 2003
Earlier work this paper cites.
Evaluating the rainbow dqn agent in hanabi with unseen partners
Rodrigo Canaan, Xianbo Gao, Youjin Chung, Julian Togelius, Andy Nealen, and Stefan Menzel · 2004
Earlier work this paper cites.
Generating and adapting to diverse ad-hoc cooperation agents in hanabi
Rodrigo Canaan, Xianbo Gao, Julian Togelius, Andy Nealen, and Stefan Menzel · 2004
Earlier work this paper cites.
Game theory: a critical text
Shaun Hargreaves Heap and Yanis Varoufakis · 2004
Earlier work this paper cites.
A framework for sequential planning in multi-agent settings
Piotr J Gmytrasiewicz and Prashant Doshi · 2005
Earlier work this paper cites.
Beliefs in repeated games
John H Nachbar · 2005
Earlier work this paper cites.
Particle filtering for dynamic agent modelling in simplified poker
Nolan Bard and Michael Bowling · 2007
Earlier work this paper cites.
A Bayesian approach to multiagent reinforcement learning and coalition formation under uncertainty
Georgios Chalkiadakis · 2007
Earlier work this paper cites.
Cooperative games with overlapping coalitions
Georgios Chalkiadakis, Edith Elkind, Evangelos Markakis, Maria Polukarov, and Nick R Jennings · 2010
Earlier work this paper cites.
Ad hoc autonomous agent teams: Collaboration without pre-coordination
Peter Stone, Gal A Kaminka, Sarit Kraus, and Jeffrey S Rosenschein · 2010
Earlier work this paper cites.
Bayes-adaptive interactive pomdps
Brenda Ng, Kofi Boakye, Carol Meyers, and Andrew Wang · 2012
Cited alongside, same era.
Bayes’ bluff: Opponent modelling in poker
Finnegan Southey, Michael P Bowling, Bryce Larson, Carmelo Piccione, Neil Burch, Darse Billings, and Chris Rayner · 2012
Cited alongside, same era.
Teamwork with limited knowledge of teammates
Samuel Barrett, Peter Stone, Sarit Kraus, and Avi Rosenfeld · 2013
Cited alongside, same era.
A general framework for interacting bayes-optimally with self-interested agents using arbitrary parametric model and model prior
Trong Nghia Hoang and Kian Hsiang Low · 2013
Cited alongside, same era.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2013
Cited alongside, same era.
Learning hierarchical features from deep generative models
Shengjia Zhao, Jiaming Song, and Stefano Ermon · 2017
Later among the works it cites.
Autonomous agents modelling other agents: A comprehensive survey and open problems
Stefano V Albrecht and Peter Stone · 2018
Later among the works it cites.
Bayesian action decoder for deep multi-agent reinforcement learning
Jakob N Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling · 2018
Later among the works it cites.
Learning policy representations in multiagent systems
Aditya Grover, Maruan Al-Shedivat, Jayesh K Gupta, Yura Burda, and Harrison Edwards · 2018
Later among the works it cites.
Actor-attention-critic for multi-agent reinforcement learning
Shariq Iqbal and Fei Sha · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
New Advances on Bayesian and Decision-Theoretic Approaches for Interactive Machine Learning
Trong Nghia Hoang · 2014
Cited alongside, same era.
Best response bayesian reinforcement learning for multiagent systems with state uncertainty
Frans A Oliehoek, Christopher Amato, et al · 2014
Cited alongside, same era.
A recurrent latent variable model for sequential data
Junyoung Chung, Kyle Kastner, Laurent Dinh, Kratarth Goel, Aaron C Courville, and Yoshua Bengio · 2015
Cited alongside, same era.
Monte carlo planning method estimates planning horizons during interactive social exchange
Andreas Hula, P Read Montague, and Peter Dayan · 2015
Cited alongside, same era.
Belief and truth in hypothesised behaviours
Stefano V Albrecht, Jacob W Crandall, and Subramanian Ramamoorthy · 2016
Cited alongside, same era.
Rl2: Fast reinforcement learning via slow reinforcement learning. 2016
Yan Duan, John Schulman, Xi Chen, Peter L Bartlett, Ilya Sutskever, and Pieter Abbeel · 2016
Cited alongside, same era.
Bayesian reinforcement learning: A survey
Mohammad Ghavamzadeh, Shie Mannor, Joelle Pineau, and Aviv Tamar · 2016
Cited alongside, same era.
Later among the works it cites.
Machine theory of mind
Neil Rabinowitz, Frank Perbet, Francis Song, Chiyuan Zhang, SM Ali Eslami, and Matthew Botvinick · 2018
Later among the works it cites.
Multi-agent generative adversarial imitation learning
Jiaming Song, Hongyu Ren, Dorsa Sadigh, and Stefano Ermon · 2018
Later among the works it cites.
Reasoning about hypothetical agent behaviours and their parameters
Stefano V Albrecht and Peter Stone · 2019
Later among the works it cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K Ho, Tom Griffiths, Sanjit Seshia, Pieter Abbeel, and Anca Dragan · 2019
Later among the works it cites.
Agent modeling as auxiliary task for deep reinforcement learning
Pablo Hernandez-Leal, Bilal Kartal, and Matthew E Taylor · 2019
Later among the works it cites.
Simplified action decoder for deep multi-agent reinforcement learning
Hengyuan Hu and Jakob N Foerster · 2019
Later among the works it cites.
Meta reinforcement learning as task inference
Jan Humplik, Alexandre Galashov, Leonard Hasenclever, Pedro A Ortega, Yee Whye Teh, and Nicolas Heess · 2019
Later among the works it cites.
Meta-learning of sequential strategies
Pedro A Ortega, Jane X Wang, Mark Rowland, Tim Genewein, Zeb Kurth-Nelson, Razvan Pascanu, Nicolas Heess, Joel Veness, Alex Pritzel, Pablo Sprechmann, et al · 2019
Later among the works it cites.
Finding friend and foe in multi-agent games
Jack Serrino, Max Kleiman-Weiner, David C Parkes, and Josh Tenenbaum · 2019
Later among the works it cites.
Variational autoencoders for opponent modeling in multi-agent systems
Georgios Papoudakis and Stefano V Albrecht · 2020
Later among the works it cites.
Learning to play against any mixture of opponents
Max Olan Smith, Thomas Anthony, Yongzhao Wang, and Michael P Wellman · 2020
Later among the works it cites.
Varibad: A very good method for bayes-adaptive deep rl via meta-learning
Luisa Zintgraf, Kyriacos Shiarlis, Maximilian Igl, Sebastian Schulze, Yarin Gal, Katja Hofmann, and Shimon Whiteson · 2020
Later among the works it cites.