Fetching the paper…
Reading the bibliography…
Deep reinforcement learning algorithms have recently been used to train multiple interacting agents in a centralised manner whilst keeping their execution decentralised.
On the theory of the brownian motion
George E Uhlenbeck and Leonard S Ornstein · 1930
Earlier work this paper cites.
The structure and function of communication in society
Harold D Lasswell · 1948
Earlier work this paper cites.
Communication in the battle of the sexes game: some experimental results
Russell Cooper, Douglas V DeJong, Robert Forsythe, and Thomas W Ross · 1989
Earlier work this paper cites.
Forward induction in coordination games
Russell Cooper, Douglas V De Jong, Robert Forsythe, and Thomas W Ross · 1992
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman · 1994
Earlier work this paper cites.
Learning without state-estimation in partially observable markovian decision processes
Satinder P Singh, Tommi Jaakkola, and Michael I Jordan · 1994
Earlier work this paper cites.
Python tutorial
Guido Van Rossum and Fred L Drake Jr · 1995
Earlier work this paper cites.
Multi-agent reinforcement learning: A modular approach
Norihiko Ono and Kenji Fukumoto · 1996
Earlier work this paper cites.
A general method for multi-agent reinforcement learning in unrestricted environments
Jurgen Schmidhuber · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Elevator group control using multiple reinforcement learning agents
Robert H Crites and Andrew G Barto · 1998
Earlier work this paper cites.
Towards collaborative and adversarial learning: A case study in robotic soccer
Peter Stone and Manuela Veloso · 1998
Earlier work this paper cites.
Introduction to reinforcement learning , volume 135
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
A probabilistic approach to collaborative multi-robot localization
Dieter Fox, Wolfram Burgard, Hannes Kruppa, and Sebastian Thrun · 2000
Earlier work this paper cites.
An algorithm for distributed reinforcement learning in cooperative multi-agent systems
Martin Lauer and Martin Riedmiller · 2000
Earlier work this paper cites.
Learning to cooperate via policy search
Leonid Peshkin, Kee-Eung Kim, Nicolas Meuleau, and Leslie Pack Kaelbling · 2000
Earlier work this paper cites.
Coverage control for mobile sensing networks
Jorge Cortes, Sonia Martinez, Timur Karatas, and Francesco Bullo · 2002
Earlier work this paper cites.
Coordinated reinforcement learning
Carlos Guestrin, Michail Lagoudakis, and Ronald Parr · 2002
Earlier work this paper cites.
Information and communication in sequential bargaining
Jeannette Brosig, Axel Ockenfels, Joachim Weimann, et al · 2003
Earlier work this paper cites.
Multi-agent systems for the simulation of land-use and land-cover change: a review
Dawn C Parker, Steven M Manson, Marco A Janssen, Matthew J Hoffmann, and Peter Deadman · 2003
Earlier work this paper cites.
Natural pragmatics and natural codes
Tim Wharton · 2003
Earlier work this paper cites.
Communication and coordination
John H Miller and Scott Moser · 2004
Earlier work this paper cites.
An experimental study of the emergence of human communication systems
Bruno Galantucci · 2005
Earlier work this paper cites.
Cooperative multi-agent learning: The state of the art
Liviu Panait and Sean Luke · 2005
Earlier work this paper cites.
Crisis management in hindsight: Cognition, communication, coordination, and control
Louise K Comfort · 2007
Earlier work this paper cites.
Hysteretic q-learning: an algorithm for decentralized reinforcement learning in cooperative multi-agent teams
Laëtitia Matignon, Guillaume Laurent, and Nadine Le Fort-Piat · 2007
Earlier work this paper cites.
Consensus and cooperation in networked multi-agent systems
Reza Olfati-Saber, J Alex Fax, and Richard M Murray · 2007
Earlier work this paper cites.
Q-value functions for decentralized pomdps
Frans A Oliehoek and Nikos Vlassis · 2007
Earlier work this paper cites.
The emergence of simple languages in an experimental coordination game
Reinhard Selten and Massimo Warglien · 2007
Cited alongside, same era.
Language, meaning, and games: A model of communication, coordination, and evolution
Stefano Demichelis and Jorgen W Weibull · 2008
Cited alongside, same era.
Multivariate analysis of variance (manova)
Aaron French, Marcelo Macedo, John Poulsen, Tyler Waterson, and Angela Yu · 2008
Cited alongside, same era.
Distributed coordination architecture for multi-robot formation control
Wei Ren and Nathan Sorensen · 2008
Cited alongside, same era.
Synchronization in networks of identical linear systems
Luca Scardovi and Rodolphe Sepulchre · 2008
Cited alongside, same era.
Communication, coordination, and camaraderie in world of warcraft
Mark G Chen · 2009
Cited alongside, same era.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Later among the works it cites.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson · 2016
Later among the works it cites.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2016
Later among the works it cites.
Multi-agent reinforcement learning as a rehearsal for decentralized planning
Landon Kraemer and Bikramjit Banerjee · 2016
Later among the works it cites.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multi-agent systems in a distributed smart grid: Design and implementation
Manisa Pipattanasomporn, Hassan Feroze, and Saifur Rahman · 2009
Cited alongside, same era.
Communication, credibility and negotiation using a cognitive hierarchy model
Michael Wunder, Michael Littman, and Matthew Stone · 2009
Cited alongside, same era.
Exploring the cognitive infrastructure of communication
Jan Peter De Ruiter, Matthijs L Noordzij, Sarah Newman-Norlund, Roger Newman-Norlund, Peter Hagoort, Stephen C Levinson, and Ivan Toni · 2010
Cited alongside, same era.
Can iterated learning explain the emergence of graphical symbols?
Simon Garrod, Nicolas Fay, Shane Rogers, Bradley Walker, and Nik Swoboda · 2010
Cited alongside, same era.
Pre-hunt communication provides context for the evolution of early human language
Szabolcs Számadó · 2010
Cited alongside, same era.
Systematicity and arbitrariness in novel communication systems
Carrie Ann Theisen, Jon Oberlander, and Simon Kirby · 2010
Cited alongside, same era.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Rob Fergus, et al · 2016
Later among the works it cites.
Xiangxiang Chu and Hangjun Ye · 2017
Later among the works it cites.
Counterfactual multi-agent policy gradients
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2017
Later among the works it cites.
Cooperative multi-agent control using deep reinforcement learning
Jayesh K Gupta, Maxim Egorov, and Mykel Kochenderfer · 2017
Later among the works it cites.
A survey of learning in multiagent environments: Dealing with non-stationarity
Pablo Hernandez-Leal, Michael Kaisers, Tim Baarslag, and Enrique Munoz de Cote · 2017
Later among the works it cites.
Revisiting the master-slave architecture in multi-agent deep reinforcement learning
Xiangyu Kong, Bo Xin, Fangchen Liu, and Yizhou Wang · 2017
Later among the works it cites.
Deep reinforcement learning: An overview
Yuxi Li · 2017
Later among the works it cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Igor Mordatch and Pieter Abbeel · 2017
Later among the works it cites.
Automatic differentiation in pytorch, 2017
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Later among the works it cites.
Multiagent bidirectionally-coordinated nets for learning to play starcraft combat games
Peng Peng, Quan Yuan, Ying Wen, Yaodong Yang, Zhenkun Tang, Haitao Long, and Jun Wang · 2017
Later among the works it cites.
Multiagent cooperation and competition with deep reinforcement learning
Ardi Tampuu, Tambet Matiisen, Dorian Kodelja, Ilya Kuzovkin, Kristjan Korjus, Juhan Aru, Jaan Aru, and Raul Vicente · 2017
Later among the works it cites.
Does communication help people coordinate?
Yevgeniy Vorobeychik, Zlatko Joveski, and Sixie Yu · 2017
Later among the works it cites.
Tarmac: Targeted multi-agent communication
Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Michael Rabbat, and Joelle Pineau · 2018
Later among the works it cites.
Deepmind ai reduces google data centre cooling bill by 40
Richard Evans and Jim Gao · 2018
Later among the works it cites.
Bayesian action decoder for deep multi-agent reinforcement learning
Jakob N Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling · 2018
Later among the works it cites.
Learning attentional communication for multi-agent cooperation
Jiechuan Jiang and Zongqing Lu · 2018
Later among the works it cites.
Adaptive multi-agents synchronization for collaborative driving of autonomous vehicles with multiple communication delays
Alberto Petrillo, Alessandro Salvi, Stefania Santini, and Antonio Saverio Valente · 2018
Later among the works it cites.
Feudal multi-agent hierarchies for cooperative reinforcement learning
Sanjeevan Ahilan and Peter Dayan · 2019
Closest in time.
Actor-attention-critic for multi-agent reinforcement learning
Shariq Iqbal and Fei Sha · 2019
Closest in time.
Learning to schedule communication in multi-agent reinforcement learning
Daewoo Kim, Sangwoo Moon, David Hostallero, Wan Ju Kang, Taeyoung Lee, Kyunghwan Son, and Yung Yi · 2019
Closest in time.
Learning when to communicate at scale in multiagent cooperative and competitive tasks
Amanpreet Singh, Tushar Jain, and Sainbayar Sukhbaatar · 2019
Closest in time.
Probabilistic recursive reasoning for multi-agent reinforcement learning
Ying Wen, Yaodong Yang, Rui Luo, Jun Wang, and Wei Pan · 2019
Closest in time.