Fetching the paper…
Reading the bibliography…
Single-Agent (SA) Reinforcement Learning systems have shown outstanding re-sults on non-stationary problems.
Improving Coordination in Small-Scale Multi-Agent Deep Reinforcement Learning through Memory-driven Communication
Emanuele Pesce and Giovanni Montana · 1901
Earlier work this paper cites.
Learning to Schedule Communication in Multi-agent Reinforcement Learning
Daewoo Kim, Sangwoo Moon, David Hostallero, Wan Ju Kang, Taeyoung Lee, Kyunghwan Son, and Yung Yi · 1902
Earlier work this paper cites.
Optimal algebraic Breadth-First Search for sparse graphs
Paul Burkhardt · 1906
Earlier work this paper cites.
Deep Reinforcement Learning meets Graph Neural Networks: exploring a routing optimization use case
Paul Almasan, José Suárez-Varela, Arnau Badia-Sampera, Krzysztof Rusek, Pere Barlet-Ros, and Albert Cabellos-Aparicio · 1910
Earlier work this paper cites.
Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar · 1911
Earlier work this paper cites.
An image synthesizer
Ken Perlin · 1985
Earlier work this paper cites.
Distributed problem-solving techniques: A survey
Keith S. Decker · 1987
Earlier work this paper cites.
Department of Computer Science University of Rochester Rochester, NY 14627 email: white@cs.rochester.edu
D Whitehead · 1991
Earlier work this paper cites.
On Team Formation
Philip Cohen, Hector Levesque, and Ira Smith · 1997
Earlier work this paper cites.
Using communication to reduce locality in distributed multiagent learning
MAJA J. MATARIC · 1998
Earlier work this paper cites.
Swarm Intelligence: From Natural to Artificial Systems
Eric Bonabeau, Marco Dorigo, and Guy Theraulaz · 1999
Earlier work this paper cites.
Learning to Cooperate via Policy Search, 2000
Leonid Peshkin, Kee-Eung Kim, Nicolas Meuleau, and Leslie Pack Kaelnling · 2000
Earlier work this paper cites.
An Algorithm for Distributed Reinforcement Learning in Cooperative Multi-Agent Systems
Martin Lauer and Martin Riedmiller · 2000
Earlier work this paper cites.
Communication decisions in multi-agent cooperation: model and experiments
Ping Xuan, Victor Lesser, and Shlomo Zilberstein · 2001
Earlier work this paper cites.
Coordinated Reinforcement Learning
Carlos Guestrin, Michail Lagoudakis, and Ronald Parr · 2002
Earlier work this paper cites.
The Communicative Multiagent Team Decision Problem: Analyzing Teamwork Theories and Models
D. V. Pynadath and M. Tambe · 2002
Earlier work this paper cites.
Communication efficiency in multi-agent systems
M. Berna-Koes, I. Nourbakhsh, and K. Sycara · 2004
Earlier work this paper cites.
Cooperative Multi-Agent Learning: The State of the Art
Liviu Panait and Sean Luke · 2005
Earlier work this paper cites.
Understanding and sharing intentions: the origins of cultural cognition
Michael Tomasello, Malinda Carpenter, Josep Call, Tanya Behne, and Henrike Moll · 2005
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini · 2009
Earlier work this paper cites.
Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations
Yoav Shoham · 2009
Earlier work this paper cites.
Wayback Machine, July 2011
Joanne Haddon · 2011
Cited alongside, same era.
Independent reinforcement learners in cooperative Markov games: a survey regarding coordination problems
Laëtitia Matignon, Guillaume J. Laurent, and Nadine Le Fort-Piat · 2012
Cited alongside, same era.
Decentralized POMDPs
Frans A. Oliehoek · 2012
Cited alongside, same era.
Analysis of the Depth First Search Algorithms
N. Kaur and D. Garg · 2012
Cited alongside, same era.
Teaching and leading an ad hoc teammate: Collaboration without pre-coordination
Peter Stone, Gal A. Kaminka, Sarit Kraus, Jeffrey S. Rosenschein, and Noa Agmon · 2013
Cited alongside, same era.
Reinforcement Learning in Robotics: A Survey
Jens Kober, J Andrew Bagnell, and Jan Peters · 2013
Cited alongside, same era.
Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments
Ryan Lowe, YI WU, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Later among the works it cites.
Gated Graph Sequence Neural Networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard Zemel · 2017
Later among the works it cites.
Convolutional Neural Networks on Graphs with Fast Localized Spectral Filtering
Michaël Defferrard, Xavier Bresson, and Pierre Vandergheynst · 2017
Later among the works it cites.
Emergence of Grounded Compositional Language in Multi-Agent Populations
Igor Mordatch and Pieter Abbeel · 2018
Later among the works it cites.
Simplified PPO-Clip Objective, July 2018
Joshua Achiam · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Coordinating Multi-Agent Reinforcement Learning with Limited Communication
Chongjie Zhang and Victor Lesser · 2013
Cited alongside, same era.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Cited alongside, same era.
DeepWalk: Online Learning of Social Representations
Bryan Perozzi, Rami Al-Rfou, and Steven Skiena · 2014
Cited alongside, same era.
Industrial Agents: Emerging Applications of Software Agents in Industry
Paulo Leitão and Stamatis Karnouskos · 2015
Cited alongside, same era.
Distributed multi-agent optimization subject to nonidentical constraints and communication delays
Peng Lin, Wei Ren, and Yongduan Song · 2016
Cited alongside, same era.
Learning Multiagent Communication with Backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2016
Cited alongside, same era.
Petar Veličković, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Liò, and Yoshua Bengio · 2018
Later among the works it cites.
Ad Hoc Teamwork With Behavior Switching Agents
Manish Ravula, Shani Alkoby, and Peter Stone · 2019
Later among the works it cites.
A Survey and Critique of Multiagent Deep Reinforcement Learning
Pablo Hernandez-Leal, Bilal Kartal, and Matthew E. Taylor · 2019
Later among the works it cites.
Human-level performance in 3D multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M. Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castañeda, Charles Beattie, Neil C. Rabinowitz, Ari S. Morcos, Avraham Ruderman, Nicolas Sonnerat, Tim Green, Louise Deason, Joel Z. Leibo, David Silver, Demis Hassabis, Koray Kavukcuoglu, and Thore Graepel · 2019
Later among the works it cites.
A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity
Pablo Hernandez-Leal, Michael Kaisers, Tim Baarslag, and Enrique Munoz de Cote · 2019
Later among the works it cites.
Message-Dropout: An Efficient Training Method for Multi-Agent Deep Reinforcement Learning
Woojun Kim, Myungsik Cho, and Youngchul Sung · 2019
Later among the works it cites.
An Introduction to Proximal Policy Optimization (PPO) in Deep Reinforcement Learning, April 2019
Udacity-DeepRL · 2019
Later among the works it cites.
The Complete Reinforcement Learning Dictionary, November 2019
Shaked Zychlinski · 2019
Later among the works it cites.
Learning to Communicate Proactively in Human-Agent Teaming
Emma M. van Zoelen, Anita Cremers, Frank P. M. Dignum, Jurriaan van Diggelen, and Marieke M. Peeters · 2020
Later among the works it cites.
A Penny for Your Thoughts: The Value of Communication in Ad Hoc Teamwork
Reuth Mirsky, William Macke, Andy Wang, Harel Yedidsion, and Peter Stone · 2020
Later among the works it cites.
Learning Decentralized Controllers for Robot Swarms with Graph Neural Networks
Ekaterina Tolstaya, Fernando Gama, James Paulos, George Pappas, Vijay Kumar, and Alejandro Ribeiro · 2020
Later among the works it cites.
Graph neural networks: A review of methods and applications
Jie Zhou, Ganqu Cui, Shengding Hu, Zhengyan Zhang, Cheng Yang, Zhiyuan Liu, Lifeng Wang, Changcheng Li, and Maosong Sun · 2020
Later among the works it cites.
Expected Value of Communication for Planning in Ad Hoc Teamwork
William Macke, Reuth Mirsky, and Peter Stone · 2021
Closest in time.
Learning Connectivity for Data Distribution in Robot Teams
Ekaterina Tolstaya, Landon Butler, Daniel Mox, James Paulos, Vijay Kumar, and Alejandro Ribeiro · 2021
Closest in time.
Proximal Policy Optimization — Spinning Up documentation, 2021
Spinning Up OpenAI · 2021
Closest in time.