Fetching the paper…
Reading the bibliography…
Role-based learning holds the promise of achieving scalable multi-agent learning by decomposing complex tasks using roles.
The wealth of nations [1776], 1937
Adam Smith · 1937
Earlier work this paper cites.
Feudal reinforcement learning
Peter Dayan and Geoffrey E Hinton · 1993
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
The organization of work in social insect colonies
Deborah M Gordon · 1996
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Earlier work this paper cites.
Soda: Societies and infrastructures in the analysis and design of agent-based systems
Andrea Omicini · 2000
Earlier work this paper cites.
The gaia methodology for agent-oriented analysis and design
Michael Wooldridge, Nicholas R Jennings, and David Kinny · 2000
Earlier work this paper cites.
Prometheus: A methodology for developing intelligent agents
Lin Padgham and Michael Winikoff · 2002
Earlier work this paper cites.
Agent oriented software engineering with ingenias
Juan Pavón and Jorge Gómez-Sanz · 2003
Earlier work this paper cites.
Tropos: An agent-oriented software development methodology
Paolo Bresciani, Anna Perini, Paolo Giorgini, Fausto Giunchiglia, and John Mylopoulos · 2004
Earlier work this paper cites.
Reinforcement learning with factored states and actions
Brian Sallans and Geoffrey E Hinton · 2004
Earlier work this paper cites.
The passi and agile passi mas meta-models compared with a unifying proposal
Massimo Cossentino, Salvatore Gaglio, Luca Sabatucci, and Valeria Seidita · 2005
Earlier work this paper cites.
Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance
Raphaël Jeanson, Penelope F Kukuk, and Jennifer H Fewell · 2005
Earlier work this paper cites.
Collaborative multiagent reinforcement learning by payoff propagation
Jelle R Kok and Nikos Vlassis · 2006
Earlier work this paper cites.
Off-policy multi-agent decomposed policy gradients
Yihan Wang, Beining Han, Tonghan Wang, Heng Dong, and Chongjie Zhang · 2007
Earlier work this paper cites.
Qplex: Duplex dueling multi-agent q-learning
Jianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu, and Chongjie Zhang · 2008
Earlier work this paper cites.
Role-based multi-agent systems
Haibin Zhu and MengChu Zhou · 2008
Earlier work this paper cites.
Using continuous action spaces to solve discrete problems
Hado Van Hasselt and Marco A Wiering · 2009
Earlier work this paper cites.
O-mase: a customisable approach to designing and building complex, adaptive multi-agent systems
Scott A DeLoach and Juan Carlos Garcia-Ojeda · 2010
Earlier work this paper cites.
Using aseme methodology for model-driven agent systems development
Nikolaos Spanoudakis and Pavlos Moraitis · 2010
Earlier work this paper cites.
Bayesian policy search for multi-agent role discovery
Aaron Wilson, Alan Fern, and Prasad Tadepalli · 2010
Earlier work this paper cites.
Self-organization for coordinating decentralized reinforcement learning
Chongjie Zhang, Victor R Lesser, and Sherief Abdallah · 2010
Earlier work this paper cites.
Generalized value functions for large action sets
Jason Pazis and Ronald Parr · 2011
Earlier work this paper cites.
Coordinated multi-agent reinforcement learning in networked distributed pomdps
Chongjie Zhang and Victor Lesser · 2011
Earlier work this paper cites.
The condensed wealth of nations
Eamonn Butler · 2012
Earlier work this paper cites.
Hierarchical relative entropy policy search
Christian Daniel, Gerhard Neumann, and Jan Peters · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Earlier work this paper cites.
Coordinating multi-agent reinforcement learning with limited communication
Chongjie Zhang and Victor Lesser · 2013
Earlier work this paper cites.
Adelfe 2.0: Handbook on agent-oriented design processes, m. cossentino, v. hilaire, a. molesini, and v. seidita, 2014
N Bonjean, W Mefteh, MP Gleizes, C Maurel, and F Migeon · 2014
Earlier work this paper cites.
Hierarchical reinforcement learning: a survey
Mostafa Al-Emran · 2015
Earlier work this paper cites.
Deep reinforcement learning in large discrete action spaces
Gabriel Dulac-Arnold, Richard Evans, Hado van Hasselt, Peter Sunehag, Timothy Lillicrap, Jonathan Hunt, Timothy Mann, Theophane Weber, Thomas Degris, and Ben Coppin · 2015
Earlier work this paper cites.
Online symbolic gradient-based optimization for factored action mdps
Hao Cui and Roni Khardon · 2016
Earlier work this paper cites.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson · 2016
Cited alongside, same era.
Karol Gregor, Danilo Jimenez Rezende, and Daan Wierstra · 2016
Cited alongside, same era.
David Ha, Andrew Dai, and Quoc V Le · 2016
Cited alongside, same era.
Deep reinforcement learning with a natural language action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf · 2016
Cited alongside, same era.
A concise introduction to decentralized POMDPs , volume 1
Frans A Oliehoek, Christopher Amato, et al · 2016
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemyslaw Debiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al · 2019
Later among the works it cites.
Learning action representations for reinforcement learning
Yash Chandak, Georgios Theocharous, James Kostas, Scott Jordan, and Philip Thomas · 2019
Later among the works it cites.
Learning action-transferable policy with action embedding
Yu Chen, Yingfeng Chen, Yu Yang, Ying Li, Jianwei Yin, and Changjie Fan · 2019
Later among the works it cites.
Tarmac: Targeted multi-agent communication
Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Mike Rabbat, and Joelle Pineau · 2019
Later among the works it cites.
Feature control as intrinsic motivation for hierarchical reinforcement learning
Nat Dilokthanakul, Christos Kaplanis, Nick Pawlowski, and Murray Shanahan · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sainbayar Sukhbaatar, Rob Fergus, et al · 2016
Cited alongside, same era.
Cooperative multi-agent control using deep reinforcement learning
Jayesh K Gupta, Maxim Egorov, and Mykel Kochenderfer · 2017
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Diff-dac: Distributed actor-critic for multitask deep reinforcement learning
Sergio Valcarcel Macua, Aleksi Tukiainen, Daniel García-Ocaña Hernández, David Baldazo, Enrique Munoz de Cote, and Santiago Zazo · 2017
Cited alongside, same era.
Sahil Sharma, Aravind Suresh, Rahul Ramesh, and Balaraman Ravindran · 2017
Cited alongside, same era.
Feudal networks for hierarchical reinforcement learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul, Nicolas Heess, Max Jaderberg, David Silver, and Koray Kavukcuoglu · 2017
Cited alongside, same era.
Later among the works it cites.
Hierarchical policy learning is sensitive to goal space design
Zach Dwiel, Madhavun Candadai, Mariano Phielipp, and Arjun K Bansal · 2019
Later among the works it cites.
Actor-attention-critic for multi-agent reinforcement learning
Shariq Iqbal and Fei Sha · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2019
Later among the works it cites.
Social influence as intrinsic motivation for multi-agent deep reinforcement learning
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro Ortega, Dj Strouse, Joel Z Leibo, and Nando De Freitas · 2019
Later among the works it cites.
Emi: Exploration with mutual information
Hyoungseok Kim, Jaekyeom Kim, Yeonwoo Jeong, Sergey Levine, and Hyun Oh Song · 2019
Later among the works it cites.
Learning to coordinate manipulation skills via skill behavior diversification
Youngwoon Lee, Jingyun Yang, and Joseph J Lim · 2019
Later among the works it cites.
Maven: Multi-agent variational exploration
Anuj Mahajan, Tabish Rashid, Mikayel Samvelyan, and Shimon Whiteson · 2019
Later among the works it cites.
Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation
Suraj Nair and Chelsea Finn · 2019
Later among the works it cites.
Planning with goal-conditioned policies
Soroush Nasiriany, Vitchyr Pong, Steven Lin, and Sergey Levine · 2019
Later among the works it cites.
When does communication learning need hierarchical multi-agent deep reinforcement learning
Marie Ossenkopf, Mackenzie Jorgensen, and Kurt Geihs · 2019
Later among the works it cites.
The starcraft multi-agent challenge
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson · 2019
Later among the works it cites.
Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Kyunghwan Son, Daewoo Kim, Wan Ju Kang, David Earl Hostallero, and Yung Yi · 2019
Later among the works it cites.
A multi-agent off-policy actor-critic algorithm for distributed reinforcement learning
Wesley Suttle, Zhuoran Yang, Kaiqing Zhang, Zhaoran Wang, Tamer Basar, and Ji Liu · 2019
Later among the works it cites.
The natural language of actions
Guy Tennenholtz and Shie Mannor · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
Probabilistic recursive reasoning for multi-agent reinforcement learning
Ying Wen, Yaodong Yang, Rui Luo, Jun Wang, and Wei Pan · 2019
Later among the works it cites.
Hierarchical cooperative multi-agent reinforcement learning with skill discovery
Jiachen Yang, Igor Borovikov, and Hongyuan Zha · 2019
Later among the works it cites.
Distributed off-policy actor-critic reinforcement learning with policy consensus
Yan Zhang and Michael M Zavlanos · 2019
Later among the works it cites.
Generalization to new actions in reinforcement learning
Jain Ayush, Szot Andrew, and J. Lim Joseph · 2020
Closest in time.
Emergent tool use from multi-agent autocurricula
Bowen Baker, Ingmar Kanitscheider, Todor Markov, Yi Wu, Glenn Powell, Bob McGrew, and Igor Mordatch · 2020
Closest in time.
Deep coordination graphs
Wendelin Böhmer, Vitaly Kurin, and Shimon Whiteson · 2020
Closest in time.
Growing action spaces
Gregory Farquhar, Laura Gustafson, Zeming Lin, Shimon Whiteson, Nicolas Usunier, and Gabriel Synnaeve · 2020
Closest in time.
Monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder De Witt, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson · 2020
Closest in time.
Learning robot skills with temporal variational inference
Tanmay Shankar and Abhinav Gupta · 2020
Closest in time.
Dynamics-aware unsupervised discovery of skills
Archit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar, and Karol Hausman · 2020
Closest in time.
Reinforcement learning with task decomposition for cooperative multiagent systems
Changyin Sun, Wenzhang Liu, and Lu Dong · 2020
Closest in time.
Options as responses: Grounding behavioural hierarchies in multi-agent rl
Alexander Sasha Vezhnevets, Yuhuai Wu, Rémi Leblond, and Joel Z Leibo · 2020
Closest in time.