Fetching the paper…
Reading the bibliography…
When should we encourage specialization in multi-agent systems versus train generalists that perform the entire task independently? We propose that specialization largely depends on task parallelizability: the potential for multiple agents to execute task components concurrently.
Validity of the single processor approach to achieving large scale computing capabilities
Gene M Amdahl · 1967
Earlier work this paper cites.
The division of cognitive labor
Philip Kitcher · 1990
Earlier work this paper cites.
Highly parallel computing
George S Almasi and Allan Gottlieb · 1994
Earlier work this paper cites.
Heterogeneous computing machines and Amdahl’s Law
David Moncrieff, Richard E. Overill, and Stephen Wilson · 1996
Earlier work this paper cites.
Is parallelism for you?
C Pancake · 1996
Earlier work this paper cites.
Size and complexity among multicellular organisms
Graham Bell and Arne O Mooers · 1997
Earlier work this paper cites.
Specialization in multi-agent systems through learning
Antonio Murciano, José del R. Millán, and Javier Zamora · 1997
Earlier work this paper cites.
Task partitioning in insect societies
Francis LW Ratnieks and Carl Anderson · 1999
Earlier work this paper cites.
Deadlock avoidance in sequential resource allocation systems with multiple resource acquisitions and flexible routings
Jonghun Park and Spyros A Reveliotis · 2001
Earlier work this paper cites.
Prometheus: A methodology for developing intelligent agents
Lin Padgham and Michael Winikoff · 2002
Earlier work this paper cites.
Q-cut—dynamic discovery of sub-goals in reinforcement learning
Ishai Menache, Shie Mannor, and Nahum Shimkin · 2002
Earlier work this paper cites.
Roma: Multi-agent reinforcement learning with emergent roles
Tonghan Wang, Heng Dong, Victor Lesser, and Chongjie Zhang · 2003
Earlier work this paper cites.
A new metric for probability distributions
Dominik Maria Endres and Johannes E Schindelin · 2003
Earlier work this paper cites.
Jensen-Shannon divergence and Hilbert space embedding
Bent Fuglede and Flemming Topsoe · 2004
Earlier work this paper cites.
The architecture of complex weighted networks
Alain Barrat, Marc Barthelemy, Romualdo Pastor-Satorras, and Alessandro Vespignani · 2004
Earlier work this paper cites.
The evolution of worker caste diversity in social insects
Else J Fjerdingstad and Ross H Crozier · 2006
Earlier work this paper cites.
Role-based multi-agent systems
Haibin Zhu and MengChu Zhou · 2008
Earlier work this paper cites.
Multiagent systems: Algorithmic, game-theoretic, and logical foundations
Yoav Shoham and Kevin Leyton-Brown · 2008
Earlier work this paper cites.
Amdahl’s law in the multicore era
Mark D Hill and Michael R Marty · 2008
Earlier work this paper cites.
Rode: Learning roles to decompose multi-agent tasks
Tonghan Wang, Tarun Gupta, Anuj Mahajan, Bei Peng, Shimon Whiteson, and Chongjie Zhang · 2010
Earlier work this paper cites.
Addressing shared resource contention in multicore processors via scheduling
Sergey Zhuravlev, Sergey Blagodurov, and Alexandra Fedorova · 2010
Earlier work this paper cites.
Self-organization for coordinating decentralized reinforcement learning
Chongjie Zhang, Victor R Lesser, and Sherief Abdallah · 2010
Earlier work this paper cites.
Beyond Amdahl’s law: An objective function that links multiprocessor performance gains to delay and energy
Andrew S Cassidy and Andreas G Andreou · 2011
Cited alongside, same era.
Structured parallel programming: Patterns for efficient computation
Michael McCool, James Reinders, and Arch Robison · 2012
Cited alongside, same era.
Evolution of functional specialization and division of labor
Claus Rueffler, Joachim Hermisson, and Günter P Wagner · 2012
Cited alongside, same era.
Optimal behavioral hierarchy
Alec Solway, Carlos Diuk, Natalia Córdova, Debbie Yee, Andrew G Barto, Yael Niv, and Matthew M Botvinick · 2014
Cited alongside, same era.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Cited alongside, same era.
Towards smart factory for industry 4.0: A self-organized multi-agent system with big data based feedback and coordination
Too many cooks: Bayesian inference for coordinating multi-agent collaboration
Sarah A Wu, Rose E Wang, James A Evans, Joshua B Tenenbaum, David C Parkes, and Max Kleiman-Weiner · 2021
Later among the works it cites.
On the importance of environments in human-robot coordination
Matthew C Fontaine, Ya-Chuan Hsu, Yulun Zhang, Bryon Tjanaka, and Stefanos Nikolaidis · 2021
Later among the works it cites.
Case-study analysis of warehouse process optimization
Margareta Živičnjak, Kristijan Rogić, and Ivona Bajor · 2022
Later among the works it cites.
Discovered policy optimisation
Chris Lu, Jakub Grudzien Kuba, Alistair Letcher, Luke Metz, Christian Schroeder de Witt, and Jakob Foerster · 2022
Later among the works it cites.
Resource sharing is sufficient for the emergence of division of labour
Jan J Kreider, Thijs Janzen, Abel Bernadou, Daniel Elsner, Boris H Kramer, and Franz J Weissing · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shiyong Wang, Jiafu Wan, Daqiang Zhang, Di Li, and Chunhua Zhang · 2016
Cited alongside, same era.
A multi-agent reinforcement learning model of common-pool resource appropriation
Julien Perolat, Joel Z Leibo, Vinicius Zambaldi, Charles Beattie, Karl Tuyls, and Thore Graepel · 2017
Cited alongside, same era.
Evaluating and mitigating bandwidth bottlenecks across the memory hierarchy in gpus
Saumay Dublish, Vijay Nagarajan, and Nigel Topham · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Improved Adam optimizer for deep neural networks
Zijun Zhang · 2018
Cited alongside, same era.
When being flexible matters: Ecological underpinnings for the evolution of collective flexibility and task allocation
Merlijn Staps and Corina E Tarnita · 2022
Later among the works it cites.
Quantifying the effects of environment and population diversity in multi-agent reinforcement learning
Kevin R McKee, Joel Z Leibo, Charlie Beattie, and Richard Everett · 2022
Later among the works it cites.
Capability-aware task allocation and team formation analysis for cooperative exploration of complex environments
Muhammad Fadhil Ginting, Kyohei Otsu, Mykel J Kochenderfer, and Ali-Akbar Agha-Mohammadi · 2022
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior. arxiv
Joon Sung Park, Joseph C O’Brien, Carrie J Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein · 2023
Later among the works it cites.
Heterogeneous multi-robot reinforcement learning
Matteo Bettini, Ajay Shankar, and Amanda Prorok · 2023
Later among the works it cites.
Cultural specialization as a double-edged sword: division into specialized guilds might promote cultural complexity at the cost of higher susceptibility to cultural loss
Yotam Ben-Oren, Oren Kolodny, and Nicole Creanza · 2023
Later among the works it cites.
Jaxmarl: Multi-agent rl environments in jax
Alexander Rutherford, Benjamin Ellis, Matteo Gallici, Jonathan Cook, Andrei Lupu, Gardar Ingvarsson, Timon Willi, Akbir Khan, Christian Schroeder de Witt, Alexandra Souly, Saptarashmi Bandyopadhyay, Mikayel Samvelyan, Minqi Jiang, Robert Tjarko Lange, Shimon Whiteson, Bruno Lacerda, Nick Hawes, Tim Rocktaschel, Chris Lu, and Jakob Nicolaus Foerster · 2023
Later among the works it cites.
Zhenhailong Wang, Shaoguang Mao, Wenshan Wu, Tao Ge, Furu Wei, and Heng Ji · 2023
Later among the works it cites.
Breaking the mold: The challenge of large scale MARL specialization
Stefan Juang, Hugh Cao, Arielle Zhou, Ruochen Liu, Nevin L Zhang, and Elvis Liu · 2024
Later among the works it cites.
Controlling behavioral diversity in multi-agent reinforcement learning
Matteo Bettini, Ryan Kortvelesy, and Amanda Prorok · 2024
Later among the works it cites.
The virtual lab: AI agents design new Sars-Cov-2 nanobodies with experimental validation
Kyle Swanson, Wesley Wu, Nash L Bulaong, John E Pak, and James Zou · 2024
Later among the works it cites.
The emergence of specialized roles within groups
Robert L Goldstone, Edgar J Andrade-Lotero, Robert D Hawkins, and Michael E Roberts · 2024
Later among the works it cites.
The rise and fall of technological development in virtual communities
Natalia Vélez, Charley M Wu, Samuel J Gershman, and Eric Schulz · 2024
Later among the works it cites.
Multi-agent reinforcement learning: Foundations and modern approaches
Stefano V Albrecht, Filippos Christianos, and Lukas Schäfer · 2024
Later among the works it cites.
People evaluate idle collaborators based on their impact on task efficiency
Elizabeth A Mieczkowski, Cameron Rouse Turner, Natalia Vélez, and Thomas L Griffiths · 2024
Later among the works it cites.
Meta-prompting: Enhancing language models with task-agnostic scaffolding
Mirac Suzgun and Adam Tauman Kalai · 2024
Later among the works it cites.
A dynamic LLM-powered agent network for task-oriented agent collaboration
Zijun Liu, Yanzhe Zhang, Peng Li, Yang Liu, and Diyi Yang · 2024
Later among the works it cites.