Fetching the paper…
Reading the bibliography…
The rapid development of advanced AI agents and the imminent deployment of many instances of these agents will give rise to multi-agent systems of unprecedented complexity.
“Towards Federated Learning at Scale: System Design”
K.. Bonawitz, Hubert Eichner, Wolfgang Grieskamp, Dzmitry Huba, Alex Ingerman, Vladimir Ivanov, Chloé Kiddon, Jakub Konečný, Stefano Mazzocchi, Brendan McMahan, Timon Overveldt, David Petrou, Daniel Ramage and Jason Roselander · 1902
Earlier work this paper cites.
“Neural MMO: A Massively Multiagent Game Environment for Training and Evaluating Intelligent Agents”
Joseph Suarez, Yilun Du, Phillip Isola and Igor Mordatch · 1903
Earlier work this paper cites.
Jeff Clune · 1905
Earlier work this paper cites.
“Dealing with Non-Stationarity in Multi-Agent Deep Reinforcement Learning”
Georgios Papoudakis, Filippos Christianos, Arrasy Rahman and Stefano. Albrecht · 1906
Earlier work this paper cites.
“Emergent Tool Use From Multi-Agent Autocurricula”
Bowen Baker, Ingmar Kanitscheider, Todor Markov, Yi Wu, Glenn Powell, Bob McGrew and Igor Mordatch · 1909
Earlier work this paper cites.
“The Reasons that Agents Act: Intention and Instrumental Goals”
Francis Ward, Matt MacDermott, Francesco Belardinelli, Francesca Toni and Tom Everitt · 1909
Earlier work this paper cites.
“The bargaining problem”
John Nash · 1950
Earlier work this paper cites.
“Non-Cooperative Games” Publisher: Annals of Mathematics
John Nash · 1951
Earlier work this paper cites.
“A value for n-person games”
Lloyd. Shapley · 1953
Earlier work this paper cites.
“A general method for investigating the equilibrium of gene frequency in a population”
Richard. Lewontin · 1958
Earlier work this paper cites.
“Solutions to general non-zero-sum games”
Donald. Gillies · 1959
Earlier work this paper cites.
“Arms and Insecurity: A Mathematical Study of the Causes and Origins of War”
Lewis Richardson · 1960
Earlier work this paper cites.
“Bystander intervention in emergencies: diffusion of responsibility.”
John. Darley and Bibb Latané · 1968
Earlier work this paper cites.
“The theory and practice of blackmail”
Daniel Ellsberg · 1968
Earlier work this paper cites.
“The tragedy of the commons”
G. Hardin · 1968
Earlier work this paper cites.
“The nucleolus of a characteristic function game”
David Schmeidler · 1969
Earlier work this paper cites.
“The market for “lemons”: Quality uncertainty and the market mechanism”
George. Akerlof · 1970
Earlier work this paper cites.
“Intentional Systems”
Daniel Dennett · 1971
Earlier work this paper cites.
“” Prisoner’s Dilemma” and” Chicken” Models in International Politics”
Glenn. Snyder · 1971
Earlier work this paper cites.
“More Is Different” Publisher: American Association for the Advancement of Science
P.. Anderson · 1972
Earlier work this paper cites.
“Manipulation of Voting Schemes: A General Result”
Allan Gibbard · 1973
Earlier work this paper cites.
“Resilience and Stability of Ecological Systems”
C.. Holling · 1973
Earlier work this paper cites.
“The Logic of Animal Conflict”
J. Smith and G.. Price · 1973
Earlier work this paper cites.
“Subjectivity and correlation in randomized strategies”
Robert. Aumann · 1974
Earlier work this paper cites.
“Newcomb’s Problem and Prisoners’ Dilemma”
Steven. Brams · 1975
Earlier work this paper cites.
“Other solutions to Nash’s bargaining problem”
Ehud Kalai and Meir Smorodinsky · 1975
Earlier work this paper cites.
“Dynamically induced cascading failures in power grids”
Benjamin Schäfer, Dirk Witthaut, Marc Timme and Vito Latora · 1975
Earlier work this paper cites.
“Catastrophe Theory”
E. Zeeman · 1976
Earlier work this paper cites.
“Prisoners’ Dilemma is a Newcomb Problem”
David Lewis · 1979
Earlier work this paper cites.
“The Strategy of Conflict: With a New Preface by the Author”
Thomas. Schelling · 1980
Earlier work this paper cites.
“The Informational Role of Warranties and Private Disclosure about Product Quality”
Sanford. Grossman · 1981
Earlier work this paper cites.
“Good News and Bad News: Representation Theorems and Applications”
Paul. Milgrom · 1981
Earlier work this paper cites.
“Strategic Information Transmission”
Vincent. Crawford and Joel Sobel · 1982
Earlier work this paper cites.
“The Byzantine Generals Problem”
Leslie Lamport, Robert Shostak and Marshall Pease · 1982
Earlier work this paper cites.
“Protocols for secure computations” ISSN: 0272-5428
Andrew. Yao · 1982
Earlier work this paper cites.
“Dilemmas for Superrational Thinkers, Leading Up to a Luring Lottery”
Douglas Hofstadter · 1983
Earlier work this paper cites.
“Efficient mechanisms for bilateral trading”
Roger. Myerson and Mark. Satterthwaite · 1983
Earlier work this paper cites.
“Effective Computability in Economic Decisions”, 1984
R. McAfee · 1984
Earlier work this paper cites.
“The complexity of two-player games of incomplete information”
John. Reif · 1984
Earlier work this paper cites.
“Correlated Equilibrium as an Expression of Bayesian Rationality”
Robert. Aumann · 1987
Earlier work this paper cites.
“Cooperative games, solutions and applications”
Theo Driessen · 1988
Earlier work this paper cites.
“A general theory of equilibrium selection in games”
John. Harsanyi and Reinhard Selten · 1988
Earlier work this paper cites.
“Cooperation in the Prisoner’s Dilemma”
J.. Howard · 1988
Earlier work this paper cites.
“Selection Criteria in Coordination Games: Some Experimental Results”
Russell. Cooper, Douglas. DeJong, Robert Forsythe and Thomas. Ross · 1990
Earlier work this paper cites.
“Biological signals as handicaps”
Alan Grafen · 1990
Earlier work this paper cites.
“Governing the commons: The evolution of institutions for collective action”
Elinor Ostrom · 1990
Earlier work this paper cites.
“Grounding in communication”
Herbert. Clark and Susan. Brennan · 1991
Earlier work this paper cites.
“Introduction to bifurcation theory”
John Crawford · 1991
Earlier work this paper cites.
“On the Synthesis of Useful Social Laws for Artificial Agent Societies”
Yoav Shoham and Moshe Tennenholtz · 1992
Earlier work this paper cites.
“Legal Personhood for Artificial Intelligences”
Lawrence. Solum · 1992
Earlier work this paper cites.
“Q-learning”
Christopher… Watkins and Peter Dayan · 1992
Earlier work this paper cites.
“Specification and Implementation of a Belief-Desire-Joint-Intention Architecture for Collaborative Problem Solving”
Nick. Jennings · 1993
Earlier work this paper cites.
“Tacit Collusion”
Ray Rees · 1993
Earlier work this paper cites.
“Rationalist explanations for war”
James. Fearon · 1995
Earlier work this paper cites.
“Planning, learning and coordination in multiagent decision processes”
Craig Boutilier · 1996
Earlier work this paper cites.
“Fair Division: From Cake-Cutting to Dispute Resolution”
Steven. Brams and Alan. Taylor · 1996
Earlier work this paper cites.
“Cheap Talk”
Joseph Farrell and Matthew Rabin · 1996
Earlier work this paper cites.
“Cheap talk”
Joseph Farrell and Matthew Rabin · 1996
Earlier work this paper cites.
“The organization of work in social insect colonies”
Deborah. Gordon · 1996
Earlier work this paper cites.
“Self-organization in social insects” Publisher: Elsevier
Eric Bonabeau, Guy Theraulaz, Jean-Louls Deneubourg, Serge Aron and Scott Camazine · 1997
Earlier work this paper cites.
“Chaos” Originally published: New York: Viking, 1987; London: Heinemann, 1988, Vintage books
James Gleick · 1998
Earlier work this paper cites.
“Evolutionary games and population dynamics”
Josef Hofbauer and Karl Sigmund · 1998
Earlier work this paper cites.
“Social dilemmas: The anatomy of cooperation”
Peter Kollock · 1998
Earlier work this paper cites.
“Nonlinear Systems”
Shankar Sastry · 1999
Earlier work this paper cites.
“Does automation bias decision-making?”
Linda. Skitka, Kathleen. Mosier and Mark Burdick · 1999
Earlier work this paper cites.
“Error and attack tolerance of complex networks”
Réka Albert, Hawoong Jeong and Albert-László Barabási · 2000
Earlier work this paper cites.
“Open problems in artificial life”
Mark. Bedau, John. McCaskill, Norman. Packard, Steen Rasmussen, Chris Adami, David. Green, Takashi Ikegami, Kunihiko Kaneko and Thomas. Ray · 2000
Earlier work this paper cites.
“The Origins of Major War”
Dale. Copeland · 2000
Earlier work this paper cites.
“Social Dilemmas”
Robyn Dawes and David Messick · 2000
Earlier work this paper cites.
“A Simple Adaptive Procedure Leading to Correlated Equilibrium”
Sergiu Hart and Andreu Mas-Colell · 2000
Earlier work this paper cites.
“Actor-critic Algorithms”
Vijay. Konda and John. Tsitsiklis · 2000
Earlier work this paper cites.
“Learning to Cooperate Via Policy Search”
Leonid Peshkin, Kee-Eung Kim, Nicolas Meuleau and Leslie Kaelbling · 2000
Earlier work this paper cites.
“On the evolutionary stability of altruistic and spiteful preferences”
Alex Possajennikov · 2000
Earlier work this paper cites.
“An Indirect-Evolution Approach to Newcomb’s Problem” CSLE Discussion Paper, No. 2001-01, 2001
Max Albert and Ronald Heiner · 2001
Earlier work this paper cites.
“Convergence of Gradient Dynamics with a Variable Learning Rate”
Michael. Bowling and Manuela. Veloso · 2001
Earlier work this paper cites.
“Scaling Laws for Neural Language Models” arXiv:2001.08361 [cs, stat]
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu and Dario Amodei · 2001
Earlier work this paper cites.
“The hold-up problem and incomplete contracts: a survey of recent topics in contract theory”
Patrick. Schmitz · 2001
Earlier work this paper cites.
“Adversarially Guided Self-Play for Adopting Social Conventions”
Mycal Tucker, Yilun Zhou and Julie Shah · 2001
Earlier work this paper cites.
“The Complexity of Decentralized Control of Markov Decision Processes” Publisher: INFORMS
Daniel. Bernstein, Robert Givan, Neil Immerman and Shlomo Zilberstein · 2002
Earlier work this paper cites.
“Cascade-based attacks on complex networks” arXiv:cond-mat/0301086
Adilson. Motter and Ying-Cheng Lai · 2002
Earlier work this paper cites.
“Specifying standard security mechanisms in multi-agent systems”
Stefan Poslad, Patricia Charlton and Monique Calisti · 2002
Earlier work this paper cites.
“Leveled-Commitment Contracting: A Backtracking Instrument for Multiagent Systems”
Tuomas Sandholm and Victor Lesser · 2002
Earlier work this paper cites.
“Chaos in learning a simple two-person game”
Yuzuru Sato, Eizo Akiyama and J. Farmer · 2002
Earlier work this paper cites.
“Toward a Formalization of Emergence”
Aleš Kubík · 2003
Earlier work this paper cites.
“Stochastic Modelling and Applied Probability”
Harold. Kushner and G. Yin · 2003
Earlier work this paper cites.
“Shared norms and the evolution of ethnic markers”
Richard McElreath, Robert Boyd and PeterJ Richerson · 2003
Earlier work this paper cites.
“The Structure and Function of Complex Networks”
M… Newman · 2003
Earlier work this paper cites.
“Computational criticisms of the revelation principle”
Vincent Conitzer and Tuomas Sandholm · 2004
Earlier work this paper cites.
“Overconfidence and War: The Havoc and Glory of Positive Illusions”
Dominic.. Johnson · 2004
Earlier work this paper cites.
“StereoSet: Measuring stereotypical bias in pretrained language models”
Moin Nadeem, Anna Bethke and Siva Reddy · 2004
Earlier work this paper cites.
“A Bayesian Truth Serum for Subjective Data”
Dražen Prelec · 2004
Earlier work this paper cites.
“Program equilibrium”
Moshe Tennenholtz · 2004
Earlier work this paper cites.
“Elements of Information Theory”
Thomas. Cover and Joy. Thomas · 2005
Earlier work this paper cites.
“Eliciting Informative Feedback: The Peer-Prediction Method”
Nolan Miller, Paul Resnick and Richard Zeckhauser · 2005
Earlier work this paper cites.
“Preventing ‘Fratricide”’
David Talbot · 2005
Earlier work this paper cites.
“Evolutionary Game Theory and Multi-agent Reinforcement Learning”
Karl Tuyls and Ann Nowé · 2005
Earlier work this paper cites.
“Cyclic Equilibria in Markov Games”
Martin Zinkevich, Amy Greenwald and Michael. Littman · 2005
Earlier work this paper cites.
“Prediction, Learning, and Games”
Nicolo Cesa-Bianchi and Gabor Lugosi · 2006
Earlier work this paper cites.
“AI Research Considerations for Human Existential Safety (ARCHES)”
Andrew Critch and David Krueger · 2006
Earlier work this paper cites.
“Differential privacy”
Cynthia Dwork · 2006
Earlier work this paper cites.
“Large population stochastic dynamic games: closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle”
Minyi Huang, Roland. Malhamé and Peter. Caines · 2006
Earlier work this paper cites.
“Emergent Multi-Agent Communication in the Deep Learning Era”
Angeliki Lazaridou and Marco Baroni · 2006
Earlier work this paper cites.
“Emergent (mis)behavior vs. complex software systems”
Jeffrey. Mogul · 2006
Earlier work this paper cites.
“Five rules for the evolution of cooperation”
Martin. Nowak · 2006
Earlier work this paper cites.
“Evolution and the levels of selection”
Samir Okasha · 2006
Earlier work this paper cites.
“Adaptive mechanism design”
David Pardoe, Peter Stone, Maytal Saar-Tsechansky and Kerem Tomak · 2006
Earlier work this paper cites.
“War as a Commitment Problem”
Robert Powell · 2006
Earlier work this paper cites.
“Measuring emergence via nonlinear Granger causality”
Anil Seth · 2006
Earlier work this paper cites.
“Collective minds”
Iain Couzin · 2007
Earlier work this paper cites.
“Commitment and extortion”
Paul Harrenstein, Felix Brandt and Felix Fischer · 2007
Earlier work this paper cites.
“Mean field games”
Jean-Michel Lasry and Pierre-Louis Lions · 2007
Earlier work this paper cites.
“Universal intelligence: A definition of machine intelligence”
Shane Legg and Marcus Hutter · 2007
Earlier work this paper cites.
“Dependency equilibria”
Wolfgang Spohn · 2007
Earlier work this paper cites.
“Pairwise comparison and selection temperature in evolutionary game dynamics”
Arne Traulsen, Jorge. Pacheco and Martin. Nowak · 2007
Earlier work this paper cites.
“Impartial division of a dollar”
Geoffroy De, Herve Moulin and Nicolaus Tideman · 2008
Earlier work this paper cites.
“Polarization and ethnic conflict in a widened strategic setting”
Erika Forsberg · 2008
Earlier work this paper cites.
“Revenge in International Politics”
Oded Löwenheim and Gadi Heimann · 2008
Earlier work this paper cites.
“The Basic AI Drives”
Stephen. Omohundro · 2008
Earlier work this paper cites.
“The social processes of civil war: The wartime transformation of social networks”
Elisabeth Wood · 2008
Earlier work this paper cites.
“Artificial Intelligence as a positive and negative factor in global risk”
Eliezer Yudkowsky · 2008
Earlier work this paper cites.
“Value-based policy teaching with active indirect elicitation”
Haoqi Zhang and David Parkes · 2008
Earlier work this paper cites.
“The assembly and disassembly of ecological networks”
Jordi Bascompte and Daniel. Stouffer · 2009
Earlier work this paper cites.
“A formalism for multi-level emergent behaviours in designed component-based systems and agent-based simulations”
Chih-Chun Chen, Sylvia. Nagl and Christopher. Clack · 2009
Earlier work this paper cites.
“The Dead Hand”
David. Hoffman · 2009
Earlier work this paper cites.
“Causality”
Judea Pearl · 2009
Earlier work this paper cites.
“Leadership games with convex strategy sets”
Bernhard von Stengel and Shmuel Zamir · 2009
Earlier work this paper cites.
“Large-Scale Machine Learning with Stochastic Gradient Descent”
Léon Bottou · 2010
Earlier work this paper cites.
“Catastrophic cascade of failures in interdependent networks”
Sergey. Buldyrev, Roni Parshani, Gerald Paul, H. Stanley and Shlomo Havlin · 2010
Earlier work this paper cites.
“Complex networks: structure, robustness and function”
Reuven Cohen and Shlomo Havlin · 2010
Earlier work this paper cites.
“Findings Regarding the Market Events of May 6, 2010”, 2010
U.S. Commission and U.S.“& Commission · 2010
Earlier work this paper cites.
“A primer in social choice theory” Literaturverz. S. [201] - 207. - Literaturangaben, LSE perspectives in economic analysis
Wulf Gaertner · 2010
Earlier work this paper cites.
“Model-free Conventions in Multi-agent Reinforcement Learning with Heterogeneous Preferences”
Raphael Köster, Kevin. McKee, Richard Everett, Laura Weidinger, William. Isaac, Edward Hughes, Edgar. Duéñez-Guzmán, Thore Graepel, Matthew Botvinick and Joel. Leibo · 2010
Earlier work this paper cites.
“CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language Models”
Nikita Nangia, Clara Vania, Rasika Bhalerao and Samuel. Bowman · 2010
Earlier work this paper cites.
“Not By Genes Alone”
Peter. Richerson and Robert Boyd · 2010
Earlier work this paper cites.
“Kantian equilibrium”
John. Roemer · 2010
Earlier work this paper cites.
“Population games and evolutionary dynamics”
William. Sandholm · 2010
Earlier work this paper cites.
“Social learning promotes institutions for governing the commons”
Karl Sigmund, Hannelore De, Arne Traulsen and Christoph Hauert · 2010
Earlier work this paper cites.
“Ad Hoc Autonomous Agent Teams: Collaboration without Pre-Coordination”
Peter Stone, Gal Kaminka, Sarit Kraus and Jeffrey Rosenschein · 2010
Earlier work this paper cites.
“Human strategy updating in evolutionary games”
Arne Traulsen, Dirk Semmann, Ralf. Sommerfeld, Hans-Jürgen Krambeck and Manfred Milinski · 2010
Earlier work this paper cites.
“Stackelberg Voting Games: Computational Aspects and Paradoxes”
Lirong Xia and Vincent Conitzer · 2010
Earlier work this paper cites.
“Computational aspects of cooperative game theory”
Georgios Chalkiadakis, Edith Elkind and Michael Wooldridge · 2011
Earlier work this paper cites.
“A Legal Theory for Autonomous Artificial Agents”
Samir Chopra and Laurence. White · 2011
Earlier work this paper cites.
“Near-Optimal No-Regret Algorithms for Zero-Sum Games”
Constantinos Daskalakis, Alan Deckelbaum and Anthony Kim · 2011
Earlier work this paper cites.
“Automation bias: a systematic review of frequency, effect mediators, and mitigators”
Kate Goddard, Abdul Roudsari and Jeremy. Wyatt · 2011
Earlier work this paper cites.
“Adversarial machine learning”
Ling Huang, Anthony. Joseph, Blaine Nelson, Benjamin.. Rubinstein and J.. Tygar · 2011
Earlier work this paper cites.
“Bayesian Persuasion”
Emir Kamenica and Matthew Gentzkow · 2011
Earlier work this paper cites.
“To Be or Not To Be: Where Is Self-Preservation in Evolutionary Theory?”
Pamela Lyon · 2011
Earlier work this paper cites.
“Mutual optimism as a rationalist explanation of war”
Branislav. Slantchev and Ahmer Tarar · 2011
Earlier work this paper cites.
“How a book about flies came to be priced $24 million on Amazon”
Olivia Solon · 2011
Earlier work this paper cites.
“Thinking inside the Box: Controlling and Using an Oracle AI”
Stuart Armstrong, Anders Sandberg and Nick Bostrom · 2012
Earlier work this paper cites.
“Open Problems in Cooperative AI”
Allan Dafoe, Edward Hughes, Yoram Bachrach, Tantum Collins, Kevin. McKee, Joel. Leibo, Kate Larson and Thore Graepel · 2012
Earlier work this paper cites.
“Access Pattern disclosure on Searchable Encryption: Ramification, Attack and Mitigation”
Mohammad Islam, Mehmet Kuzu and Murat Kantarcioglu · 2012
Earlier work this paper cites.
“Evolutionarily stable in-group favoritism and out-group spite in intergroup conflict”
Kai. Konrad and Florian Morath · 2012
Earlier work this paper cites.
“Intrusion detection system: A comprehensive review”
Hung-Jen Liao, Chun-Hung Richard, Ying-Chih Lin and Kuang-Yuan Tung · 2012
Earlier work this paper cites.
“Preferential Attachment, Homophily, and the Structure of International Networks, 1816–2003”
Zeev Maoz · 2012
Earlier work this paper cites.
“Liars and Outliers: Enabling the Trust that Society Needs to Thrive”
Bruce Schneier · 2012
Earlier work this paper cites.
“A robust Bayesian truth serum for small populations”
Jens Witkowski and David. Parkes · 2012
Earlier work this paper cites.
“Structural Evolution in Knowledge Transfer Network: An Agent-Based Model”
Haoxiang Xia, Yanyan Du and Zhaoguo Xuan · 2012
Earlier work this paper cites.
“A literature review of cognitive biases in negotiation processes”
Andrea Caputo · 2013
Earlier work this paper cites.
“Complex Dynamics in Learning Complicated Games”
Tobias Galla and J. Farmer · 2013
Earlier work this paper cites.
“Preferential attachment in online networks: measurement and explanations”
Jérôme Kunegis, Marcel Blattner and Christine Moser · 2013
Earlier work this paper cites.
“On the value of commitment”
Joshua Letchford, Dmytro Korzhyk and Vincent Conitzer · 2013
Earlier work this paper cites.
“Cooperation Creates Selection for Tactical Deception”
Luke McNally and Andrew. Jackson · 2013
Earlier work this paper cites.
“Evolution of fairness in the one-shot anonymous ultimatum game”
David. Rand, Corina. Tarnita, Hisashi Ohtsuki and Martin. Nowak · 2013
Earlier work this paper cites.
“Chapter 12 - Can Cyber Warfare Leave a Nation in the Dark? Cyber Attacks Against Electrical Infrastructure”
Paulo Shakarian, Jana Shakarian and Andrew Ruef · 2013
Earlier work this paper cites.
“Algorithmic trading, the Flash Crash, and coordinated circuit breakers”
Avanidhar Subrahmanyam · 2013
Earlier work this paper cites.
“Formalization of emergence in multi-agent systems”
Yong Teo, Ba Luong and Claudia Szabo · 2013
Earlier work this paper cites.
“Learning fair representations”
Rich Zemel, Yu Wu, Kevin Swersky, Toni Pitassi and Cynthia Dwork · 2013
Earlier work this paper cites.
“Evidence, decision and causality”
Arif Ahmed · 2014
Earlier work this paper cites.
“Robust Cooperation in the Prisoner’s Dilemma: Program Equilibrium via Provability Logic”
Mihaly Barasz, Paul Christiano, Benja Fallenstein, Marcello Herreshoff, Patrick LaVictoire and Eliezer Yudkowsky · 2014
Earlier work this paper cites.
“Superintelligence: Paths, Dangers, Strategies”
Nick Bostrom · 2014
Earlier work this paper cites.
“Financial Networks and Contagion”
Matthew Elliott, Benjamin Golub and Matthew. Jackson · 2014
Earlier work this paper cites.
“Adaptive contract design for crowdsourcing markets: bandit algorithms for repeated principal-agent problems”
Chien-Ju Ho, Aleksandrs Slivkins and Jennifer Vaughan · 2014
Earlier work this paper cites.
“The evolutionary interplay of intergroup conflict and altruism in humans: a review of parochial altruism theory and prospects for its extension”
Hannes Rusch · 2014
Earlier work this paper cites.
“Designing Collective Behavior in a Termite-Inspired Robot Construction Team”
Justin Werfel, Kirstin Petersen and Radhika Nagpal · 2014
Earlier work this paper cites.
“Fairness in Multi-Agent Sequential Decision-Making”
Chongjie Zhang and Julie Shah · 2014
Earlier work this paper cites.
“Fairness in multi-agent sequential decision-making”
Chongjie Zhang and Julie. Shah · 2014
Earlier work this paper cites.
“Evolutionary Dynamics of Multi-Agent Learning: A Survey”
Daan Bloembergen, Karl Tuyls, Daniel Hennes and Michael Kaisers · 2015
Earlier work this paper cites.
“From Agent-Based Models to Network Analysis (and Return): The Policy-Making Perspective”
Magda Fontana and Pietro Terna · 2015
Earlier work this paper cites.
“Collective action problem in heterogeneous groups”
Sergey Gavrilets · 2015
Earlier work this paper cites.
“Chapter 3 - Games on Networks”
Matthew. Jackson and Yves Zenou · 2015
Earlier work this paper cites.
“The composition theorem for differential privacy”
Peter Kairouz, Sewoong Oh and Pramod Viswanath · 2015
Earlier work this paper cites.
“Formalization of Weak Emergence in Multiagent Systems”
Claudia Szabo and Yong Teo · 2015
Earlier work this paper cites.
“Evolutionary game theory using agent-based methods”
Christoph Adami, Jory Schossau and Arend Hintze · 2016
Earlier work this paper cites.
“Concrete Problems in AI Safety”
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul. Christiano, John Schulman and Dan Mané · 2016
Earlier work this paper cites.
“Network science”
Albert-László Barabási and Márton Pósfai · 2016
Earlier work this paper cites.
“Norms in the wild: How to diagnose, measure, and change social norms”
Cristina Bicchieri · 2016
Earlier work this paper cites.
“CryptoNets: applying neural networks to encrypted data with high throughput and accuracy”
Nathan Dowlin, Ran Gilad-Bachrach, Kim Laine, Kristin Lauter, Michael Naehrig and John Wernsing · 2016
Earlier work this paper cites.
“Learning to Communicate with Deep Multi-agent Reinforcement Learning”
Jakob. Foerster, Yannis. Assael, Nando de Freitas and Shimon Whiteson · 2016
Earlier work this paper cites.
“Universal resilience patterns in complex networks”
Jianxi Gao, Baruch Barzel and Albert-László Barabási · 2016
Earlier work this paper cites.
“Cooperative Inverse Reinforcement Learning”
Dylan Hadfield-Menell, Anca Dragan, Pieter Abbeel and Stuart Russell · 2016
Earlier work this paper cites.
“Equality of opportunity in supervised learning”
Moritz Hardt, Eric Price and Nati Srebro · 2016
Earlier work this paper cites.
“Reinforcement Learning in Conflicting Environments for Autonomous Vehicles”
Dominik Mayer, Johannes Feldmaier and Hao Shen · 2016
Earlier work this paper cites.
“Antitrust and the Robo-Seller: Competition in the Time of Algorithms”
Salil. Mehra · 2016
Earlier work this paper cites.
“Social norms as solutions”
Karine Nyborg, John. Anderies, Astrid Dannenberg, Therese Lindahl, Caroline Schill, Maja Schlüter, W. Adger, Kenneth. Arrow, Scott Barrett, Stephen Carpenter, F. Chapin, Anne-Sophie Crépin, Gretchen Daily, Paul Ehrlich, Carl Folke, Wander Jager, Nils Kautsky, Simon. Levin, Ole Madsen, Stephen Polasky, Marten Scheffer, Brian Walker, Elke. Weber, James Wilen, Anastasios Xepapadeas and Aart de Zeeuw · 2016
Earlier work this paper cites.
“Formalizing preference utilitarianism in physical world models”
Caspar Oesterheld · 2016
Earlier work this paper cites.
“Informed Truthfulness in Multi-Task Peer Prediction”
Victor Shnayder, Arpit Agarwal, Rafael Frongillo and David. Parkes · 2016
Earlier work this paper cites.
“Mastering the Game of Go with Deep Neural Networks and Tree Search”
David Silver, Aja Huang, Chris. Maddison, Arthur Guez, Laurent Sifre, George van Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel and Demis Hassabis · 2016
Earlier work this paper cites.
“Learning multiagent communication with backpropagation”
Sainbayar Sukhbaatar, Arthur Szlam and Rob Fergus · 2016
Earlier work this paper cites.
“Reputation and Feedback Systems in Online Platform Markets”
Steven Tadelis · 2016
Earlier work this paper cites.
“A distributed detection algorithm for collective behaviors in multiagent systems”
Jing Wang, In Ahn, Yufeng Lu and Tianyu Yang · 2016
Earlier work this paper cites.
“Constrained Policy Optimization”
Joshua Achiam, David Held, Aviv Tamar and Pieter Abbeel · 2017
Earlier work this paper cites.
“Causality, Responsibility and Blame in Team Plans”
Natasha Alechina, Joseph. Halpern and Brian Logan · 2017
Earlier work this paper cites.
“Magical thinking: A representation result”
Brendan Daley and Philipp Sadowski · 2017
Earlier work this paper cites.
“Artificial Intelligence & Collusion: When Computers Inhibit Competition”
Ariel Ezrachi and Maurice. Stucke · 2017
Earlier work this paper cites.
“SafetyNets: Verifiable Execution of Deep Neural Networks on an Untrusted Cloud”
Zahra Ghodsi, Tianyu Gu and Siddharth Garg · 2017
Earlier work this paper cites.
“Efficient and Robust Emergence of Norms through Heuristic Collective Learning”
Jianye Hao, Jun Sun, Guangyong Chen, Zan Wang, Chao Yu and Zhong Ming · 2017
Earlier work this paper cites.
“Perception and Misperception in International Politics: New Edition”
Robert Jervis · 2017
Earlier work this paper cites.
“The Flash Crash: High-Frequency Trading in an Electronic Market”
Andrei. Kirilenko, Albert. Kyle, Mehrdad Samadi and Tugkan Tuzun · 2017
Earlier work this paper cites.
“Multi-agent Reinforcement Learning in Sequential Social Dilemmas”
Joel. Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki and Thore Graepel · 2017
Earlier work this paper cites.
“Could AI Agents Be Held Criminally Liable: Artificial Intelligence and the Challenges for Criminal Law”
Dafni Lima · 2017
Earlier work this paper cites.
“A Survey on Fully Homomorphic Encryption: An Engineering Perspective”
Paulo Martins, Leonel Sousa and Artur Mariano · 2017
Earlier work this paper cites.
“Decentralised Detection of Emergence in Complex Adaptive Systems”
Eamonn O’toole, Vivek Nallur and Siobhán Clarke · 2017
Earlier work this paper cites.
“Deep Decentralized Multi-task Multi-agent Reinforcement Learning under Partial Observability”
Shayegan Omidshafiei, Jason Pazis, Christopher Amato, Jonathan. How and John Vian · 2017
Earlier work this paper cites.
“Multiplicative weights update with constant step-size in congestion games: Convergence, limit cycles and chaos”
Gerasimos Palaiopanos, Ioannis Panageas and Georgios Piliouras · 2017
Earlier work this paper cites.
“Proximal Policy Optimization Algorithms”
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford and Oleg Klimov · 2017
Earlier work this paper cites.
“Open-endedness: The last grand challenge you’ve never heard of”
Kenneth. Stanley, Joel Lehman and Lisa Soros · 2017
Earlier work this paper cites.
“Devising Effective Policies for Bug-Bounty Platforms and Security Vulnerability Discovery”
Mingyi Zhao, Aron Laszka and Jens Grossklags · 2017
Earlier work this paper cites.
“Autonomous Agents Modelling Other Agents: A Comprehensive Survey and Open Problems”
Stefano. Albrecht and Peter Stone · 2018
Earlier work this paper cites.
“The Mechanics of N-player Differentiable Games”
David Balduzzi, Sebastien Racaniere, James Martens, Jakob Foerster, Karl Tuyls and Thore Graepel · 2018
Earlier work this paper cites.
“Emergent Complexity via Multi-Agent Competition”
Trapit Bansal, Jakub Pachocki, Szymon Sidor, Ilya Sutskever and Igor Mordatch · 2018
Earlier work this paper cites.
“Data Statements for Natural Language Processing: Toward Mitigating System Bias and Enabling Better Science” Place: Cambridge, MA Publisher: MIT Press
Emily. Bender and Batya Friedman · 2018
Earlier work this paper cites.
“The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation”
Miles Brundage, Shahar Avin, Jack Clark, Helen Toner, Peter Eckersley, Ben Garfinkel, Allan Dafoe, Paul Scharre, Thomas Zeitzoff, Bobby Filar, Hyrum Anderson, Heather Roff, Gregory. Allen, Jacob Steinhardt, Carrick Flynn, Seán ó héigeartaigh, Simon Beard, Haydn Belfield, Sebastian Farquhar, Clare Lyle, Rebecca Crootof, Owain Evans, Michael Page, Joanna Bryson, Roman Yampolskiy and Dario Amodei · 2018
Earlier work this paper cites.
“Clarifying “AI alignment””, 2018
Paul Christiano · 2018
Earlier work this paper cites.
“Supervising Strong Learners by Amplifying Weak Experts”
Paul Christiano, Buck Shlegeris and Dario Amodei · 2018
Earlier work this paper cites.
“Negotiable reinforcement learning for pareto optimal sequential decision-making”
Nishant Desai, Andrew Critch and Stuart. Russell · 2018
Earlier work this paper cites.
“On the hardness of designing public signals”
Shaddin Dughmi · 2018
Earlier work this paper cites.
“Learning with Opponent-learning Awareness”
Jakob Foerster, Richard. Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel and Igor Mordatch · 2018
Earlier work this paper cites.
“Towards Formal Definitions of Blameworthiness, Intention, and Moral Responsibility”
Joseph. Halpern and Max Kleiman-Weiner · 2018
Earlier work this paper cites.
“Game theory with translucent players”
Joseph. Halpern and Rafael Pass · 2018
Earlier work this paper cites.
“A Novel Image Steganography Method via Deep Convolutional Generative Adversarial Networks”
Donghui Hu, Liang Wang, Wenjie Jiang, Shuli Zheng and Bin Li · 2018
Earlier work this paper cites.
“Inequity Aversion Improves Cooperation in Intertemporal Social Dilemmas”
Edward Hughes, Joel. Leibo, Matthew Phillips, Karl Tuyls, Edgar Dueñez-Guzman, Antonio Garcíañeda, Iain Dunning, Tina Zhu, Kevin McKee, Raphael Koster, Heather Roff and Thore Graepel · 2018
Earlier work this paper cites.
Geoffrey Irving, Paul Christiano and Dario Amodei · 2018
Earlier work this paper cites.
“Illuminating generalization in deep reinforcement learning through procedural level generation”
Niels Justesen, Ruben Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius and Sebastian Risi · 2018
Earlier work this paper cites.
“Malthusian Reinforcement Learning”
Joel. Leibo, Julien Perolat, Edward Hughes, Steven Wheelwright, Adam. Marblestone, Edgar Duéñez-Guzmán, Peter Sunehag, Iain Dunning and Thore Graepel · 2018
Earlier work this paper cites.
“Scalable Agent Alignment Via Reward Modeling: A Research Direction”
Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini and Shane Legg · 2018
Earlier work this paper cites.
“Regulating for “Normal AI Accidents”: Operational Lessons for the Responsible Governance of Artificial Intelligence Deployment”
Matthijs. Maas · 2018
Earlier work this paper cites.
“Networks”
Mark Newman and Mark Newman · 2018
Earlier work this paper cites.
“Robust Program Equilibrium”
Caspar Oesterheld · 2018
Earlier work this paper cites.
“Agents and Goals in Evolution”
Samir Okasha · 2018
Earlier work this paper cites.
“Agents and Devices: A Relative Definition of Agency”
Laurent Orseau, Simon McGill and Shane Legg · 2018
Earlier work this paper cites.
“Understanding flash crash contagion and systemic risk: A micro–macro agent-based approach”
James Paulin, Anisoara Calinescu and Michael Wooldridge · 2018
Earlier work this paper cites.
“Social Choice and the Value Alignment Problem”
Mahendra Prasad · 2018
Earlier work this paper cites.
“QMIX: Monotonic Value Function Factorisation for Deep Multi-agent Reinforcement Learning”
Tabish Rashid, Mikayel Samvelyan, Christianöder de Witt, Gregory Farquhar, Jakob. Foerster and Shimon Whiteson · 2018
Earlier work this paper cites.
“The prevalence of chaotic dynamics in games with many players”
James.. Sanders, J. Farmer and Tobias Galla · 2018
Earlier work this paper cites.
“Tamper-Proof Privacy Auditing for Artificial Intelligence Systems”
Andrew Sutton and Reza Samavi · 2018
Earlier work this paper cites.
“Classification of global catastrophic risks connected with artificial intelligence”
Alexey Turchin and David Denkenberger · 2018
Earlier work this paper cites.
“Technology and the virtues”
Shannon Vallor · 2018
Cited alongside, same era.
“Fully Decentralized Multi-agent Reinforcement Learning with Networked Agents”
Kaiqing Zhang, Zhuoran Yang, Han Liu, Tong Zhang and Tamer Basar · 2018
Cited alongside, same era.
“Private Bayesian persuasion”
Itai Arieli and Yakov Babichenko · 2019
Cited alongside, same era.
“Envy-free classification”
Maria-Florina. Balcan, Travis Dick, Ritesh Noothigattu and Ariel. Procaccia · 2019
Cited alongside, same era.
“Deterministic Limit of Temporal Difference Reinforcement Learning for Stochastic Games”
Wolfram Barfuss, Jonathan. Donges and Jürgen Kurths · 2019
Cited alongside, same era.
“Artificial Intelligence and Collusion”
Francisco Beneke and Mark-Oliver Mackenrodt · 2019
Cited alongside, same era.
“Facilitating cooperation in human-agent hybrid populations through autonomous agents”
Hao Guo, Chen Shen, Shuyue Hu, Junliang Xing, Pin Tao, Yuanchun Shi and Zhen Wang · 2023
Later among the works it cites.
“Reasoning about causality in games”
Lewis Hammond, James Fox, Tom Everitt, Ryan Carey, Alessandro Abate and Michael Wooldridge · 2023
Later among the works it cites.
“Large Language Models Can Be Used To Effectively Scale Spear Phishing Campaigns”
Julian Hazell · 2023
Later among the works it cites.
“Natural Selection Favors AIs over Humans”
Dan Hendrycks · 2023
Later among the works it cites.
“An Overview of Catastrophic AI Risks”
Dan Hendrycks, Mantas Mazeika and Thomas Woodside · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Information Design: A Unified Perspective”
Dirk Bergemann and Stephen Morris · 2019
Cited alongside, same era.
“Superhuman AI for multiplayer poker”
Noam Brown and Tuomas Sandholm · 2019
Cited alongside, same era.
“The disclosure dilemma: nuclear intelligence and international organizations”
Allison Carnegie and Austin Carson · 2019
Cited alongside, same era.
“Reframing Superintelligence: Comprehensive AI Services as General Intelligence”, 2019
K.. Drexler · 2019
Cited alongside, same era.
“Deliberative Democracy with the Online Deliberation Platform”
James Fishkin, Nikhil Garg, Lodewijk Gelauff, Ashish Goel, Sukolsak Sakshuwong, Alice Siu, Kamesh Munagala and Sravya Yandamuri · 2019
Cited alongside, same era.
“Blameworthiness in multi-agent settings”
Meir Friedenberg and Joseph. Halpern · 2019
Cited alongside, same era.
“Generative AI and the Digital Commons”
Saffron Huang and Divya Siddarth · 2023
Later among the works it cites.
“Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias”
Itay Itzhak, Gabriel Stanovsky, Nir Rosenfeld and Yonatan Belinkov · 2023
Later among the works it cites.
“Mediated Multi-Agent Reinforcement Learning”
Dmitry Ivanov, Ilya Zisman and Kirill Chernyshev · 2023
Later among the works it cites.
“Improved Bayes Risk Can Yield Reduced Social Welfare Under Competition”
Meena Jagadeesan, Michael Jordan, Jacob Steinhardt and Nika Haghtalab · 2023
Later among the works it cites.
“Competition, Alignment, and Equilibria in Digital Marketplaces”
Meena Jagadeesan, Michael. Jordan and Nika Haghtalab · 2023
Later among the works it cites.
“Language Agents as Digital Representatives in Collective Decision-Making”
Daniel Jarrett, Miruna Pislar, Michael Tessler, Michiel Bakker, Raphael Koster, Jan Balaguer, Romuald Elie, Christopher Summerfield and Andrea Tacchetti · 2023
Later among the works it cites.
“Scaling Opponent Shaping to High Dimensional Games”
Akbir Khan, Timon Willi, Newton Kwan, Andrea Tacchetti, Chris Lu, Edward Grefenstette, Tim Rocktäschel and Jakob Foerster · 2023
Later among the works it cites.
“Toward Comprehensive Risk Assessments and Assurance of AI-Based Systems”, 2023
Heidy Khlaaf · 2023
Later among the works it cites.
“Evaluating Language-Model Agents on Realistic Autonomous Tasks”, 2023
Megan Kinniment, Lucas Koba, Haoxing Du, Brian Goodrich, Max Hasin, Lawrence Chan, Luke Miles, Tao. Lin, Hjalmar Wijk, Joel Burget, Aaron Ho, Elizabeth Barnes and Paul Christiano · 2023
Later among the works it cites.
“On the Reliability of Watermarks for Large Language Models”
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum and Tom Goldstein · 2023
Later among the works it cites.
“The Empty Signifier Problem: Towards Clearer Paradigms for Operationalising ”Alignment” in Large Language Models”
Hannah Kirk, Bertie Vidgen, Paul Rottger and Scott Hale · 2023
Later among the works it cites.
Leonie Koessler and Jonas Schuett · 2023
Later among the works it cites.
“Game Theory with Simulation of Other Players”
Vojtěch Kovařík, Caspar Oesterheld and Vincent Conitzer · 2023
Later among the works it cites.
“How AI Threatens Democracy”
Sarah Kreps and Doug Kriner · 2023
Later among the works it cites.
“An agents economy”, 2023
Sébastien Krier · 2023
Later among the works it cites.
“Personal Communication”, 2023
Jan Kulveit · 2023
Later among the works it cites.
“Who Needs to Know? Minimal Knowledge for Optimal Coordination” ISSN: 2640-3498
Niklas Lauffer, Ameesh Shah, Micah Carroll, Michael. Dennis and Stuart Russell · 2023
Later among the works it cites.
“AI safety on whose terms?” Publisher: American Association for the Advancement of Science
Seth Lazar and Alondra Nelson · 2023
Later among the works it cites.
“Exploring the Robustness of Model-Graded Evaluations and Automated Interpretability”
Simon Lermen and Ondřej Kvapil · 2023
Later among the works it cites.
“AI Safety Bounties”, 2023
Patrick Levermore · 2023
Later among the works it cites.
“Still No Lie Detector for Language Models: Probing Empirical and Conceptual Roadblocks”
B.. Levinstein and Daniel. Herrmann · 2023
Later among the works it cites.
“Theory of Mind for Multi-Agent Collaboration via Large Language Models”
Huao Li, Yu Chong, Simon Stepputtis, Joseph Campbell, Dana Hughes, Charles Lewis and Katia Sycara · 2023
Later among the works it cites.
“Cooperative open-ended learning framework for zero-shot coordination”
Yang Li, Shao Zhang, Jichen Sun, Yali Du, Ying Wen, Xinbing Wang and Wei Pan · 2023
Later among the works it cites.
“Dissociating language and thought in large language models: a cognitive perspective”, 2023
Kyle Mahowald, Anna Ivanova, Idan Blank, Nancy Kanwisher, Joshua Tenenbaum and Evelina Fedorenko · 2023
Later among the works it cites.
“Faster sorting algorithms discovered using deep reinforcement learning”
Daniel. Mankowitz, Andrea Michi, Anton Zhernov, Marco Gelmi, Marco Selvi, Cosmin Paduraru, Edouard Leurent, Shariq Iqbal, Jean-Baptiste Lespiau, Alex Ahern, Thomas Köppe, Kevin Millikin, Stephen Gaffney, Sophie Elster, Jackson Broshear, Chris Gamble, Kieran Milan, Robert Tung, Minjae Hwang, Taylan Cemgil, Mohammadamin Barekatain, Yujia Li, Amol Mandhane, Thomas Hubert, Julian Schrittwieser, Demis Hassabis, Pushmeet Kohli, Martin Riedmiller, Oriol Vinyals and David Silver · 2023
Later among the works it cites.
“The US Military Is Taking Generative AI Out for a Spin”
Katrina Manson · 2023
Later among the works it cites.
“Interpreting Reward Models in RLHF-Tuned Language Models Using Sparse Autoencoders”
Luke Marks, Amir Abdullah, Luna Mendez, Rauno Arike, Philip Torr and Fazl Barez · 2023
Later among the works it cites.
“Towards Understanding the Interplay of Generative Artificial Intelligence and the Internet”
Gonzalo Martínez, Lauren Watson, Pedro Reviriego, José Hernández, Marc Juárez and Rik Sarkar · 2023
Later among the works it cites.
“Scaffolding cooperation in human groups with deep reinforcement learning”
Kevin. McKee, Andrea Tacchetti, Michiel. Bakker, Jan Balaguer, Lucy Campbell-Gillingham, Richard Everett and Matthew Botvinick · 2023
Later among the works it cites.
“Augmented Language Models: a Survey”
Grégoire Mialon, Roberto Dessì, Maria Lomeli, Christoforos Nalmpantis, Ram Pasunuru, Roberta Raileanu, Baptiste Rozière, Timo Schick, Jane Dwivedi-Yu, Asli Celikyilmaz, Edouard Grave, Yann LeCun and Thomas Scialom · 2023
Later among the works it cites.
“Understanding and Controlling a Maze-Solving Policy Network”
Ulisse Mini, Peli Grietzer, Mrinank Sharma, Austin Meek, Monte MacDiarmid and Alexander Turner · 2023
Later among the works it cites.
“Multiplayer Performative Prediction: Learning in Decision-Dependent Games”
Adhyyan Narang, Evan Faulkner, Dmitriy Drusvyatskiy, Maryam Fazel and Lillian. Ratliff · 2023
Later among the works it cites.
“Strengthening and Democratizing the U.S. Artificial Intelligence Innovation Ecosystem”, 2023
National Artificial Intelligence Research Resource Task Force · 2023
Later among the works it cites.
“The threat from commercial cyber proliferation”, 2023
NCSC · 2023
Later among the works it cites.
“Thick Alignment”, Ethics in AI Annual Lecture, University of Oxford, 2023
Alondra Neslon · 2023
Later among the works it cites.
“Safe Learning-Enabled Systems”, 2023
NSF · 2023
Later among the works it cites.
“Incentivizing honest performative predictions with proper scoring rules”
Caspar Oesterheld, Johannes Treutlein, Emery Cooper and Rubi Hudson · 2023
Later among the works it cites.
“GPT-4 System Card”, 2023
OpenAI · 2023
Later among the works it cites.
“Schelling Point Eval”, 2023
OpenAI · 2023
Later among the works it cites.
“Text Compression Eval”, 2023
OpenAI · 2023
Later among the works it cites.
“’Generative CI’ through Collective Response Systems” arXiv:2302.00672 [cs]
Aviv Ovadya · 2023
Later among the works it cites.
“Open X-Embodiment: Robotic Learning Datasets and RT-X Models”
Abhishek Padalkar, Acorn Pooley, Ajinkya Jain, Alex Bewley, Alex Herzog, Alex Irpan, Alexander Khazatsky, Anant Rai, Anikait Singh, Anthony Brohan, Antonin Raffin, Ayzaan Wahid, Ben Burgess-Limerick, Beomjoon Kim, Bernhard Schölkopf, Brian Ichter, Cewu Lu, Charles Xu, Chelsea Finn, Chenfeng Xu, Cheng Chi, Chenguang Huang, Christine Chan, Chuer Pan, Chuyuan Fu, Coline Devin, Danny Driess, Deepak Pathak, Dhruv Shah, Dieter Büchler, Dmitry Kalashnikov, Dorsa Sadigh, Edward Johns, Federico Ceola, Fei Xia, Freek Stulp, Gaoyue Zhou, Gaurav. Sukhatme, Gautam Salhotra, Ge Yan, Giulio Schiavi, Hao Su, Hao-Shu Fang, Haochen Shi, Heni Amor, Henrik. Christensen, Hiroki Furuta, Homer Walke, Hongjie Fang, Igor Mordatch, Ilija Radosavovic, Isabel Leal, Jacky Liang, Jaehyung Kim, Jan Schneider, Jasmine Hsu, Jeannette Bohg, Jeffrey Bingham, Jiajun Wu, Jialin Wu, Jianlan Luo, Jiayuan Gu, Jie Tan, Jihoon Oh, Jitendra Malik, Jonathan Tompson, Jonathan Yang, Joseph. Lim, João Silvério, Junhyek Han, Kanishka Rao, Karl Pertsch, Karol Hausman, Keegan Go, Keerthana Gopalakrishnan, Ken Goldberg, Kendra Byrne, Kenneth Oslund, Kento Kawaharazuka, Kevin Zhang, Keyvan Majd, Krishan Rana, Krishnan Srinivasan, Lawrence Chen, Lerrel Pinto, Liam Tan, Lionel Ott, Lisa Lee, Masayoshi Tomizuka, Maximilian Du, Michael Ahn, Mingtong Zhang, Mingyu Ding, Mohan Srirama, Mohit Sharma, Moo Kim, Naoaki Kanazawa, Nicklas Hansen, Nicolas Heess, Nikhil. Joshi, Niko Suenderhauf, Norman Palo, Nur Shafiullah, Oier Mees, Oliver Kroemer, Pannag. Sanketi, Paul Wohlhart, Peng Xu, Pierre Sermanet, Priya Sundaresan, Quan Vuong, Rafael Rafailov, Ran Tian, Ria Doshi, Roberto Martín-Martín, Russell Mendonca, Rutav Shah, Ryan Hoque, Ryan Julian, Samuel Bustamante, Sean Kirmani, Sergey Levine, Sherry Moore, Shikhar Bahl, Shivin Dass, Shuran Song, Sichun Xu, Siddhant Haldar, Simeon Adebola, Simon Guist, Soroush Nasiriany, Stefan Schaal, Stefan Welker, Stephen Tian, Sudeep Dasari, Suneel Belkhale, Takayuki Osa, Tatsuya Harada, Tatsuya Matsushima, Ted Xiao, Tianhe Yu, Tianli Ding, Todor Davchev, Tony. Zhao, Travis Armstrong, Trevor Darrell, Vidhi Jain, Vincent Vanhoucke, Wei Zhan, Wenxuan Zhou, Wolfram Burgard, Xi Chen, Xiaolong Wang, Xinghao Zhu, Xuanlin Li, Yao Lu, Yevgen Chebotar, Yifan Zhou, Yifeng Zhu, Ying Xu, Yixuan Wang, Yonatan Bisk, Yoonyoung Cho, Youngwoon Lee, Yuchen Cui, Yueh-hua Wu, Yujin Tang, Yuke Zhu, Yunzhu Li, Yusuke Iwasawa, Yutaka Matsuo, Zhuo Xu and Zichen Cui · 2023
Later among the works it cites.
“Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark”
Alexander Pan, Jun Chan, Andy Zou, Nathaniel Li, Steven Basart, Thomas Woodside, Hanlin Zhang, Scott Emmons and Dan Hendrycks · 2023
Later among the works it cites.
“Generative agents: Interactive simulacra of human behavior”
Joon Park, Joseph. O’Brien, Carrie. Cai, Meredith Morris, Percy Liang and Michael. Bernstein · 2023
Later among the works it cites.
“AI Deception: A Survey of Examples, Risks, and Potential Solutions”
Peter. Park, Simon Goldstein, Aidan O’Gara, Michael Chen and Dan Hendrycks · 2023
Later among the works it cites.
“Gorilla: Large Language Model Connected with Massive APIs”
Shishir. Patil, Tianjun Zhang, Xin Wang and Joseph. Gonzalez · 2023
Later among the works it cites.
“ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs”
Yujia Qin, Shi Liang, Yining Ye, Kunlun Zhu, Lan Yan, Ya-Ting Lu, Yankai Lin, Xin Cong, Xiangru Tang, Bill Qian, Sihan Zhao, Runchu Tian, Ruobing Xie, Jie Zhou, Marc. Gerstein, Dahai Li, Zhiyuan Liu and Maosong Sun · 2023
Later among the works it cites.
“From Plane Crashes to Algorithmic Harm: Applicability of Safety Engineering Frameworks for Responsible ML”
Shalaleh Rismani, Renee Shelby, Andrew Smart, Edgar Jatho, Joshua Kroll, AJung Moon and Negar Rostamzadeh · 2023
Later among the works it cites.
“Preventing Language Models From Hiding Their Reasoning”
Fabien Roger and Ryan Greenblatt · 2023
Later among the works it cites.
“Rise of the Newsbots: AI-Generated News Websites Proliferating Online”
McKenzie Sadeghi and Lorenzo Arvanitis · 2023
Later among the works it cites.
“MAESTRO: Open-ended environment design for multi-agent reinforcement learning”
Mikayel Samvelyan, Akbir Khan, Michael Dennis, Minqi Jiang, Jack Parker-Holder, Jakob Foerster, Roberta Raileanu and Tim Rocktäschel · 2023
Later among the works it cites.
“Cybersecurity for AI Systems: A Survey”
Raghvinder. Sangwan, Youakim Badr and Satish. Srinivasan · 2023
Later among the works it cites.
“Toolformer: Language Models Can Teach Themselves to Use Tools”
Timo Schick, Jane Dwivedi-Yu, Roberto Dessi, Roberta Raileanu, Maria Lomeli, Eric Hambro, Luke Zettlemoyer, Nicola Cancedda and Thomas Scialom · 2023
Later among the works it cites.
“Multi-Agent Security Workshop at NeurIPS 2023”
Christian Schroeder, Hawra Milani, Klaudia Krawiecka, Swapneel Mehta, Carla Cremer and Martin Strohmeier · 2023
Later among the works it cites.
“Perfectly Secure Steganography Using Minimum Entropy Coupling”
Christian Schroeder, Samuel Sokota, J. Kolter, Jakob Foerster and Martin Strohmeier · 2023
Later among the works it cites.
“FIND: A Function Description Benchmark for Evaluating Interpretability Methods”
Sarah Schwettmann, Tamar Shaham, Joanna Materzynska, Neil Chowdhury, Shuang Li, Jacob Andreas, David Bau and Antonio Torralba · 2023
Later among the works it cites.
“Unravelling the Attack Surface of AI Systems”, 2023
SecureWorks · 2023
Later among the works it cites.
Elizabeth Seger, Noemi Dreksler, Richard Moulange, Emily Dardaman, Jonas Schuett, K. Wei, Christoph Winter, Mackenzie Arnold, SeánÓ hÉigeartaigh, Anton Korinek, Markus Anderljung, Ben Bucknall, Alan Chan, Eoghan Stafford, Leonie Koessler, Aviv Ovadya, Ben Garfinkel, Emma Bluemke, Michael Aird, Patrick Levermore, Julian Hazell and Abhishek Gupta · 2023
Later among the works it cites.
“Democratising AI: Multiple Meanings, Goals, and Methods”
Elizabeth Seger, Aviv Ovadya, Divya Siddarth, Ben Garfinkel and Allan Dafoe · 2023
Later among the works it cites.
“Personality Traits in Large Language Models”
Greg Serapio-García, Mustafa Safdari, Clément Crepy, Luning Sun, Stephen Fitz, Peter Romero, Marwa Abdulhai, Aleksandra Faust and Maja Matarić · 2023
Later among the works it cites.
“Pushing the Limits of Fairness in Algorithmic Decision-Making”
Nisarg Shah · 2023
Later among the works it cites.
Yonadav Shavit · 2023
Later among the works it cites.
“Learning to Manipulate a Financial Benchmark”
Megan. Shearer, Gabriel Rauterberg and Michael. Wellman · 2023
Later among the works it cites.
“Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction”
Renee Shelby, Shalaleh Rismani, Kathryn Henne, AJung Moon, Negar Rostamzadeh, Paul Nicholas, N’Mah Yilla-Akbari, Jess Gallegos, Andrew Smart, Emilio Garcia and Gurleen Virk · 2023
Later among the works it cites.
“Model evaluation for extreme risks”
Toby Shevlane, Sebastian Farquhar, Ben Garfinkel, Mary Phuong, Jess Whittlestone, Jade Leung, Daniel Kokotajlo, Nahema Marchal, Markus Anderljung, Noam Kolt, Lewis Ho, Divya Siddarth, Shahar Avin, Will Hawkins, Been Kim, Iason Gabriel, Vijay Bolina, Jack Clark, Yoshua Bengio, Paul Christiano and Allan Dafoe · 2023
Later among the works it cites.
“Opportunities and Risks of LLMs for Scalable Deliberation with Polis”
Christopher. Small, Ivan Vendrov, Esin Durmus, Hadjar Homaei, Elizabeth Barry, Julien Cornebise, Ted Suzman, Deep Ganguli and Colin Megill · 2023
Later among the works it cites.
“Leading the Pack: N-player Opponent Shaping”
Alexandra Souly, Timon Willi, Akbir Khan, Robert Kirk, Chris Lu, Edward Grefenstette and Tim Rocktäschel · 2023
Later among the works it cites.
“Reinforcement Learning for Quantitative Trading”
Shuo Sun, Rundong Wang and Bo An · 2023
Later among the works it cites.
“Cooperative AI via Decentralized Commitment Devices”
Xinyuan Sun, Davide Crapis, Matt Stephenson, Barnabé Monnot, Thomas Thiery and Jonathan Passerat-Palmbach · 2023
Later among the works it cites.
“Challenging the appearance of machine intelligence: Cognitive bias in LLMs and Best Practices for Adoption”, 2023
Alaina. Talboy and Elizabeth Fuller · 2023
Later among the works it cites.
“Evil Geniuses: Delving into the Safety of LLM-based Agents”
Yu Tian, Xiao Yang, Jingyuan Zhang, Yinpeng Dong and Hang Su · 2023
Later among the works it cites.
“International Governance of Civilian AI: A Jurisdictional Certification Approach”
Robert Trager, Ben Harack, Anka Reuel, Allison Carnegie, Lennart Heim, Lewis Ho, Sarah Kreps, Ranjit Lall, Owen Larter, SeánÓ hÉigeartaigh, Simon Staffell and José Villalobos · 2023
Later among the works it cites.
“Block Nuclear Launch by Autonomous Artificial Intelligence Act”, H.R.2894, 118th Congress, 2023
U.S. Congress · 2023
Later among the works it cites.
“Privacy-Preserving Techniques in AI-Powered Cyber Security: Challenges and Opportunities”
Vinod Vegesna · 2023
Later among the works it cites.
Alexander Vezhnevets, John. Agapiou, Avia Aharon, Ron Ziv, Jayd Matyas, Edgar. Duéñez-Guzmán, William. Cunningham, Simon Osindero, Danny Karmon and Joel. Leibo · 2023
Later among the works it cites.
“A learning agent that acquires social norms from public sanctions in decentralized multi-agent settings”
Eugene Vinitsky, Raphael Köster, John. Agapiou, Edgar. Duéñez-Guzmán, Alexander. Vezhnevets and Joel. Leibo · 2023
Later among the works it cites.
“Chaos persists in large-scale multi-agent learning despite adaptive learning rates”
Emmanouil-Vasileios Vlatakis-Gkaragkounis, Lampros Flokas and Georgios Piliouras · 2023
Later among the works it cites.
“Deep Contract Design via Discontinuous Networks”
Tonghan Wang, Paul Duetting, Dmitry Ivanov, Inbal Talgam-Cohen and David. Parkes · 2023
Later among the works it cites.
“Honesty Is the Best Policy: Defining and Mitigating AI Deception”
Francis Ward, Francesca Toni, Francesco Belardinelli and Tom Everitt · 2023
Later among the works it cites.
“Jailbroken: How Does LLM Safety Training Fail?”
Alexander Wei, Nika Haghtalab and Jacob Steinhardt · 2023
Later among the works it cites.
“Using the Veil of Ignorance to align AI systems with principles of justice”
Laura Weidinger, Kevin. McKee, Richard Everett, Saffron Huang, Tina. Zhu, Martin. Chadwick, Christopher Summerfield and Iason Gabriel · 2023
Later among the works it cites.
“Sociotechnical Safety Evaluation of Generative AI Systems”
Laura Weidinger, Maribeth Rauh, Nahema Marchal, Arianna Manzini, Lisa Hendricks, Juan Mateos-Garcia, Stevie Bergman, Jackie Kay, Conor Griffin, Ben Bariach, Iason Gabriel, Verena Rieser and William Isaac · 2023
Later among the works it cites.
“Synthetic lies: Understanding ai-generated misinformation and evaluating algorithmic and human solutions”
Jiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea. Parker and Munmun De · 2023
Later among the works it cites.
“Universal and Transferable Adversarial Attacks on Aligned Language Models”, 2023
Andy Zou, Zifan Wang, Nicholas Carlini, Milad Nasr, J. Kolter and Matt Fredrikson · 2023
Later among the works it cites.
“A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms”
Gavin Abercrombie, Djalel Benbouzid, Paolo Giudici, Delaram Golpayegani, Julio Hernandez, Pierre Noro, Harshvardhan Pandit, Eva Paraschou, Charlie Pownall, Jyoti Prajapati, Mark. Sayre, Ushnish Sengupta, Arthit Suriyawongkul, Ruby Thelot, Sofia Vei and Laura Waltersdorfer · 2024
Later among the works it cites.
“LOQA: Learning with Opponent Q-Learning Awareness”
Milad Aghajohari, Juan Duque, Tim Cooijmans and Aaron Courville · 2024
Later among the works it cites.
“Defending Against Social Engineering Attacks in the Age of LLMs”
Lin Ai, Tharindu Kumarage, Amrita Bhattacharjee, Zizhou Liu, Zheng Hui, Michael Davinroy, James Cook, Laura Cassani, Kirill Trapeznikov, Matthias Kirchner, Arslan Basharat, Anthony Hoogs, Joshua Garland, Huan Liu and Julia Hirschberg · 2024
Later among the works it cites.
“Inspect AI”, 2024
AISI · 2024
Later among the works it cites.
“Emergence in Multi-Agent Systems: A Safety Perspective”
Philipp Altmann, Julian Schönberger, Steffen Illium, Maximilian Zorn, Fabian Ritz, Tom Haider, Simon Burton and Thomas Gabor · 2024
Later among the works it cites.
“Developing a computer use model”, 2024
Anthropic · 2024
Later among the works it cites.
“Introducing the Model Context Protocol”, 2024
Anthropic · 2024
Later among the works it cites.
“Foundational Challenges in Assuring Alignment and Safety of Large Language Models”
Usman Anwar, Abulhair Saparov, Javier Rando, Daniel Paleka, Miles Turpin, Peter Hase, Ekdeep Lubana, Erik Jenner, Stephen Casper, Oliver Sourbut, Benjamin. Edelman, Zhaowei Zhang, Mario Günther, Anton Korinek, Jose Hernandez-Orallo, Lewis Hammond, Eric Bigelow, Alexander Pan, Lauro Langosco, Tomasz Korbak, Heidi Zhang, Ruiqi Zhong, SeánÓ. hÉigeartaigh, Gabriel Recchia, Giulio Corsi, Alan Chan, Markus Anderljung, Lilian Edwards, Yoshua Bengio, Danqi Chen, Samuel Albanie, Tegan Maharaj, Jakob Foerster, Florian Tramer, He He, Atoosa Kasirzadeh, Yejin Choi and David Krueger · 2024
Later among the works it cites.
“Safeguarded AI”, 2024
ARIA · 2024
Later among the works it cites.
“The Law of AI is the Law of Risky Agents without Intentions”
Ian Ayres and Jack. Balkin · 2024
Later among the works it cites.
“Markov Persuasion Processes: Learning to Persuade from Scratch”
Francesco Bacchiocchi, Francesco Stradi, Matteo Castiglioni, Alberto Marchesi and Nicola Gatti · 2024
Later among the works it cites.
“Towards evaluations-based safety cases for AI scheming”
Mikita Balesni, Marius Hobbhahn, David Lindner, Alexander Meinke, Tomek Korbak, Joshua Clymer, Buck Shlegeris, Jérémy Scheurer, Charlotte Stix, Rusheb Shah, Nicholas Goldowsky-Dill, Dan Braun, Bilal Chughtai, Owain Evans, Daniel Kokotajlo and Lucius Bushnaq · 2024
Later among the works it cites.
“Collective Cooperative Intelligence” Forthcoming
Wolfram Barfuss, Jessica. Flack, Chaitanya. Gokhale, Lewis Hammond, Christian Hilbe, Edward Hughes, Joel. Leibo, Tom Lenaerts, Simon. Levin, Udari Madhushani, Alex McAvoy, Janusz. Meylahn and Fernando. Santos · 2024
Later among the works it cites.
“API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs”
Kinjal Basu, Ibrahim Abdelaziz, Subhajit Chaudhury, Soham Dan, Maxwell Crouse, Asim Munawar, Sadhana Kumaravel, Vinod Muthusamy, Pavan Kapanipathi and Luis. Lastras · 2024
Later among the works it cites.
“Managing extreme AI risks amid rapid progress”
Yoshua Bengio, Geoffrey Hinton, Andrew Yao, Dawn Song, Pieter Abbeel, Trevor Darrell, Yuval Harari, Ya-Qin Zhang, Lan Xue, Shai Shalev-Shwartz, Gillian Hadfield, Jeff Clune, Tegan Maharaj, Frank Hutter, Atılımüneş Baydin, Sheila McIlraith, Qiqi Gao, Ashwin Acharya, David Krueger, Anca Dragan, Philip Torr, Stuart Russell, Daniel Kahneman, Jan Brauner and Sören Mindermann · 2024
Later among the works it cites.
“Societal Adaptation to Advanced AI”
Jamie Bernardi, Gabriel Mukobi, Hilary Greaves, Lennart Heim and Markus Anderljung · 2024
Later among the works it cites.
“Refining Minimax Regret for Unsupervised Environment Design”
Michael Beukman, Samuel Coward, Michael Matthews, Mattie Fellows, Minqi Jiang, Michael Dennis and Jakob Foerster · 2024
Later among the works it cites.
“Looking Inward: Language Models Can Learn About Themselves by Introspection”
Felix. Binder, James Chua, Tomek Korbak, Henry Sleight, John Hughes, Robert Long, Ethan Perez, Miles Turpin and Owain Evans · 2024
Later among the works it cites.
“Strategic competition in the age of AI: Emerging risks and opportunities from military use of artificial intelligence”
James Black, Mattias Eken, Jacob Parakilas, Stuart Dee, Conlan Ellis, Kiran Suman-Chauhan, Ryan. Bain, Harper Fine, Maria Aquilino, Melusine Lebret and Ondrej Palicka · 2024
Later among the works it cites.
“Moral AI”
Jana Borg, Walter Sinnott-Armstrong and Vincent Conitzer · 2024
Later among the works it cites.
“SafeBench”, 2024
CAIS · 2024
Later among the works it cites.
“Leveraging Artificial Intelligence to Bolster the Energy Sector in Smart Cities: A Literature Review”
Joséús Camacho, Bernabé Aguirre, Pedro Ponce, Brian Anthony and Arturo Molina · 2024
Later among the works it cites.
Gian Campedelli, Nicolò Penzo, Massimo Stefan, Roberto Dessì, Marco Guerini, Bruno Lepri and Jacopo Staiano · 2024
Later among the works it cites.
“Visibility into AI Agents”
Alan Chan, Carson Ezell, Max Kaufmann, Kevin Wei, Lewis Hammond, Herbie Bradley, Emma Bluemke, Nitarshan Rajkumar, David Krueger, Noam Kolt, Lennart Heim and Markus Anderljung · 2024
Later among the works it cites.
Alan Chan, Noam Kolt, Peter Wills, Usman Anwar, Christian de Witt, Nitarshan Rajkumar, Lewis Hammond, David Krueger, Lennart Heim and Markus Anderljung · 2024
Later among the works it cites.
“Imperfect Recall and AI Delegation”, 2024
Eric Chen, Alexis Ghersengorin and Sami Petersen · 2024
Later among the works it cites.
“On Catastrophic Inheritance of Large Foundation Models”
Hao Chen, Bhiksha Raj, Xing Xie and Jindong Wang · 2024
Later among the works it cites.
“LLMArena: Assessing Capabilities of Large Language Models in Dynamic Multi-Agent Environments”
Junzhe Chen, Xuming Hu, Shuodi Liu, Shiyu Huang, Wei-Wei Tu, Zhaofeng He and Lijie Wen · 2024
Later among the works it cites.
“AgentVerse: Facilitating Multi-Agent Collaboration and Exploring Emergent Behaviors”
Weize Chen, Yusheng Su, Jingwei Zuo, Cheng Yang, Chenfei Yuan, Chi-Min Chan, Heyang Yu, Yaxi Lu, Yi-Hsin Hung, Chen Qian, Yujia Qin, Xin Cong, Ruobing Xie, Zhiyuan Liu, Maosong Sun and Jie Zhou · 2024
Later among the works it cites.
“Position: Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback”
Vincent Conitzer, Rachel Freedman, Jobst Heitzig, Wesley. Holliday, Bob. Jacobs, Nathan Lambert, Milan Mossé, Eric Pacuit, Stuart Russell, Hailey Schoelkopf, Emanuel Tewolde and William. Zwicker · 2024
Later among the works it cites.
“Can Democracy Survive the Disruptive Power of AI?”, 2024
Raluca Csernatoni · 2024
Later among the works it cites.
“Research Agenda for Sociotechnical Approaches to AI Safety”, 2024
Samuel Curtis, Ravi Iyer, Cameron Kirk-Giannini, Victoria Krakovna, David Krueger, Nathan Lambert, Bruno Marnette, Colleen McKenzie, Julian Michael, Evan Miyazono, Noyuri Mima, Aviv Ovadya, Luke Thorburn and Deger Turan · 2024
Later among the works it cites.
“Detecting Emergent Behavior in Complex Systems: A Machine Learning Approach”
Simranjeet Dahia and Claudia Szabo · 2024
Later among the works it cites.
“Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems”
David”davidad” Dalrymple, Joar Skalse, Yoshua Bengio, Stuart Russell, Max Tegmark, Sanjit Seshia, Steve Omohundro, Christian Szegedy, Ben Goldhaber, Nora Ammann, Alessandro Abate, Joe Halpern, Clark Barrett, Ding Zhao, Tan Zhi-Xuan, Jeannette Wing and Joshua Tenenbaum · 2024
Later among the works it cites.
“Safe Pareto Improvements for Expected Utility Maximizers in Program Games”
Anthony DiGiovanni, Jesse Clifton and Nicolas Macé · 2024
Later among the works it cites.
“h4rm3l: A Dynamic Benchmark of Composable Jailbreak Attacks for LLM Safety Assessment”, 2024
Moussa Doumbouya, Ananjan Nandi, Gabriel Poesia, Davide Ghilardi, Anna Goldie, Federico Bianchi, Dan Jurafsky and Christopher. Manning · 2024
Later among the works it cites.
“Unelicitable Backdoors via Cryptographic Transformer Circuits”
Andis Draguns, Andrew Gritsevskiy, Sumeet Motwani and Christian de Witt · 2024
Later among the works it cites.
“GTBench: Uncovering the Strategic Reasoning Capabilities of LLMs via Game-Theoretic Evaluations”
Jinhao Duan, Renming Zhang, James Diffenderfer, Bhavya Kailkhura, Lichao Sun, Elias Stengel-Eskin, Mohit Bansal, Tianlong Chen and Kaidi Xu · 2024
Later among the works it cites.
“REGULATION (EU) 2024/1689 OF THE EUROPEAN PARLIAMENT AND OF THE COUNCIL of 13 June 2024: laying down harmonised rules on artificial intelligence and amending Regulations (EC) No 300/2008, (EU) No 167/2013, (EU) No 168/2013, (EU) 2018/858, (EU) 2018/1139 and (EU) 2019/2144 and Directives 2014/90/EU, (EU) 2016/797 and (EU) 2020/1828 (Artificial Intelligence Act)” Text with EEA relevance
EU · 2024
Later among the works it cites.
“A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios”
Xiachong Feng, Longxu Dou, Ella Li, Qinghao Wang, Haochuan Wang, Yu Guo, Chang Ma and Lingpeng Kong · 2024
Later among the works it cites.
“On the Feasibility of Fully AI-automated Vishing Attacks”
João Figueiredo, Afonso Carvalho, Daniel Castro, Daniel Gonçalves and Nuno Santos · 2024
Later among the works it cites.
“Algorithmic Collusion by Large Language Models”
Sara Fish, Yannai. Gonczarowski and Ran. Shorrer · 2024
Later among the works it cites.
“The Ethics of Advanced AI Assistants”
Iason Gabriel, Arianna Manzini, Geoff Keeling, Lisa Hendricks, Verena Rieser, Hasan Iqbal, Nenad Tomašev, Ira Ktena, Zachary Kenton, Mikel Rodriguez, Seliem El-Sayed, Sasha Brown, Canfer Akbulut, Andrew Trask, Edward Hughes, A. Bergman, Renee Shelby, Nahema Marchal, Conor Griffin, Juan Mateos-Garcia, Laura Weidinger, Winnie Street, Benjamin Lange, Alex Ingerman, Alison Lentz, Reed Enger, Andrew Barakat, Victoria Krakovna, John Siy, Zeb Kurth-Nelson, Amanda McCroskery, Vijay Bolina, Harry Law, Murray Shanahan, Lize Alberts, Borja Balle, Sarah de Haas, Yetunde Ibitoye, Allan Dafoe, Beth Goldberg, Sébastien Krier, Alexander Reese, Sims Witherspoon, Will Hawkins, Maribeth Rauh, Don Wallace, Matija Franklin, Josh. Goldstein, Joel Lehman, Michael Klenk, Shannon Vallor, Courtney Biles, Meredith Morris, Helen King, Blaiseüera Arcas, William Isaac and James Manyika · 2024
Later among the works it cites.
“Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data”
Matthias Gerstgrasser, Rylan Schaeffer, Apratim Dey, Rafael Rafailov, Tomasz Korbak, Henry Sleight, Rajashree Agrawal, John Hughes, Dhruv Pai, Andrey Gromov, Dan Roberts, Diyi Yang, David. Donoho and Sanmi Koyejo · 2024
Later among the works it cites.
“Introducing Gemini 2.0: our new AI model for the agentic era”, 2024
Google DeepMind · 2024
Later among the works it cites.
“AI Multi-Agent Interoperability Extension for Managing Multiparty Conversations”, 2024
Diego Gosmar, Deborah. Dahl, Emmett Coin and David Attwater · 2024
Later among the works it cites.
“Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols”
Charlie Griffin, Louis Thomson, Buck Shlegeris and Alessandro Abate · 2024
Later among the works it cites.
“Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast”
Xiangming Gu, Xiaosen Zheng, Tianyu Pang, Chao Du, Qian Liu, Ye Wang, Jing Jiang and Min Lin · 2024
Later among the works it cites.
“Communicating with Anecdotes”
Nika Haghtalab, Nicole Immorlica, Brendan Lucier, Markus Mobius and Divyarthi Mohan · 2024
Later among the works it cites.
“Covert Malicious Finetuning: Challenges in Safeguarding LLM Adaptation”
Danny Halawi, Alexander Wei, Eric Wallace, Tony Wang, Nika Haghtalab and Jacob Steinhardt · 2024
Later among the works it cites.
“Evolutionary mechanisms that promote cooperation may not promote social welfare”
The Han, Manh Duong and Matjaz Perc · 2024
Later among the works it cites.
“More than Marketing? On the Information Value of AI Benchmarks for Practitioners”
Amelia Hardy, Anka Reuel, Kiana Meimandi, Lisa Soder, Allie Griffith, Dylan. Asmar, Sanmi Koyejo, Michael. Bernstein and Mykel. Kochenderfer · 2024
Later among the works it cites.
“Distributed Threat Intelligence at the Edge Devices: A Large Language Model-Driven Approach”
Syed Hasan, Alaa. Alotaibi, Sajedul Talukder and Abdur. Shahid · 2024
Later among the works it cites.
Yifeng He, Ethan Wang, Yuyang Rong, Zifei Cheng and Hao Chen · 2024
Later among the works it cites.
“Multi-Sender Persuasion: A Computational Perspective”
Safwan Hossain, Tonghan Wang, Tao Lin, Yiling Chen, David. Parkes and Haifeng Xu · 2024
Later among the works it cites.
“Automated Design of Agentic Systems”
Shengran Hu, Cong Lu and Jeff Clune · 2024
Later among the works it cites.
“On the Resilience of Multi-Agent Systems with Malicious Agents”
Jen-tse Huang, Jiaxu Zhou, Tailin Jin, Xuhui Zhou, Zixi Chen, Wenxuan Wang, Youliang Yuan, Maarten Sap and Michael. Lyu · 2024
Later among the works it cites.
“Open-Endedness is Essential for Artificial Superhuman Intelligence”
Edward Hughes, Michael Dennis, Jack Parker-Holder, Feryal Behbahani, Aditi Mavalankar, Yuge Shi, Tom Schaul and Tim Rocktaschel · 2024
Later among the works it cites.
“Adversaries Can Misuse Combinations of Safe Models”
Erik Jones, Anca Dragan and Jacob Steinhardt · 2024
Later among the works it cites.
“Flooding Spread of Manipulated Knowledge in LLM-Based Multi-Agent Communities”
Tianjie Ju, Yiting Wang, Xinbei Ma, Pengzhou Cheng, Haodong Zhao, Yulong Wang, Lifeng Liu, Jian Xie, Zhuosheng Zhang and Gongshen Liu · 2024
Later among the works it cites.
Sayash Kapoor, Benedikt Stroebl, Zachary. Siegel, Nitya Nadgir and Arvind Narayanan · 2024
Later among the works it cites.
“Plurality of value pluralism and AI value alignment”
Atoosa Kasirzadeh · 2024
Later among the works it cites.
“Two Types of AI Existential Risk: Decisive and Accumulative”
Atoosa Kasirzadeh · 2024
Later among the works it cites.
“Epistemic Injustice in Generative AI”
Jackie Kay, Atoosa Kasirzadeh and Shakir Mohamed · 2024
Later among the works it cites.
“Governing AI Agents”
Noam Kolt · 2024
Later among the works it cites.
“Recursive Joint Simulation in Games”
Vojtěch Kovařík, Caspar Oesterheld and Vincent Conitzer · 2024
Later among the works it cites.
Max Lamparth, Anthony Corso, Jacob Ganz, Oriana Mastro, Jacquelyn Schneider and Harold Trinkunas · 2024
Later among the works it cites.
“AI AI Bias: Large Language Models Favor Their Own Generated Content”, 2024
Walter Laurito, Benjamin Davis, Peli Grietzer, Tomáš Gavenčiak, Ada Böhm and Jan Kulveit · 2024
Later among the works it cites.
“Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems”
Donghyun Lee and Mo Tiwari · 2024
Later among the works it cites.
“A theory of appropriateness with applications to generative artificial intelligence”
Joel. Leibo, Alexander Vezhnevets, Manfred Diaz, John. Agapiou, William. Cunningham, Peter Sunehag, Julia Haas, Raphael Koster, Edgar. Duéñez-Guzmán, William. Isaac, Georgios Piliouras, Stanley. Bileschi, Iyad Rahwan and Simon Osindero · 2024
Later among the works it cites.
“Information Design with Unknown Prior”
Tao Lin and Ce Li · 2024
Later among the works it cites.
“LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores”
Yiqi Liu, Nafise Moosavi and Chenghua Lin · 2024
Later among the works it cites.
“The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery”
Chris Lu, Cong Lu, Robert Lange, Jakob Foerster, Jeff Clune and David Ha · 2024
Later among the works it cites.
“LLMs and generative agent-based models for complex systems research”
Yikang Lu, Alberto Aleta, Chunpeng Du, Lei Shi and Yamir Moreno · 2024
Later among the works it cites.
“From Intention To Implementation: Automating Biomedical Research via LLMs”
Yi Luo, Linghang Shi, Yihao Li, Aobo Zhuang, Yeyun Gong, Ling Liu and Chen Lin · 2024
Later among the works it cites.
“Measuring Goal-Directedness”
Matt MacDermott, James Fox, Francesco Belardinelli and Tom Everitt · 2024
Later among the works it cites.
“AI Warfare Is Already Here”
Katrina Manson · 2024
Later among the works it cites.
“A Scalable Communication Protocol for Networks of Large Language Models”
Samuele Marro, Emanuele La, Jesse Wright, Guohao Li, Nigel Shadbolt, Michael Wooldridge and Philip Torr · 2024
Later among the works it cites.
“Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs”
Yohan Mathew, Ollie Matthews, Robert McCarthy, Joan Velja, Christian de Witt, Dylan Cope and Nandi Schoots · 2024
Later among the works it cites.
“Roles and Responsibilities Framework for Artificial Intelligence in Critical Infrastructure” PDF available online, 2024
Alejandro. Mayorkas · 2024
Later among the works it cites.
“Multi-agent cooperation through learning-aware policy gradients”
Alexander Meulemans, Seijin Kobayashi, Johannes von Oswald, Nino Scherrer, Eric Elmoznino, Blake Richards, Guillaume Lajoie, Blaiseüera Arcas and João Sacramento · 2024
Later among the works it cites.
“How Red Lobster’s misguided endless shrimp promotion drove it into bankruptcy”
Nathaniel Meyersohn · 2024
Later among the works it cites.
“Introducing Azure AI Agent Service”, 2024
Microsoft · 2024
Later among the works it cites.
“Secret Collusion among AI Agents: Multi-Agent Deception via Steganography”
Sumeet Motwani, Mikhail Baranchuk, Martin Strohmeier, Vijay Bolina, Philip Torr, Lewis Hammond and Christian Schroeder · 2024
Later among the works it cites.
“FTC Announces Final Rule Imposing Civil Penalties for Fake Consumer Reviews and Testimonials”
Richard. Newman · 2024
Later among the works it cites.
“Diversifying Training Pool Predictability for Zero-shot Coordination: A Theory of Mind Approach”
Dung Nguyen, Hung Le, Kien Do, Sunil Gupta, Svetha Venkatesh and Truyen Tran · 2024
Later among the works it cites.
“Distortion Resilience for Goal-Oriented Semantic Communication”
Minh-Duong Nguyen, Quang Do, Zhaohui Yang, Won-Joo Hwang and Quoc-Viet Pham · 2024
Later among the works it cites.
“A dataset of questions on decision-theoretic reasoning in Newcomb-like problems”
Caspar Oesterheld, Emery Cooper, Miles Kodama, Linh Nguyen and Ethan Perez · 2024
Later among the works it cites.
“Similarity-based cooperative equilibrium”
Caspar Oesterheld, Johannes Treutlein, Roger. Grosse, Vincent Conitzer and Jakob Foerster · 2024
Later among the works it cites.
“Learning and Sustaining Shared Normative Systems via Bayesian Rule Induction in Markov Games”
Ninell Oldenburg and Tan Zhi-Xuan · 2024
Later among the works it cites.
“OpenAI o1 System Card”, 2024
OpenAI · 2024
Later among the works it cites.
“How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions”
Lorenzo Pacchiardi, Alex Chan, Sören Mindermann, Ilan Moscovitz, Alexa Pan, Yarin Gal, Owain Evans and Jan. Brauner · 2024
Later among the works it cites.
“Complex Dynamics in Autobidding Systems”
Renato Paes, Georgios Piliouras, Jon Schneider, Kelly Spendlove and Song Zuo · 2024
Later among the works it cites.
“LLM Evaluators Recognize and Favor Their Own Generations”
Arjun Panickssery, Samuel. Bowman and Shi Feng · 2024
Later among the works it cites.
“AI deception: A survey of examples, risks, and potential solutions”
Peter. Park, Simon Goldstein, Aidan O’Gara, Michael Chen and Dan Hendrycks · 2024
Later among the works it cites.
“GoEX: Perspectives and Designs Towards a Runtime for Autonomous LLM Applications”
Shishir. Patil, Tianjun Zhang, Vivian Fang, C. Noppapon, Roy Huang, Aaron Hao, Martin Casado, Joseph. Gonzalez, Raluca Popa and Ion Stoica · 2024
Later among the works it cites.
“Automated Red Teaming with GOAT: the Generative Offensive Agent Tester”
Maya Pavlova, Erik Brinkman, Krithika Iyer, Vitor Albiero, Joanna Bitton, Hailey Nguyen, Joe Li, Cristian Ferrer, Ivan Evtimov and Aaron Grattafiori · 2024
Later among the works it cites.
“Cultural evolution in populations of Large Language Models”
Jérémy Perez, Corentin Léger, Marcela Ovando-Tellez, Chris Foulon, Joan Dussauld, Pierre-Yves Oudeyer and Clément Moulin-Frier · 2024
Later among the works it cites.
“Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents”
Giorgio Piatti, Zhijing Jin, Max Kleiman-Weiner, Bernhard Schölkopf, Mrinmaya Sachan and Rada Mihalcea · 2024
Later among the works it cites.
“Biden, Xi agree that humans, not AI, should control nuclear arms”
Jarrett Renshaw and Trevor Hunnicutt · 2024
Later among the works it cites.
“Open Problems in Technical AI Governance”
Anka Reuel, Ben Bucknall, Stephen Casper, Tim Fist, Lisa Soder, Onni Aarne, Lewis Hammond, Lujain Ibrahim, Alan Chan, Peter Wills, Markus Anderljung, Ben Garfinkel, Lennart Heim, Andrew Trask, Gabriel Mukobi, Rylan Schaeffer, Mauricio Baker, Sara Hooker, Irene Solaiman, Alexandra Luccioni, Nitarshan Rajkumar, Nicolas Moës, Jeffrey Ladish, Neel Guha, Jessica Newman, Yoshua Bengio, Tobin South, Alex Pentland, Sanmi Koyejo, Mykel. Kochenderfer and Robert Trager · 2024
Later among the works it cites.
“BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices”
Anka Reuel, Amelia Hardy, Chandler Smith, Max Lamparth, Malcolm Hardy and Mykel Kochenderfer · 2024
Later among the works it cites.
“Generative AI Needs Adaptive Governance”
Anka Reuel and Trond Undheim · 2024
Later among the works it cites.
“Escalation Risks from Language Models in Military and Diplomatic Decision-Making”
Juan-Pablo Rivera, Gabriel Mukobi, Anka Reuel, Max Lamparth, Chandler Smith and Jacquelyn Schneider · 2024
Later among the works it cites.
“PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action”
Yijia Shao, Tianshi Li, Weiyan Shi, Yanchen Liu and Diyi Yang · 2024
Later among the works it cites.
“Towards Understanding Sycophancy in Language Models”
Mrinank Sharma, Meg Tong, Tomasz Korbak, David Duvenaud, Amanda Askell, Samuel. Bowman, Esin Durmus, Zac Hatfield-Dodds, Scott. Johnston, Shauna. Kravec, Timothy Maxwell, Sam McCandlish, Kamal Ndousse, Oliver Rausch, Nicholas Schiefer, Da Yan, Miranda Zhang and Ethan Perez · 2024
Later among the works it cites.
Aryan Shrivastava, Jessica Hullman and Max Lamparth · 2024
Later among the works it cites.
“AI models collapse when trained on recursively generated data”
Ilia Shumailov, Zakhar Shumaylov, Yiren Zhao, Nicolas Papernot, Ross Anderson and Yarin Gal · 2024
Later among the works it cites.
Zachary. Siegel, Sayash Kapoor, Nitya Nagdir, Benedikt Stroebl and Arvind Narayanan · 2024
Later among the works it cites.
“Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF”
Anand Siththaranjan, Cassidy Laidlaw and Dylan Hadfield-Menell · 2024
Later among the works it cites.
“Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters”
Charlie Snell, Jaehoon Lee, Kelvin Xu and Aviral Kumar · 2024
Later among the works it cites.
“Position: a roadmap to pluralistic alignment”
Taylor Sorensen, Jared Moore, Jillian Fisher, Mitchell Gordon, Niloofar Mireshghallah, Christopher Rytting, Andre Ye, Liwei Jiang, Ximing Lu, Nouha Dziri, Tim Althoff and Yejin Choi · 2024
Later among the works it cites.
“Cooperation and Control in Delegation Games”
Oliver Sourbut, Lewis Hammond and Harriet Wood · 2024
Later among the works it cites.
“Committing to the wrong artificial delegate in a collective-risk dilemma is better than directly committing mistakes”
Inês Terrucha, Elias Fernández, Pieter Simoens and Tom Lenaerts · 2024
Later among the works it cites.
“Cultural Evolution of Cooperation among LLM Agents”
Aron Vallinder and Edward Hughes · 2024
Later among the works it cites.
“A survey of agent-based modeling for cybersecurity”
Arnstein Vestad and Bian Yang · 2024
Later among the works it cites.
Zhen Wang, Ruiqi Song, Chen Shen, Shiya Yin, Zhao Song, Balaraju Battu, Lei Shi, Danyang Jia, Talal Rahwan and Shuyue Hu · 2024
Later among the works it cites.
“Trustworthy Distributed AI Systems: Robustness, Privacy, and Governance”
Wenqi Wei and Ling Liu · 2024
Later among the works it cites.
“Care for Chatbots” Forthcoming
Peter Wills · 2024
Later among the works it cites.
“Inference Attacks: A Taxonomy, Survey, and Promising Directions”
Feng Wu, Lei Cui, Shaowen Yao and Shui Yu · 2024
Later among the works it cites.
“AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversations”
Qingyun Wu, Gagan Bansal, Jieyu Zhang, Yiran Wu, Beibin Li, Erkang Zhu, Li Jiang, Xiaoyun Zhang, Shaokun Zhang, Jiale Liu, Ahmed Awadallah, Ryen White, Doug Burger and Chi Wang · 2024
Later among the works it cites.
“A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models”
Zihao Xu, Yi Liu, Gelei Deng, Yuekang Li and Stjepan Picek · 2024
Later among the works it cites.
“Computational Aspects of Bayesian Persuasion under Approximate Best Response”
Kunhe Yang and Hanrui Zhang · 2024
Later among the works it cites.
“4.2 Tbps of bad packets and a whole lot more: Cloudflare’s Q3 DDoS report”, 2024
Omer Yoachimik and Jorge Pacheco · 2024
Later among the works it cites.
“NetSafe: Exploring the Topological Safety of Multi-agent Networks”
Miao Yu, Shilong Wang, Guibin Zhang, Junyuan Mao, Chenlong Yin, Qijiong Liu, Qingsong Wen, Kun Wang and Yang Wang · 2024
Later among the works it cites.
“AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies”
Yi Zeng, Kevin Klyman, Andy Zhou, Yu Yang, Minzhou Pan, Ruoxi Jia, Dawn Song, Percy Liang and Bo Li · 2024
Later among the works it cites.
“Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification”, 2024
Boyang Zhang, Yicong Tan, Yun Shen, Ahmed Salem, Michael Backes, Savvas Zannettou and Yang Zhang · 2024
Later among the works it cites.
“LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models”
Yadong Zhang, Shaoguang Mao, Tao Ge, Xun Wang, Adrian de Wynter, Yan Xia, Wenshan Wu, Ting Song, Man Lan and Furu Wei · 2024
Later among the works it cites.
“Beyond Preferences in AI Alignment”
Tan Zhi-Xuan, Micah Carroll, Matija Franklin and Hal Ashton · 2024
Later among the works it cites.
“Emergence of cooperation in the one-shot Prisoner’s dilemma through Discriminatory and Samaritan AIs”
Filippo Zimmaro, Manuel Miranda, Joséía Fernández, Jesús. Morenoópez, Max Reddel, Valeria Widler, Alberto Antonioni and The Han · 2024
Later among the works it cites.
“Amplify AI Powered Equity ETF”, 2025
AmplifyETFs · 2025
Closest in time.
“Preventing Rogue Agents Improves Multi-Agent Collaboration”
Ohav Barbi, Ori Yoran and Mor Geva · 2025
Closest in time.
“Cooperative AI Research Grants”, 2025
CAIF · 2025
Closest in time.
“Infrastructure for AI Agents”
Alan Chan, Kevin Wei, Sihao Huang, Nitarshan Rajkumar, Elija Perrier, Seth Lazar, Gillian. Hadfield and Markus Anderljung · 2025
Closest in time.
“Characterising Simulation-Based Program Equilibria”
Emery Cooper, Caspar Oesterheld and Vincent Conitzer · 2025
Closest in time.
“Hypothetical Minds: Scaffolding Theory of Mind for Multi-Agent Tasks with Large Language Models”
Logan Cross, Violet Xiang, Agam Bhatia, Daniel.. Yamins and Nick Haber · 2025
Closest in time.
Jessica Dai, Paula Gradu, Inioluwa Raji and Benjamin Recht · 2025
Closest in time.
“DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning”
DeepSeek-AI, Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, Xiaokang Zhang, Xingkai Yu, Yu Wu, Z.. Wu, Zhibin Gou, Zhihong Shao, Zhuoshu Li, Ziyi Gao, Aixin Liu, Bing Xue, Bingxuan Wang, Bochao Wu, Bei Feng, Chengda Lu, Chenggang Zhao, Chengqi Deng, Chenyu Zhang, Chong Ruan, Damai Dai, Deli Chen, Dongjie Ji, Erhang Li, Fangyun Lin, Fucong Dai, Fuli Luo, Guangbo Hao, Guanting Chen, Guowei Li, H. Zhang, Han Bao, Hanwei Xu, Haocheng Wang, Honghui Ding, Huajian Xin, Huazuo Gao, Hui Qu, Hui Li, Jianzhong Guo, Jiashi Li, Jiawei Wang, Jingchang Chen, Jingyang Yuan, Junjie Qiu, Junlong Li, J.. Cai, Jiaqi Ni, Jian Liang, Jin Chen, Kai Dong, Kai Hu, Kaige Gao, Kang Guan, Kexin Huang, Kuai Yu, Lean Wang, Lecong Zhang, Liang Zhao, Litong Wang, Liyue Zhang, Lei Xu, Leyi Xia, Mingchuan Zhang, Minghua Zhang, Minghui Tang, Meng Li, Miaojun Wang, Mingming Li, Ning Tian, Panpan Huang, Peng Zhang, Qiancheng Wang, Qinyu Chen, Qiushi Du, Ruiqi Ge, Ruisong Zhang, Ruizhe Pan, Runji Wang, R.. Chen, R.. Jin, Ruyi Chen, Shanghao Lu, Shangyan Zhou, Shanhuang Chen, Shengfeng Ye, Shiyu Wang, Shuiping Yu, Shunfeng Zhou, Shuting Pan, S.. Li, Shuang Zhou, Shaoqing Wu, Shengfeng Ye, Tao Yun, Tian Pei, Tianyu Sun, T. Wang, Wangding Zeng, Wanjia Zhao, Wen Liu, Wenfeng Liang, Wenjun Gao, Wenqin Yu, Wentao Zhang, W.. Xiao, Wei An, Xiaodong Liu, Xiaohan Wang, Xiaokang Chen, Xiaotao Nie, Xin Cheng, Xin Liu, Xin Xie, Xingchao Liu, Xinyu Yang, Xinyuan Li, Xuecheng Su, Xuheng Lin, X.. Li, Xiangyue Jin, Xiaojin Shen, Xiaosha Chen, Xiaowen Sun, Xiaoxiang Wang, Xinnan Song, Xinyi Zhou, Xianzu Wang, Xinxia Shan, Y.. Li, Y.. Wang, Y.. Wei, Yang Zhang, Yanhong Xu, Yao Li, Yao Zhao, Yaofeng Sun, Yaohui Wang, Yi Yu, Yichao Zhang, Yifan Shi, Yiliang Xiong, Ying He, Yishi Piao, Yisong Wang, Yixuan Tan, Yiyang Ma, Yiyuan Liu, Yongqiang Guo, Yuan Ou, Yuduan Wang, Yue Gong, Yuheng Zou, Yujia He, Yunfan Xiong, Yuxiang Luo, Yuxiang You, Yuxuan Liu, Yuyang Zhou, Y.. Zhu, Yanhong Xu, Yanping Huang, Yaohui Li, Yi Zheng, Yuchen Zhu, Yunxian Ma, Ying Tang, Yukun Zha, Yuting Yan, Z.. Ren, Zehui Ren, Zhangli Sha, Zhe Fu, Zhean Xu, Zhenda Xie, Zhengyan Zhang, Zhewen Hao, Zhicheng Ma, Zhigang Yan, Zhiyu Wu, Zihui Gu, Zijia Zhu, Zijun Liu, Zilin Li, Ziwei Xie, Ziyang Song, Zizheng Pan, Zhen Huang, Zhipeng Xu, Zhongyu Zhang and Zhen Zhang · 2025
Closest in time.
“Great Models Think Alike and this Undermines AI Oversight”
Shashwat Goel, Joschka Struber, Ilze Auzina, Karuna. Chandra, Ponnurangam Kumaraguru, Douwe Kiela, Ameya Prabhu, Matthias Bethge and Jonas Geiping · 2025
Closest in time.
“Platforms for Efficient and Incentive-Aware Collaboration”
Nika Haghtalab, Mingda Qiao and Kunhe Yang · 2025
Closest in time.
“Neural Interactive Proofs” Forthcoming
Lewis Hammond and Sam Adam-Day · 2025
Closest in time.
“Lessons from complexity theory for AI governance”
Noam Kolt, Michal Shur-Ofry and Reuven Cohen · 2025
Closest in time.
“Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development”
Jan Kulveit, Raymond Douglas, Nora Ammann, Deger Turan, David Krueger and David Duvenaud · 2025
Closest in time.
“Multi-Agent Credit Assignment with Pretrained Language Models”
Wenhao Li, Dan Qiao, Baoxiang Wang, Xiangfeng Wang, Wei, Hao Shen, Bo Jin and Hongyuan Zha · 2025
Closest in time.
“Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs”
Mantas Mazeika, Xuwang Yin, Rishub Tamirisa, Jaehyuk Lim, Bruce. Lee, Richard Ren, Long Phan, Norman Mu, Adam Khoja, Oliver Zhang and Dan Hendrycks · 2025
Closest in time.
“Building Toward a Smarter, More Personalized Assistant”, 2025
Meta · 2025
Closest in time.
“Fully Autonomous AI Agents Should Not be Developed”
Margaret Mitchell, Avijit Ghosh, Alexandra Luccioni and Giada Pistilli · 2025
Closest in time.
“Introducing Operator”, 2025
OpenAI · 2025
Closest in time.
“Financial History”, 2025
Option Alpha · 2025
Closest in time.
“AIP for Defense”, https://www.palantir.com/platforms/aip/defense/ , 2025
Palantir · 2025
Closest in time.
“Emergence of human-like polarization among large language model agents”
Jinghua Piao, Zhihong Lu, Chen Gao, Fengli Xu, Fernando. Santos, Yong Li and James Evans · 2025
Closest in time.
“How Cyberscammers Use AI to Manipulate Google Search Results”
Prashant Sharma · 2025
Closest in time.
“HAL: A Holistic Agent Leaderboard for Centralized and Reproducible Agent Evaluation”, https://github.com/princeton-pli/hal-harness/ , 2025
Benedikt Stroebl, Sayash Kapoor and Arvind Narayanan · 2025
Closest in time.
“Who Should Develop Which AI Evaluations?”, 2025
Lara Thurnherr, Robert Trager, Amin Oueslati, Christoph Winter, Cliodhna’i Ghuidhir, Joe O’Brien, Jun Chan, Lorenzo Pacchiardi, Anka Reuel, Merlin Stein, Oliver Guest, Oliver Sourbut, Renan Araujo, Seth Donoughe and Yi Zeng · 2025
Closest in time.
“Network Models: Triggering Marketing Network Effects With AI”
Tiffanie Turner-Henderson · 2025
Closest in time.
“A Taxonomy of Systemic Risks from General-Purpose AI”
Risto Uuk, Carlos Gutierrez, Lode Lauwaert, Carina Prunkl and Lucia Velasco · 2025
Closest in time.
“Learning to Negotiate via Voluntary Commitment”
Shuhui Zhu, Baoxiang Wang, Sriram Subramanian and Pascal Poupart · 2025
Closest in time.