Fetching the paper…
Reading the bibliography…
AI agents plan and execute interactions in open-ended environments.
The digital library project volume 1: the world of knowbots (DRAFT)
Robert E Kahn and Vinton G Cerf · 1988
Earlier work this paper cites.
Provision of Public Goods: Fully Implementing the Core through Private Contributions
Mark Bagnoli and Barton L. Lipman · 1989
Earlier work this paper cites.
Agents that reduce work and information overload
Pattie Maes · 1994
Earlier work this paper cites.
Rationalist Explanations for War
James D. Fearon · 1995
Earlier work this paper cites.
Artificial life meets entertainment: lifelike autonomous agents
Pattie Maes · 1995
Earlier work this paper cites.
21 CFR Part 820 - Quality System Regulation
Health and Human Services Department and Food and Drug Administration · 1996
Earlier work this paper cites.
Autonomous interface agents
Henry Lieberman · 1997
Earlier work this paper cites.
Humans and Automation: Use, Misuse, Disuse, Abuse
Raja Parasuraman and Victor Riley · 1997
Earlier work this paper cites.
A roadmap of agent research and development
Nicholas R Jennings, Katia Sycara, and Michael Wooldridge · 1998
Earlier work this paper cites.
The Private Provision of Public Goods via Dominant Assurance Contracts
Alexander Tabarrok · 1998
Earlier work this paper cites.
Assessing anonymous communication on the internet: Policy deliberations
Rob Kling, Ya-ching Lee, Al Teich, and Mark S Frankel · 1999
Earlier work this paper cites.
Does automation bias decision-making?
LINDA J. Skitka, KATHLEEN L. Mosier, and MARK Burdick · 1999
Earlier work this paper cites.
Unmasking Jane and John Doe: Online Anonymity and the First Amendment
Victoria Smith Ekstrand · 2003
Earlier work this paper cites.
War as a Commitment Problem
Robert Powell · 2006
Earlier work this paper cites.
RFC 4271: A border gateway protocol 4 (BGP-4), 2006
Yakov Rekhter, Tony Li, and Susan Hares · 2006
Earlier work this paper cites.
The Generative Internet
Jonathan L. Zittrain · 2006
Earlier work this paper cites.
Specifying protocols for multi-agent systems interaction
Stefan Poslad · 2007
Earlier work this paper cites.
Adverse selection in online "trust" certifications
Benjamin Edelman · 2009
Earlier work this paper cites.
An Introduction to Multiagent Systems
Michael Wooldridge · 2009
Earlier work this paper cites.
Verification and validation in scientific computing
William L Oberkampf and Christopher J Roy · 2010
Earlier work this paper cites.
Software Agents, Anticipatory Ethics, and Accountability
Deborah G. Johnson · 2011
Earlier work this paper cites.
The Anonymous Internet
Bryan H. Choi · 2012
Earlier work this paper cites.
Automation bias: a systematic review of frequency, effect mediators, and mitigators
Kate Goddard, Abdul Roudsari, and Jeremy C Wyatt · 2012
Earlier work this paper cites.
Why do people seek anonymity on the internet? informing policy and design
Ruogu Kang, Stephanie Brown, and Sara Kiesler · 2013
Earlier work this paper cites.
Playing Atari with Deep Reinforcement Learning, December 2013
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Privacy Self-Regulation in Crisis? – TRUSTe’s ‘Deceptive’ Practices, December 2014
Chris Connolly, Graham Greenleaf, and Nigel Waters · 2014
Earlier work this paper cites.
Where in the world is the internet? Locating political power in internet infrastructure
Ashwin Jacob Mathew · 2014
Earlier work this paper cites.
Information Supplement: Guidance for PCI DSS Scoping and Network Segmentation
PCI Security Standards Council · 2016
Earlier work this paper cites.
Cooperative Inverse Reinforcement Learning
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2016
Earlier work this paper cites.
Security issues with certificate authorities
Jake A. Berkowsky and Thaier Hayajneh · 2017
Earlier work this paper cites.
Deep Reinforcement Learning from Human Preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Earlier work this paper cites.
Rogers Communications Inc. v. Voltage Pictures, September 2018
2018
Earlier work this paper cites.
Scalable agent alignment via reward modeling: a research direction, November 2018
Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini, and Shane Legg · 2018
Earlier work this paper cites.
OpenAI Charter, 2018
OpenAI · 2018
Earlier work this paper cites.
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis · 2018
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S. Sutton and Andrew G. Barto · 2018
Earlier work this paper cites.
The Role of Cooperation in Responsible AI Development
Amanda Askell, Miles Brundage, and Gillian Hadfield · 2019
Earlier work this paper cites.
Learning Existing Social Conventions via Observationally Augmented Self-Play
Adam Lerer and Alexander Peysakhovich · 2019
Earlier work this paper cites.
Digital Market Perfection
Rory Van Loo · 2019
Earlier work this paper cites.
AI loyalty: A New Paradigm for Aligning Stakeholder Interests, March 2020
Anthony Aguirre, Gaia Dempsey, Harry Surden, and Peter B. Reiner · 2020
Earlier work this paper cites.
Agent57: Outperforming the Atari Human Benchmark, March 2020
Adrià Puigdomènech Badia, Bilal Piot, Steven Kapturowski, Pablo Sprechmann, Alex Vitvitskyi, Daniel Guo, and Charles Blundell · 2020
Earlier work this paper cites.
An Object Detection based Solver for {Google’s} Image {reCAPTCHA} v2
Md Imran Hossen, Yazhou Tu, Md Fazle Rabby, Md Nazmul Islam, Hui Cao, and Xiali Hei · 2020
Earlier work this paper cites.
"Other-Play " for zero-shot coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster · 2020
Earlier work this paper cites.
Closing the AI accountability gap: defining an end-to-end framework for internal algorithmic auditing
Inioluwa Deborah Raji, Andrew Smart, Rebecca N. White, Margaret Mitchell, Timnit Gebru, Ben Hutchinson, Jamila Smith-Loud, Daniel Theron, and Parker Barnes · 2020
Earlier work this paper cites.
Cooperative AI: machines must learn to find common ground
Allan Dafoe, Yoram Bachrach, Gillian Hadfield, Eric Horvitz, Kate Larson, and Thore Graepel · 2021
Earlier work this paper cites.
A Low-Cost Attack against the hCaptcha System
Md Imran Hossen and Xiali Hei · 2021
Earlier work this paper cites.
Scalable Evaluation of Multi-Agent Reinforcement Learning with Melting Pot, July 2021
Joel Z. Leibo, Edgar Duéñez-Guzmán, Alexander Sasha Vezhnevets, John P. Agapiou, Peter Sunehag, Raphael Koster, Jayd Matyas, Charles Beattie, Igor Mordatch, and Thore Graepel · 2021
Earlier work this paper cites.
Artificial Intelligence: A Modern Approach
Stuart J. Russell and Peter Norvig · 2021
Earlier work this paper cites.
Constitutional AI: Harmlessness from AI Feedback, December 2022
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosuite, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemi Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan · 2022
Earlier work this paper cites.
LangChain 0.0.77 Docs, 2022
Harrison Chase · 2022
Earlier work this paper cites.
When would AGIs engage in conflict?, October 2022
Jesse Clifton, Samuel Martin, and Anthony DiGiovanni · 2022
Earlier work this paper cites.
Who Audits the Auditors? Recommendations from a field scan of the algorithmic auditing ecosystem
Sasha Costanza-Chock, Inioluwa Deborah Raji, and Joy Buolamwini · 2022
Earlier work this paper cites.
RFC 9293: Transmission control protocol (tcp), 2022
W Eddy · 2022
Earlier work this paper cites.
RFC 9110: HTTP semantics, 2022
R Fielding, M Nottingham, and J Reschke · 2022
Cited alongside, same era.
Where to Report a Cyber Incident, May 2022
Government of the United Kingdom · 2022
Cited alongside, same era.
The flaws of policies requiring human oversight of government algorithms
Ben Green · 2022
Cited alongside, same era.
Preparing for the (Non-Existent?) Future of Work, June 2022
Anton Korinek and Megan Juelfs · 2022
Cited alongside, same era.
Outsider Oversight: Designing a Third Party Audit Ecosystem for AI Governance
Inioluwa Deborah Raji, Peggy Xu, Colleen Honigsberg, and Daniel Ho · 2022
Cited alongside, same era.
15 U.S. Code § 1666 - Correction of billing errors, 2023
2023
Cited alongside, same era.
TrustAgent: Towards Safe and Trustworthy LLM-based Agents through Agent Constitution, August 2024
Wenyue Hua, Xianjun Yang, Mingyu Jin, Wei Cheng, Ruixiang Tang, and Yongfeng Zhang · 2024
Later among the works it cites.
Collective Constitutional AI: Aligning a Language Model with Public Input
Saffron Huang, Divya Siddarth, Liane Lovitt, Thomas I. Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli · 2024
Later among the works it cites.
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?, April 2024
Carlos E. Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan · 2024
Later among the works it cites.
Adversaries Can Misuse Combinations of Safe Models, June 2024
Erik Jones, Anca Dragan, and Jacob Steinhardt · 2024
Later among the works it cites.
Adding payments to your LLM agentic workflows, November 2024
Steve Kaliski · 2024
Later among the works it cites.
AI Agents That Matter, July 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
AI Safety Summit - Enhancing Frontier AI Safety, 2023
AWS · 2023
Cited alongside, same era.
Designing Fiduciary Artificial Intelligence, July 2023
Sebastian Benthall and David Shekman · 2023
Cited alongside, same era.
Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models, December 2023
Manish Bhatt, Sahana Chennabasappa, Cyrus Nikolaidis, Shengye Wan, Ivan Evtimov, Dominik Gabi, Daniel Song, Faizan Ahmad, Cornelius Aschermann, Lorenzo Fontana, Sasha Frolov, Ravi Prakash Giri, Dhaval Kapil, Yiannis Kozyrakis, David LeBlanc, James Milazzo, Aleksandar Straumann, Gabriel Synnaeve, Varun Vontimitta, Spencer Whitman, and Joshua Saxe · 2023
Cited alongside, same era.
C2PA Technical Specification, 2023
C2PA · 2023
Cited alongside, same era.
Harms from Increasingly Agentic Algorithmic Systems
Alan Chan, Rebecca Salganik, Alva Markelius, Chris Pang, Nitarshan Rajkumar, Dmitrii Krasheninnikov, Lauro Langosco, Zhonghao He, Yawen Duan, Micah Carroll, Michelle Lin, Alex Mayhew, Katherine Collins, Maryam Molamohammadi, John Burden, Wanru Zhao, Shalaleh Rismani, Konstantinos Voudouris, Umang Bhatt, Adrian Weller, David Krueger, and Tegan Maharaj · 2023
Cited alongside, same era.
Data, privacy, and security for Azure OpenAI Service - Azure AI services, June 2023
ChrisHMSFT, PatrickFarley, mrbullwinkle, eric urban, and aahill · 2023
Cited alongside, same era.
Sayash Kapoor, Benedikt Stroebl, Zachary S. Siegel, Nitya Nadgir, and Arvind Narayanan · 2024
Later among the works it cites.
The PRISM Alignment Project: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models, April 2024
Hannah Rose Kirk, Alexander Whitefield, Paul Röttger, Andrew Bean, Katerina Margatina, Juan Ciro, Rafael Mosquera, Max Bartolo, Adina Williams, He He, Bertie Vidgen, and Scott A. Hale · 2024
Later among the works it cites.
Governing AI Agents, April 2024
Noam Kolt · 2024
Later among the works it cites.
Scenarios for the Transition to AGI, March 2024
Anton Korinek and Donghyun Suh · 2024
Later among the works it cites.
Frontier AI Ethics: Anticipating and Evaluating the Societal Impacts of Generative Agents, April 2024
Seth Lazar · 2024
Later among the works it cites.
The Moral Case for Using Language Model Agents for Recommendation, October 2024
Seth Lazar, Luke Thorburn, Tian Jin, and Luca Belli · 2024
Later among the works it cites.
Questionable practices in machine learning, July 2024
Gavin Leech, Juan J. Vazquez, Misha Yagudin, Niclas Kupper, and Laurence Aitchison · 2024
Later among the works it cites.
ToolSandbox: A Stateful, Conversational, Interactive Evaluation Benchmark for LLM Tool Use Capabilities, August 2024
Jiarui Lu, Thomas Holleis, Yizhe Zhang, Bernhard Aumayer, Feng Nan, Felix Bai, Shuang Ma, Shen Ma, Mengyu Li, Guoli Yin, Zirui Wang, and Ruoming Pang · 2024
Later among the works it cites.
A Scalable Communication Protocol for Networks of Large Language Models, October 2024
Samuele Marro, Emanuele La Malfa, Jesse Wright, Guohao Li, Nigel Shadbolt, Michael Wooldridge, and Philip Torr · 2024
Later among the works it cites.
MultiOn AI, 2024
MultiOn · 2024
Later among the works it cites.
GPT Actions, 2024
OpenAI · 2024
Later among the works it cites.
How should I report a GPT?, 2024
OpenAI · 2024
Later among the works it cites.
Model Spec, May 2024
OpenAI · 2024
Later among the works it cites.
GoEX: Perspectives and Designs Towards a Runtime for Autonomous LLM Applications, April 2024
Shishir G. Patil, Tianjun Zhang, Vivian Fang, Noppapon C., Roy Huang, Aaron Hao, Martin Casado, Joseph E. Gonzalez, Raluca Ada Popa, and Ion Stoica · 2024
Later among the works it cites.
Language Models Can Reduce Asymmetry in Information Markets, March 2024
Nasim Rahaman, Martin Weiss, Manuel Wüthrich, Yoshua Bengio, Li Erran Li, Chris Pal, and Bernhard Schölkopf · 2024
Later among the works it cites.
Artificial Intelligence Incident Database, 2024
Responsible AI Collaborative · 2024
Later among the works it cites.
AI Rights for Human Safety, 2024
Peter Salib and Simon Goldstein · 2024
Later among the works it cites.
The SEC Whistleblower Program Is Dominating Regulatory Enforcement, October 2024
Bruce Schneier and Nathan Sanders · 2024
Later among the works it cites.
Whistleblower Program, August 2024
Securities and Exchange Commission · 2024
Later among the works it cites.
Auto-GPT-Plugins, 2024
Significant-Gravitas · 2024
Later among the works it cites.
Beyond Browsing: API-Based Web Agents, October 2024
Yueqi Song, Frank Xu, Shuyan Zhou, and Graham Neubig · 2024
Later among the works it cites.
Verifiable Credentials Data Model v2.0
Manu Sporny, Dave Longley, David Chadwick, and Orie Steele · 2024
Later among the works it cites.
eCommerce - Worldwide, 2024
Statista · 2024
Later among the works it cites.
Brave New World? Human Welfare and Paternalistic AI, 2024
Cass R. Sunstein · 2024
Later among the works it cites.
Tamper-Resistant Safeguards for Open-Weight LLMs, August 2024
Rishub Tamirisa, Bhrugu Bharathi, Long Phan, Andy Zhou, Alice Gatti, Tarun Suresh, Maxwell Lin, Justin Wang, Rowan Wang, Ron Arel, Andy Zou, Dawn Song, Bo Li, Dan Hendrycks, and Mantas Mazeika · 2024
Later among the works it cites.
Identifying Current Barriers in RPKI Adoption, September 2024
Cecilia Testart, Josephine Wolff, Deepak Gouda, and Romain Fontugne · 2024
Later among the works it cites.
Beyond Privacy Trade-offs with Structured Transparency, March 2024
Andrew Trask, Emma Bluemke, Teddy Collins, Ben Garfinkel Eric Drexler, Claudia Ghezzou Cuervas-Mons, Iason Gabriel, Allan Dafoe, and William Isaac · 2024
Later among the works it cites.
The Rise of AI Agent Infrastructure, June 2024
Jon Turow · 2024
Later among the works it cites.
North Carolina Musician Charged With Music Streaming Fraud Aided By Artificial Intelligence, September 2024
Southern District of New York U.S. Attorney’s Office · 2024
Later among the works it cites.
The instruction hierarchy: Training llms to prioritize privileged instructions
Eric Wallace, Kai Xiao, Reimar Leike, Lilian Weng, Johannes Heidecke, and Alex Beutel · 2024
Later among the works it cites.
Designing Incident Reporting Systems for Harms from AI
Kevin Wei and Lennart Heim · 2024
Later among the works it cites.
RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts, November 2024
Hjalmar Wijk, Tao Lin, Joel Becker, Sami Jawhar, Neev Parikh, Thomas Broadley, Lawrence Chan, Michael Chen, Josh Clymer, Jai Dhyani, Elena Ericheva, Katharyn Garcia, Brian Goodrich, Nikola Jurkovic, Megan Kinniment, Aron Lajko, Seraphina Nix, Lucas Sato, William Saunders, Maksym Taran, Ben West, and Elizabeth Barnes · 2024
Later among the works it cites.
Care for Chatbots, May 2024
Peter Wills · 2024
Later among the works it cites.
Introducing Devin, the first AI software engineer, March 2024
Scott Wu · 2024
Later among the works it cites.
OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments, April 2024
Tianbao Xie, Danyang Zhang, Jixuan Chen, Xiaochuan Li, Siheng Zhao, Ruisheng Cao, Toh Jing Hua, Zhoujun Cheng, Dongchan Shin, Fangyu Lei, Yitao Liu, Yiheng Xu, Shuyan Zhou, Silvio Savarese, Caiming Xiong, Victor Zhong, and Tao Yu · 2024
Later among the works it cites.
Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risk of Language Models, August 2024
Andy K. Zhang, Neil Perry, Riya Dulepet, Eliot Jones, Justin W. Lin, Joey Ji, Celeste Menders, Gashon Hussein, Samantha Liu, Donovan Jasper, Pura Peetathawatchai, Ari Glenn, Vikram Sivashankar, Daniel Zamoshchin, Leo Glikbarg, Derek Askaryar, Mike Yang, Teddy Zhang, Rishi Alluri, Nathan Tran, Rinnara Sangpisit, Polycarpos Yiorkadjis, Kenny Osele, Gautham Raghupathi, Dan Boneh, Daniel E. Ho, and Percy Liang · 2024
Later among the works it cites.
Improving Alignment and Robustness with Circuit Breakers, July 2024
Andy Zou, Long Phan, Justin Wang, Derek Duenas, Maxwell Lin, Maksym Andriushchenko, Rowan Wang, Zico Kolter, Matt Fredrikson, and Dan Hendrycks · 2024
Later among the works it cites.
Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation, March 2025
Bowen Baker, Joost Huizinga, Leo Gao, Zehao Dou, Melody Y. Guan, Aleksander Madry, Wojciech Zaremba, Jakub Pachocki, and David Farhi · 2025
Closest in time.
Dasher Identity Verification FAQ, 2025
DoorDash · 2025
Closest in time.
Multi-Agent Risks from Advanced AI, February 2025
Lewis Hammond, Alan Chan, Jesse Clifton, Jason Hoelscher-Obermaier, Akbir Khan, Euan McLean, Chandler Smith, Wolfram Barfuss, Jakob Foerster, Tomáš Gavenčiak, The Anh Han, Edward Hughes, Vojtěch Kovařík, Jan Kulveit, Joel Z. Leibo, Caspar Oesterheld, Christian Schroeder de Witt, Nisarg Shah, Michael Wellman, Paolo Bova, Theodor Cimpeanu, Carson Ezell, Quentin Feuillade-Montixi, Matija Franklin, Esben Kran, Igor Krawczuk, Max Lamparth, Niklas Lauffer, Alexander Meinke, Sumeet Motwani, Anka Reuel, Vincent Conitzer, Michael Dennis, Iason Gabriel, Adam Gleave, Gillian Hadfield, Nika Haghtalab, Atoosa Kasirzadeh, Sébastien Krier, Kate Larson, Joel Lehman, David C. Parkes, Georgios Piliouras, and Iyad Rahwan · 2025
Closest in time.
CVE: Common Vulnerabilities and Exposures, 2025
MITRE Corporation · 2025
Closest in time.
Law-Following AI: Designing AI Agents to Obey Human Laws, 2025
Cullen O’Keefe, Ketan Ramakrishnan, Janna Tay, and Christoph Winter · 2025
Closest in time.
Classical and Language Model Multi-Agent Infrastructure: A Comparative Review
Elija Perrier · 2025
Closest in time.
Language Model Agent Ontologies
Elija Perrier and Seth Lazar · 2025
Closest in time.
Announcing the Agent2Agent Protocol (A2A) - Google Developers Blog, April 2025
Rao Surapaneni, Miku Jha, Michael Vakoc, and Todd Segal · 2025
Closest in time.
Trapping misbehaving bots in an AI Labyrinth, March 2025
Reid Tatoris, Harsh Saxena, and Luis Miglietti · 2025
Closest in time.
Identity Verification Checks, 2025
Uber · 2025
Closest in time.