Fetching the paper…
Reading the bibliography…
Increased delegation of commercial, scientific, governmental, and personal activities to AI agents -- systems capable of pursuing complex goals with limited supervision -- may exacerbate existing societal risks and introduce new risks.
Theory of the firm: Managerial behavior, agency costs and ownership structure
Michael C. Jensen and William H. Meckling. 1976 · 1976
Earlier work this paper cites.
Micromotives and Macrobehavior
Thomas C. Schelling. 1978 · 1978
Earlier work this paper cites.
Moral Hazard and Observability
Bengt Holmström. 1979 · 1979
Earlier work this paper cites.
Agency Problems and Residual Claims
Eugene F. Fama and Michael C. Jensen. 1983 · 1983
Earlier work this paper cites.
Multitask Principal-Agent Analyses: Incentive Contracts, Asset Ownership, and Job Design
Bengt Holmstrom and Paul Milgrom. 1991 · 1991
Earlier work this paper cites.
The Out-of-the-Loop Performance Problem and Level of Control in Automation
Mica R. Endsley and Esin O. Kiris. 1995 · 1995
Earlier work this paper cites.
Accountability in a computerized society
Helen Nissenbaum. 1996 · 1996
Earlier work this paper cites.
Autonomous interface agents. In Proceedings of the ACM SIGCHI Conference on Human factors in computing systems (CHI ’97) . Association for Computing Machinery, New York, NY, USA, 67–74
Henry Lieberman. 1997 · 1997
Earlier work this paper cites.
Level of automation effects on performance, situation awareness and workload in a dynamic control task
M. R. Endsley and D. B. Kaber. 1999 · 1999
Earlier work this paper cites.
Scaling Laws for Neural Language Models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2001
Earlier work this paper cites.
The Theory of Incentives: The Principal-Agent Model
Jean-Jacques Laffont and David Martimort. 2002 · 2002
Earlier work this paper cites.
Seeking Truth for Power: Informational Strategy and Regulatory Policy Making
Cary Coglianese, Richard Zeckhauser, and Edward Parson. 2004 · 2004
Earlier work this paper cites.
Automation Bias in Intelligent Time Critical Decision Support Systems
Mary Cummings. 2004 · 2004
Earlier work this paper cites.
Freedom of Information and Openness: Fundamental Human Rights?
Patrick Birkinshaw. 2006 · 2006
Earlier work this paper cites.
Regulatory capture: A review
Ernesto Dal Bó. 2006 · 2006
Earlier work this paper cites.
Guide to Broker-Dealer Registration
Division of Trading and Markets. 2008 · 2008
Earlier work this paper cites.
How markets fail: The logic of economic calamities
John Cassidy. 2009 · 2009
Earlier work this paper cites.
Workplace surveillance: an overview
Kirstie Ball. 2010 · 2010
Earlier work this paper cites.
Complex networks: structure, robustness and function
Reuven Cohen and Shlomo Havlin. 2010 · 2010
Earlier work this paper cites.
A 61-million-person experiment in social influence and political mobilization
Robert M. Bond, Christopher J. Fariss, Jason J. Jones, Adam D. I. Kramer, Cameron Marlow, Jaime E. Settle, and James H. Fowler. 2012 · 2012
Earlier work this paper cites.
NSA Prism program taps in to user data of Apple, Google and others
Glenn Greenwald and Ewen MacAskill. 2013 · 2013
Earlier work this paper cites.
Corrigibility. In Workshops at the Twenty-Ninth AAAI Conference on Artificial Intelligence
Nate Soares, Benja Fallenstein, Stuart Armstrong, and Eliezer Yudkowsky. 2015 · 2015
Earlier work this paper cites.
Artificial Intelligence, Automation, and Work
Daron Acemoglu and Pascual Restrepo. 2018 · 2018
Earlier work this paper cites.
Seeing without knowing: Limitations of the transparency ideal and its application to algorithmic accountability
Mike Ananny and Kate Crawford. 2018 · 2018
Earlier work this paper cites.
Cooperation or resistance?: The role of tech companies in government surveillance
Chloe Goodwin. 2018 · 2018
Earlier work this paper cites.
Intelligent Autonomous Agents are Key to Cyber Defense of the Future Army Networks
Alexander Kott. 2018 · 2018
Earlier work this paper cites.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2019 · 2018
Earlier work this paper cites.
Reinforcement learning: An introduction (second edition ed.)
Richard S. Sutton and Andrew G. Barto. 2018 · 2018
Earlier work this paper cites.
Automation and New Tasks: How Technology Displaces and Reinstates Labor
Daron Acemoglu and Pascual Restrepo. 2019 · 2019
Earlier work this paper cites.
Bottom-up data Trusts: disturbing the ‘one size fits all’ approach to data governance
Sylvie Delacroix and Neil D Lawrence. 2019 · 2019
Earlier work this paper cites.
A New Law Makes Bots Identify Themselves—That’s the Problem
Renee DiResta. 2019 · 2019
Earlier work this paper cites.
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru. 2019 · 2019
Earlier work this paper cites.
Regulatory Monitors: Policing Firms in the Compliance Era
Rory Van Loo. 2019 · 2019
Earlier work this paper cites.
Thinking about risks from AI: Accidents, misuse and structure
Remco Zwetsloot and Allan Dafoe. 2019 · 2019
Earlier work this paper cites.
Monitoring approaches for health-care workers during the COVID-19 pandemic
Julia A. Bielicki, Xavier Duval, Nina Gobat, Herman Goossens, Marion Koopmans, Evelina Tacconelli, and Sylvie van der Werf. 2020 · 2020
Earlier work this paper cites.
States’ Automated Systems Are Trapping Citizens in Bureaucratic Nightmares With Their Lives on the Line
Alejandro De La Garza. 2020 · 2020
Earlier work this paper cites.
Democratising the Digital Revolution: The Role of Data Governance
Sylvie Delacroix, Joelle Pineau, and Jessica Montgomery. 2020 · 2020
Earlier work this paper cites.
They who must not be identified—distinguishing personal from non-personal data under the GDPR
Michèle Finck and Frank Pallas. 2020 · 2020
Earlier work this paper cites.
Monitoring Misuse for Accountable ’Artificial Intelligence as a Service’. In Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society (AIES ’20) . Association for Computing Machinery, New York, NY, USA, 300–306
Seyyed Ahmad Javadi, Richard Cloete, Jennifer Cobbe, Michelle Seng Ah Lee, and Jatinder Singh. 2020 · 2020
Earlier work this paper cites.
Mitigating bias in algorithmic hiring: evaluating claims and practices. In Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency (FAT* ’20) . Association for Computing Machinery, New York, NY, USA, 469–481
Manish Raghavan, Solon Barocas, Jon Kleinberg, and Karen Levy. 2020 · 2020
Earlier work this paper cites.
An introduction to complex systems science and its applications
Alexander F Siegenfeld and Yaneer Bar-Yam. 2020 · 2020
Earlier work this paper cites.
The philosophical basis of algorithmic recourse. In Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency . ACM
Suresh Venkatasubramanian and Mark Alfano. 2020 · 2020
Earlier work this paper cites.
Proposal for a Regulation of the European Parliament and of the Council Laying Down Harmonised Rules on Artificial Intelligence (Artificial Intelligence Act) and Amending Certain Union Legislative Acts
2021 · 2021
Earlier work this paper cites.
Proactive Monitoring of Markets and Institutions
Board of Governors of the Federal Reserve System. 2021 · 2021
Earlier work this paper cites.
The Society of Algorithms
Jenna Burrell and Marion Fourcade. 2021 · 2021
Earlier work this paper cites.
Algorithmic collusion: A critical review
Florian E. Dorner. 2021 · 2021
Earlier work this paper cites.
Datasheets for datasets
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé Iii, and Kate Crawford. 2021 · 2021
Earlier work this paper cites.
Surveillance on Healthcare Workers During the First Wave of SARS-CoV-2 Pandemic in Italy: The Experience of a Tertiary Care Pediatric Hospital
Valentina Guarnieri, Maria Moriondo, Mattia Giovannini, Lorenzo Lodi, Silvia Ricci, Laura Pisano, Paola Barbacci, Costanza Bini, Giuseppe Indolfi, Alberto Zanobini, and Chiara Azzari. 2021 · 2021
Earlier work this paper cites.
AI and International Stability: Risks and Confidence-Building Measures
Michael Horowitz and Paul Scharre. 2021 · 2021
Earlier work this paper cites.
Monitoring AI Services for Misuse. In Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’21) . Association for Computing Machinery, New York, NY, USA, 597–607
Seyyed Ahmad Javadi, Chris Norval, Richard Cloete, and Jatinder Singh. 2021 · 2021
Earlier work this paper cites.
Algorithmic Recourse: From Counterfactual Explanations to Interventions. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’21) . Association for Computing Machinery, New York, NY, USA, 353–362
Amir-Hossein Karimi, Bernhard Schölkopf, and Isabel Valera. 2021 · 2021
Earlier work this paper cites.
Artificial Intelligence: A Modern Approach (4 ed.)
Stuart J. Russell and Peter Norvig. 2021 · 2021
Earlier work this paper cites.
Debunking the AI Arms Race Theory (Summer 2021)
Paul Scharre. 2021 · 2021
Earlier work this paper cites.
Is Instagram Causing Poorer Mental Health Among Teen Girls? - Is Instagram Causing Poorer Mental Health Among Teen Girls? - United States Joint Economic Committee
Rachel Sheffield and Catherine Francois. 2021 · 2021
Earlier work this paper cites.
Data Hiding with Deep Learning: A Survey Unifying Digital Watermarking and Steganography
Zihan Wang, Olivia Byrnes, Hu Wang, Ruoxi Sun, Congbo Ma, Huaming Chen, Qi Wu, and Minhui Xue. 2021 · 2021
Earlier work this paper cites.
12 CFR § 1026.25 - Record retention
2022 · 2022
Earlier work this paper cites.
Constitutional AI: Harmlessness from AI Feedback
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosuite, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemi Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan. 2022 · 2022
Earlier work this paper cites.
PowerGridworld: a framework for multi-agent reinforcement learning in power systems. In Proceedings of the Thirteenth ACM International Conference on Future Energy Systems (e-Energy ’22) . Association for Computing Machinery, New York, NY, USA, 565–570
David Biagioni, Xiangyu Zhang, Dylan Wald, Deepthi Vaidhynathan, Rohit Chintala, Jennifer King, and Ahmed S. Zamzam. 2022 · 2022
Earlier work this paper cites.
Picking on the Same Person: Does Algorithmic Monoculture lead to Outcome Homogenization?
Rishi Bommasani, Kathleen A. Creel, Ananya Kumar, Dan Jurafsky, and Percy Liang. 2022 · 2022
Cited alongside, same era.
The Dangers of Underclaiming: Reasons for Caution When Reporting How NLP Systems Fail. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Dublin, Ireland, 7484–7499
Samuel Bowman. 2022 · 2022
Cited alongside, same era.
LangChain 0.0.77 Docs
Harrison Chase. 2022 · 2022
Cited alongside, same era.
17 CFR § 240.17a-3 - Records to be made by certain exchange members, brokers and dealers
Securities and Exchange Commission. 2022a · 2022
Cited alongside, same era.
17 CFR § 240.17a-4 - Records to be preserved by certain exchange members, brokers and dealers
Securities and Exchange Commission. 2022b · 2022
Trends in machine learning hardware
Marius Hobbhahn, Lennart Heim, and Gökçe Aydos. 2023 · 2023
Later among the works it cites.
National Cybersecurity Strategy
The White House. 2023 · 2023
Later among the works it cites.
Generative AI and the Digital Commons
Saffron Huang and Divya Siddarth. 2023 · 2023
Later among the works it cites.
Co-Writing with Opinionated Language Models Affects Users’ Views. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (CHI ’23) . Association for Computing Machinery, New York, NY, USA, 1–15
Maurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson, and Mor Naaman. 2023 · 2023
Later among the works it cites.
Working With AI to Persuade: Examining a Large Language Model’s Ability to Generate Pro-Vaccination Messages
Elise Karinshak, Sunny Xun Liu, Joon Sung Park, and Jeffrey T. Hancock. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Accountability in an Algorithmic Society: Relationality, Responsibility, and Robustness in Machine Learning. In 2022 ACM Conference on Fairness, Accountability, and Transparency . ACM
A. Feder Cooper, Emanuel Moss, Benjamin Laufer, and Helen Nissenbaum. 2022 · 2022
Cited alongside, same era.
Payment Card Industry Data Security Standard
PCI Security Standars Council. 2022 · 2022
Cited alongside, same era.
Magnetic control of tokamak plasmas through deep reinforcement learning
Jonas Degrave, Federico Felici, Jonas Buchli, Michael Neunert, Brendan Tracey, Francesco Carpanese, Timo Ewalds, Roland Hafner, Abbas Abdolmaleki, Diego de las Casas, Craig Donner, Leslie Fritz, Cristian Galperti, Andrea Huber, James Keeling, Maria Tsimpoukelli, Jackie Kay, Antoine Merle, Jean-Marc Moret, Seb Noury, Federico Pesamosca, David Pfau, Olivier Sauter, Cristian Sommariva, Stefano Coda, Basil Duval, Ambrogio Fasoli, Pushmeet Kohli, Koray Kavukcuoglu, Demis Hassabis, and Martin Riedmiller. 2022 · 2022
Cited alongside, same era.
Predictability and Surprise in Large Generative Models. In Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’22) . Association for Computing Machinery, New York, NY, USA, 1747–1764
Deep Ganguli, Danny Hernandez, Liane Lovitt, Amanda Askell, Yuntao Bai, Anna Chen, Tom Conerly, Nova Dassarma, Dawn Drain, Nelson Elhage, Sheer El Showk, Stanislav Fort, Zac Hatfield-Dodds, Tom Henighan, Scott Johnston, Andy Jones, Nicholas Joseph, Jackson Kernian, Shauna Kravec, Ben Mann, Neel Nanda, Kamal Ndousse, Catherine Olsson, Daniela Amodei, Tom Brown, Jared Kaplan, Sam McCandlish, Christopher Olah, Dario Amodei, and Jack Clark. 2022 · 2022
Cited alongside, same era.
The flaws of policies requiring human oversight of government algorithms
Ben Green. 2022 · 2022
Cited alongside, same era.
Trends in GPU Price-Performance
Marius Hobbhahn and Tamay Besiroglu. 2022 · 2022
Cited alongside, same era.
Training Compute-Optimal Large Language Models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, Tom Hennigan, Eric Noland, Katie Millican, George van den Driessche, Bogdan Damoc, Aurelia Guy, Simon Osindero, Karen Simonyan, Erich Elsen, Jack W. Rae, Oriol Vinyals, and Laurent Sifre. 2022 · 2022
Cited alongside, same era.
Zachary Kenton, Ramana Kumar, Sebastian Farquhar, Jonathan Richens, Matt MacDermott, and Tom Everitt. 2023 · 2023
Later among the works it cites.
Evaluating Language-Model Agents on Realistic Autonomous Tasks
Megan Kinniment, Lucas Jun Koba Sato, Haoxing Du, Brian Goodrich, Max Hasin, Lawrence Chan, Luke Harold Miles, Tao R. Lin, Hjalmar Wijk, Joel Burget, Aaron Ho, Elizabeth Barnes, and Paul Christiano. 2023 · 2023
Later among the works it cites.
Self-preferencing by platforms: A literature review
Yuta Kittaka, Susumu Sato, and Yusuke Zennyo. 2023 · 2023
Later among the works it cites.
Leonie Koessler and Jonas Schuett. 2023 · 2023
Later among the works it cites.
AgentBench: Evaluating LLMs as Agents
Xiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu, Xuanyu Lei, Hanyu Lai, Yu Gu, Hangliang Ding, Kaiwen Men, Kejuan Yang, Shudan Zhang, Xiang Deng, Aohan Zeng, Zhengxiao Du, Chenhui Zhang, Sheng Shen, Tianjun Zhang, Yu Su, Huan Sun, Minlie Huang, Yuxiao Dong, and Jie Tang. 2023 · 2023
Later among the works it cites.
GAIA: a benchmark for General AI Assistants
Grégoire Mialon, Clémentine Fourrier, Craig Swift, Thomas Wolf, Yann LeCun, and Thomas Scialom. 2023 · 2023
Later among the works it cites.
Engagement, User Satisfaction, and the Amplification of Divisive Content on Social Media
Smitha Milli, Micah Carroll, Yike Wang, Sashrika Pandey, Sebastian Zhao, and Anca D. Dragan. 2023 · 2023
Later among the works it cites.
Levels of AGI: Operationalizing Progress on the Path to AGI
Meredith Ringel Morris, Jascha Sohl-dickstein, Noah Fiedel, Tris Warkentin, Allan Dafoe, Aleksandra Faust, Clement Farabet, and Shane Legg. 2023 · 2023
Later among the works it cites.
Azure OpenAI Service abuse monitoring - Azure OpenAI
mrbullwinkle and eric urban. 2023 · 2023
Later among the works it cites.
Auditing Large Language Models: A Three-Layered Approach
Jakob Mökander, Jonas Schuett, Hannah Rose Kirk, and Luciano Floridi. 2023 · 2023
Later among the works it cites.
Testing Language Model Agents Safely in the Wild
Silen Naihin, David Atkinson, Marc Green, Merwane Hamadi, Craig Swift, Douglas Schonholtz, Adam Tauman Kalai, and David Bau. 2023 · 2023
Later among the works it cites.
Deployment Corrections: An incident response framework for frontier AI models
Joe O’Brien, Shaun Ee, and Zoe Williams. 2023 · 2023
Later among the works it cites.
Guidance Regarding Methods for De-identification of Protected Health Information in Accordance with the Health Insurance Portability and Accountability Act (HIPAA) Privacy Rule
U.S. Department of Human and Health Services. 2022 · 2023
Later among the works it cites.
’Generative CI’ through Collective Response Systems
Aviv Ovadya. 2023 · 2023
Later among the works it cites.
Alexander Pan, Chan Jun Shern, Andy Zou, Nathaniel Li, Steven Basart, and others. 2023 · 2023
Later among the works it cites.
Generative Agents: Interactive Simulacra of Human Behavior. In Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (UIST ’23) . Association for Computing Machinery, New York, NY, USA, 1–22
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2023 · 2023
Later among the works it cites.
Auto-GPT: An Autonomous GPT-4 Experiment
Toran Bruce Richards. 2023 · 2023
Later among the works it cites.
Infographic: Amazon Maintains Lead in the Cloud Market
Felix Richter. 2023 · 2023
Later among the works it cites.
From Plane Crashes to Algorithmic Harm: Applicability of Safety Engineering Frameworks for Responsible ML. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (CHI ’23) . Association for Computing Machinery, New York, NY, USA, 1–18
Shalaleh Rismani, Renee Shelby, Andrew Smart, Edgar Jatho, Joshua Kroll, AJung Moon, and Negar Rostamzadeh. 2023 · 2023
Later among the works it cites.
Preventing Language Models From Hiding Their Reasoning
Fabien Roger and Ryan Greenblatt. 2023 · 2023
Later among the works it cites.
Identifying the Risks of LM Agents with an LM-Emulated Sandbox
Yangjun Ruan, Honghua Dong, Andrew Wang, Silviu Pitis, Yongchao Zhou, Jimmy Ba, Yann Dubois, Chris J. Maddison, and Tatsunori Hashimoto. 2023 · 2023
Later among the works it cites.
Jonas B. Sandbrink. 2023 · 2023
Later among the works it cites.
Toolformer: Language Models Can Teach Themselves to Use Tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Later among the works it cites.
Democratising AI: Multiple Meanings, Goals, and Methods. In Proceedings of the 2023 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’23) . Association for Computing Machinery, New York, NY, USA, 715–722
Elizabeth Seger, Aviv Ovadya, Divya Siddarth, Ben Garfinkel, and Allan Dafoe. 2023 · 2023
Later among the works it cites.
A Causal Framework for AI Regulation and Auditing
Lee Sharkey, Clíodhna Ní Ghuidhir, Dan Braun, Jérémy Scheurer, Mikita Balesni, Lucius Bushnaq, Charlotte Stix, and Marius Hobbhahn. 2023 · 2023
Later among the works it cites.
Practices for Governing Agentic AI Systems
Yonadav Shavit, Sandhini Agarwal, Miles Brundage, Steven Adler, Cullen O’Keefe, Rosie Campbell, Teddy Lee, Pamela Mishkin, Tyna Eloundou, Alan Hickey, Katarina Slama, Lama Ahmad, Paul McMillan, Alex Beutel, Alexandre Passos, and David G. Robinson. 2023 · 2023
Later among the works it cites.
Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction. In Proceedings of the 2023 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’23) . Association for Computing Machinery, New York, NY, USA, 723–741
Renee Shelby, Shalaleh Rismani, Kathryn Henne, AJung Moon, Negar Rostamzadeh, Paul Nicholas, N’Mah Yilla-Akbari, Jess Gallegos, Andrew Smart, Emilio Garcia, and Gurleen Virk. 2023 · 2023
Later among the works it cites.
HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face
Yongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li, Weiming Lu, and Yueting Zhuang. 2023 · 2023
Later among the works it cites.
Model evaluation for extreme risks
Toby Shevlane, Sebastian Farquhar, Ben Garfinkel, Mary Phuong, Jess Whittlestone, Jade Leung, Daniel Kokotajlo, Nahema Marchal, Markus Anderljung, Noam Kolt, Lewis Ho, Divya Siddarth, Shahar Avin, Will Hawkins, Been Kim, Iason Gabriel, Vijay Bolina, Jack Clark, Yoshua Bengio, Paul Christiano, and Allan Dafoe. 2023 · 2023
Later among the works it cites.
Can large language models democratize access to dual-use biotechnology?
Emily H. Soice, Rafael Rocha, Kimberlee Cordova, Michael Specter, and Kevin M. Esvelt. 2023 · 2023
Later among the works it cites.
The Gradient of Generative AI Release: Methods and Considerations
Irene Solaiman. 2023 · 2023
Later among the works it cites.
Cognitive Architectures for Language Agents
Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan, and Thomas L. Griffiths. 2023 · 2023
Later among the works it cites.
The rise of AI fake news is creating a ‘misinformation superspreader’
Pranshu Verma. 2023a · 2023
Later among the works it cites.
They thought loved ones were calling for help. It was an AI scam
Pranshu Verma. 2023b · 2023
Later among the works it cites.
A Survey on Large Language Model based Autonomous Agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, Wayne Xin Zhao, Zhewei Wei, and Ji-Rong Wen. 2023 · 2023
Later among the works it cites.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Later among the works it cites.
Sociotechnical Safety Evaluation of Generative AI Systems
Laura Weidinger, Maribeth Rauh, Nahema Marchal, Arianna Manzini, Lisa Anne Hendricks, Juan Mateos-Garcia, Stevie Bergman, Jackie Kay, Conor Griffin, Ben Bariach, Iason Gabriel, Verena Rieser, and William Isaac. 2023 · 2023
Later among the works it cites.
Open (for Business): Big Tech, Concentrated Power, and the Political Economy of Open AI
David Gray Widder, Meredith Whittaker, and Sarah Myers West. 2023 · 2023
Later among the works it cites.
Prompt injection: What’s the worst that can happen?
Simon Willison. 2023 · 2023
Later among the works it cites.
AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation Framework
Qingyun Wu, Gagan Bansal, Jieyu Zhang, Yiran Wu, Beibin Li, Erkang Zhu, Li Jiang, Xiaoyun Zhang, Shaokun Zhang, Jiale Liu, Ahmed Hassan Awadallah, Ryen W. White, Doug Burger, and Chi Wang. 2023 · 2023
Later among the works it cites.
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
Zelai Xu, Chao Yu, Fei Fang, Yu Wang, and Yi Wu. 2023 · 2023
Later among the works it cites.
Watermarks in the Sand: Impossibility of Strong Watermarking for Generative Models
Hanlin Zhang, Benjamin L. Edelman, Danilo Francati, Daniele Venturi, Giuseppe Ateniese, and Boaz Barak. 2023 · 2023
Later among the works it cites.
Universal and Transferable Adversarial Attacks on Aligned Language Models
Andy Zou, Zifan Wang, Nicholas Carlini, Milad Nasr, J. Zico Kolter, and Matt Fredrikson. 2023 · 2023
Later among the works it cites.
Artificial Intelligence Can Persuade Humans on Political Issues
(Max) Hui Bai, Jan G. Voelkel, johannes C. Eichstaedt, and Robb Willer. 2024 · 2024
Closest in time.
Let’s Encrypt Stats - Let’s Encrypt
Let’s Encrypt. 2024 · 2024
Closest in time.
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening
Chengguang Gan, Qinghao Zhang, and Tatsunori Mori. 2024 · 2024
Closest in time.
AI Control: Improving Safety Despite Intentional Subversion
Ryan Greenblatt, Buck Shlegeris, Kshitij Sachan, and Fabien Roger. 2024 · 2024
Closest in time.
Multi-Agent Risks from Advanced AI
Lewis Hammond and TBD. 2024 · 2024
Closest in time.
A Survey of Text Watermarking in the Era of Large Language Models
Aiwei Liu, Leyi Pan, Yijian Lu, Jingjing Li, Xuming Hu, Lijie Wen, Irwin King, and Philip S. Yu. 2024 · 2024
Closest in time.
Judges in England and Wales Given Cautious Approval to Use AI
Brian Melley. 2024 · 2024
Closest in time.
Introducing the GPT Store
OpenAI. 2024 · 2024
Closest in time.
Microsoft’s new Copilot Pro brings AI-powered Office features to the rest of us
Tom Warren. 2024 · 2024
Closest in time.