Fetching the paper…
Reading the bibliography…
Recent policy proposals aim to improve the safety of general-purpose AI, but there is little understanding of the efficacy of different regulatory approaches to AI safety.
The bargaining problem
John F Nash et al · 1950
Earlier work this paper cites.
The economic theory of agency: The principal’s problem
Stephen A Ross · 1973
Earlier work this paper cites.
An analysis of the principal-agent problem
Sanford J Grossman and Oliver D Hart · 1992
Earlier work this paper cites.
Product liability, research and development, and innovation
W Kip Viscusi and Michael J Moore · 1993
Earlier work this paper cites.
On the division of profit in sequential innovation
Jerry R Green and Suzanne Scotchmer · 1995
Earlier work this paper cites.
Supply chain coordination with contracts
Gérard P Cachon · 2003
Earlier work this paper cites.
Sequential innovation, patents, and imitation
James Bessen and Eric Maskin · 2009
Earlier work this paper cites.
Robustness and linear contracts
Gabriel Carroll · 2015
Earlier work this paper cites.
Strategic classification
Moritz Hardt, Nimrod Megiddo, Christos Papadimitriou, and Mary Wootters · 2016
Earlier work this paper cites.
Simple versus optimal contracts
Paul Dütting, Tim Roughgarden, and Inbal Talgam-Cohen · 2019
Earlier work this paper cites.
One for one, or all for all: Equilibria and optimality of collaboration in federated learning
Avrim Blum, Nika Haghtalab, Richard Lanas Phillips, and Han Shao · 2021
Earlier work this paper cites.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Earlier work this paper cites.
Model-sharing games: Analyzing federated learning under voluntary participation
Kate Donahue and Jon Kleinberg · 2021
Earlier work this paper cites.
Stateful strategic regression
Keegan Harris, Hoda Heidari, and Steven Z Wu · 2021
Cited alongside, same era.
Approximate optimality of linear contracts under uncertainty
Tal Alon, Paul Dütting, Yingkai Li, and Inbal Talgam-Cohen · 2022
Cited alongside, same era.
Strategic ranking
Lydia T Liu, Nikhil Garg, and Christian Borgs · 2022
Cited alongside, same era.
Taxonomy of Risks posed by Language Models
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, Amelia Glaese, Borja Balle, Atoosa Kasirzadeh, Courtney Biles, Sasha Brown, Zac Kenton, Will Hawkins, Tom Stepleton, Abeba Birhane, Lisa Anne Hendricks, Laura Rimell, William Isaac, Julia Haas, Sean Legassick, Geoffrey Irving, and Iason Gabriel · 2022
Cited alongside, same era.
Frontier ai regulation: Managing emerging risks to public safety
Markus Anderljung, Joslyn Barnhart, Anton Korinek, Jade Leung, Cullen O’Keefe, Jess Whittlestone, Shahar Avin, Miles Brundage, Justin Bullock, Duncan Cass-Beggs, et al · 2023
Cited alongside, same era.
Safety vs. performance: How multi-objective learning reduces barriers to market entry
Meena Jagadeesan, Michael I Jordan, and Jacob Steinhardt · 2024
Later among the works it cites.
Fine-tuning games: Bargaining and adaptation for general-purpose models
Benjamin Laufer, Jon Kleinberg, and Hoda Heidari · 2024
Later among the works it cites.
On evaluating the durability of safeguards for open-weight llms
Xiangyu Qi, Boyi Wei, Nicholas Carlini, Yangsibo Huang, Tinghao Xie, Luxi He, Matthew Jagielski, Milad Nasr, Prateek Mittal, and Peter Henderson · 2024
Later among the works it cites.
The economics of ai foundation models: Openness, competition, and governance
Fasheng Xu, Xiaoyu Wang, Wei Chen, and Karen Xie · 2024
Later among the works it cites.
International ai safety report
Yoshua Bengio, Sören Mindermann, Daniel Privitera, Tamay Besiroglu, Rishi Bommasani, Stephen Casper, Yejin Choi, Philip Fox, Ben Garfinkel, Danielle Goldfarb, et al · 2025
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ai supply chains
Sarah Huiyi Cen, Aspen Hopkins, Andrew Ilyas, Aleksander Madry, Isabella Struckman, and Luis Videgaray Caso · 2023
Cited alongside, same era.
Multi-agent contracts
Paul Dütting, Tomer Ezra, Michal Feldman, and Thomas Kesselheim · 2023
Cited alongside, same era.
Regulation priorities for artificial intelligence foundation models
Matthew R Gaske · 2023
Cited alongside, same era.
Fine-tuning aligned language models compromises safety, even when users do not intend to!
Xiangyu Qi, Yi Zeng, Tinghao Xie, Pin-Yu Chen, Ruoxi Jia, Prateek Mittal, and Peter Henderson · 2023
Cited alongside, same era.
Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction, July 2023
Renee Shelby, Shalaleh Rismani, Kathryn Henne, AJung Moon, Negar Rostamzadeh, Paul Nicholas, N’Mah Yilla, Jess Gallegos, Andrew Smart, Emilio Garcia, and Gurleen Virk · 2023
Cited alongside, same era.
A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms, July 2024
Gavin Abercrombie, Djalel Benbouzid, Paolo Giudici, Delaram Golpayegani, Julio Hernandez, Pierre Noro, Harshvardhan Pandit, Eva Paraschou, Charlie Pownall, Jyoti Prajapati, Mark A. Sayre, Ushnish Sengupta, Arthit Suriyawongkul, Ruby Thelot, Sofia Vei, and Laura Waltersdorfer · 2024
Cited alongside, same era.
Accounting for ai and users shaping one another: The role of mathematical models
Sarah Dean, Evan Dong, Meena Jagadeesan, and Liu Leqi · 2024
Cited alongside, same era.
Closest in time.
Draft report of the joint california policy working group on ai frontier models
Jennifer Tour Chayes, Mariano-Florentino Cuéllar, and Li Fei-Fei · 2025
Closest in time.
Multi-agent combinatorial contracts
Paul Dütting, Tomer Ezra, Michal Feldman, and Thomas Kesselheim · 2025
Closest in time.
In-house evaluation is not enough: Towards robust third-party flaw disclosure for general-purpose ai
Shayne Longpre, Kevin Klyman, Ruth E Appel, Sayash Kapoor, Rishi Bommasani, Michelle Sahar, Sean McGregor, Avijit Ghosh, Borhane Blili-Hamelin, Nathan Butters, et al · 2025
Closest in time.
Game theory meets large language models: A systematic survey
Haoran Sun, Yusen Wu, Yukun Cheng, and Xu Chu · 2025
Closest in time.
Selective response strategies for genai
Boaz Taitler and Omer Ben-Porat · 2025
Closest in time.
Data sharing with a generative ai competitor
Boaz Taitler, Omer Madmon, Moshe Tennenholtz, and Omer Ben-Porat · 2025
Closest in time.
Navigating the deployment dilemma and innovation paradox: Open-source versus closed-source models
Yanxuan Wu, Haihan Duan, Xitong Li, and Xiping Hu · 2025
Closest in time.