Fetching the paper…
Reading the bibliography…
As foundation models have accumulated hundreds of millions of users, developers have begun to take steps to prevent harmful types of uses.
Harm reduction: Come as you are
G Alan Marlatt · 1996
Earlier work this paper cites.
Internet acceptable use policies: Navigating the management, legal, and technical issues
Farley Stewart · 2000
Earlier work this paper cites.
Acceptable internet use policy
Keng Siau, Fiona Fui-Hoon Nah, and Limei Teng · 2002
Earlier work this paper cites.
Reconstructing the software license
Michael J Madison · 2003
Earlier work this paper cites.
Sex-work harm reduction
Michael L Rekart · 2005
Earlier work this paper cites.
The qualitative content analysis process
Satu Elo and Helvi Kyngäs · 2007
Earlier work this paper cites.
Understanding compliance with internet use policy from the perspective of rational choice theory
Han Li, Jie Zhang, and Rathindra Sarathy · 2009
Earlier work this paper cites.
Reinforcing the security of corporate information resources: A critical review of the role of the acceptable use policy
Neil Doherty, Leonidas Anastasakis, and Heather Fulford · 2010
Earlier work this paper cites.
Crows-pairs: A challenge dataset for measuring social biases in masked language models, 2020
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel R. Bowman · 2010
Earlier work this paper cites.
Ethical decision making: Improving the quality of acceptable use policies
A.B. Ruighaver, S.B. Maynard, and M. Warren · 2010
Earlier work this paper cites.
Social media access in k-12 schools: Intractable policy controversies in an evolving world
June Ahn, Lauren K. Bivona, and Jeffrey DiScala · 2011
Earlier work this paper cites.
Busting the ghost guns: A technical, statutory, and practical approach to the 3-d printed weapon problem
Katherine E. Beyer · 2014
Earlier work this paper cites.
3d printing in libraries: A view from within the american library association: Privacy, intellectual freedom and ethical policy framework
Barbara M. Jones · 2015
Earlier work this paper cites.
Qualitative Content Analysis: Theoretical Background and Procedures
Philipp Mayring, Angelika Bikner-Ahsbahs, Christine Knipping, and Norma Presmeg · 2015
Earlier work this paper cites.
The Library’s Legal Answers for Makerspaces
Mary Minow, Tomas A. Lipinski, Gretchen McCord, et al · 2016
Earlier work this paper cites.
Characterizations of online harassment: Comparing policies across social media platforms
Jessica A. Pater, Moon K. Kim, Elizabeth D. Mynatt, and Casey Fiesler · 2016
Earlier work this paper cites.
Privacy and security in online social networks: A survey
Imrul Kayes and Adriana Iamnitchi · 2017
Earlier work this paper cites.
Custodians of the Internet: Platforms, Content Moderation, and the Hidden Decisions that Shape Social Media
T. Gillespie · 2018
Earlier work this paper cites.
Online harassment and content moderation: The case of blocklists
Shagun Jhaver, Sucheta Ghoshal, Amy Bruckman, and Eric Gilbert · 2018
Earlier work this paper cites.
The biggest lie on the internet: ignoring the privacy policies and terms of service policies of social networking services
Jonathan A. Obar and Anne Oeldorf-Hirsch · 2018
Earlier work this paper cites.
Internet sex work: Beyond the gaze
Teela Sanders, Jane Scoular, Rosie Campbell, Jane Pitcher, and Stewart Cunningham · 2018
Earlier work this paper cites.
Safe for work: Feminist porn, corporate regulation and community standards
Zahra Stardust · 2018
Earlier work this paper cites.
Gender bias in coreference resolution: Evaluation and debiasing methods, 2018
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang · 2018
Earlier work this paper cites.
The global expansion of AI surveillance
Steven Feldstein · 2019
Earlier work this paper cites.
Translating principles into practices of digital ethics: Five risks of being unethical
Luciano Floridi · 2019
Earlier work this paper cites.
Ghost Work: How to Stop Silicon Valley from Building a New Global Underclass
M.L. Gray and S. Suri · 2019
Earlier work this paper cites.
Hacking in the cloud
Nick Gregorio, Janahan Mathanamohan, Qusay H. Mahmoud, and May AlTaei · 2019
Earlier work this paper cites.
The global landscape of ai ethics guidelines
Anna Jobin, Marcello Ienca, and Effy Vayena · 2019
Earlier work this paper cites.
Public library digital services: Emergent issues of access and acceptable use, 2019
D. McMenemy, University of Strathclyde. Department of Computer, and Information Sciences · 2019
Earlier work this paper cites.
Model cards for model reporting
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru · 2019
Earlier work this paper cites.
Acceptable Use Policies
W. Ian O’Byrne · 2019
Earlier work this paper cites.
Behind the Screen: Content Moderation in the Shadows of Social Media
S.T. Roberts · 2019
Earlier work this paper cites.
Technologies for social justice: Lessons from sex workers on the front lines
Angelika Strohmayer, Jenn Clamen, and Mary Laing · 2019
Earlier work this paper cites.
A right to reasonable inferences: re-thinking data protection law in the age of big data and ai
Sandra Wachter and Brent Mittelstadt · 2019
Earlier work this paper cites.
The acceptable state: An analysis of the current state of acceptable use policies in academic institutions
Jake Weidman and Jens Grossklags · 2019
Earlier work this paper cites.
An institutionalist approach to ai ethics: Justifying the priority of government regulation over self-regulation
Thomas Ferretti · 2020
Earlier work this paper cites.
Principled artificial intelligence: Mapping consensus in ethical and rights-based approaches to principles for ai
Jessica Fjeld, Nele Achten, Hannah Hilligoss, Adam Nagy, and Madhulika Srikumar · 2020
Earlier work this paper cites.
Camming: Money, Power, and Pleasure in the Sex Work Industry
A. Jones · 2020
Earlier work this paper cites.
The ai-based cyber threat landscape: A survey
Nektaria Kaloudi and Jingyue Li · 2020
Earlier work this paper cites.
‘to be understood as to understand’: A readability analysis of public library acceptable use policies
Elaine Robinson and David McMenemy · 2020
Earlier work this paper cites.
Algorithmic warfare and the reinvention of accuracy
Lucy Suchman · 2020
Earlier work this paper cites.
The use of information and communication technologies by sex workers to manage occupational health and safety: Scoping review
T. Bernier, A. Shah, L. E. Ross, C. H. Logie, and E. Seto · 2021
Earlier work this paper cites.
Deplatforming sexual speech in the age of fosta/sesta
C. Bronstein · 2021
Earlier work this paper cites.
The long fuse: Misinformation and the 2020 election, 2021
Center for an Informed Public, Digital Forensic Research Lab, Graphika, and Stanford Internet Observatory · 2021
Earlier work this paper cites.
Reconfiguring diversity and inclusion for ai ethics
Nicole Chi, Emma Lurie, and Deirdre K. Mulligan · 2021
Earlier work this paper cites.
Disproportionate removals and differing content moderation experiences for conservative, transgender, and black social media users: Marginalization and moderation gray areas
Oliver L Haimson, Daniel Delmonaco, Peipei Nie, and Andrea Wegner · 2021
Earlier work this paper cites.
Amplification and its discontents: Why regulating the reach of online content is hard
Daphne Keller · 2021
Earlier work this paper cites.
Conceptualizing ai literacy: An exploratory review
Davy Tsz Kit Ng, Jac Ka Lok Leung, Samuel Kai Wah Chu, and Maggie Shen Qiao · 2021
Earlier work this paper cites.
The panoptic principle: privacy and surveillance in the public library as evidenced in the acceptable use policy
Elaine Robinson · 2021
Earlier work this paper cites.
Governing hate: Facebook and digital racism
Eugenia Siapera and Paloma Viejo-Otero · 2021
Earlier work this paper cites.
Privacy, data sharing, and data security policies of women’s mhealth apps: Scoping review and content analysis
Nouf Alfawzan, Markus Christen, Giovanni Spitale, and Nikola Biller-Andorno · 2022
Earlier work this paper cites.
Constitutional ai: Harmlessness from ai feedback, 2022
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosuite, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemi Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan · 2022
Earlier work this paper cites.
Power to the people? opportunities and challenges for participatory ai
Abeba Birhane, William Isaac, Vinodkumar Prabhakaran, Mark Diaz, Madeleine Clare Elish, Iason Gabriel, and Shakir Mohamed · 2022
Earlier work this paper cites.
Fusing finetuned models for better pretraining, 2022
Leshem Choshen, Elad Venezian, Noam Slonim, and Yoav Katz · 2022
Earlier work this paper cites.
Best practices for deploying language models, Jul 2022
Cohere, OpenAI, and AI21 Labs · 2022
Earlier work this paper cites.
Behavioral use licensing for responsible ai
Danish Contractor, Daniel McDuff, Julia Katherine Haines, Jenny Lee, Christopher Hines, Brent Hecht, Nicholas Vincent, and Hanlin Li · 2022
Earlier work this paper cites.
Accountability in an algorithmic society: Relationality, responsibility, and robustness in machine learning
A. Feder Cooper, Emanuel Moss, Benjamin Laufer, and Helen Nissenbaum · 2022
Earlier work this paper cites.
Meme Wars: The Untold Story of the Online Battles Upending Democracy in America
J. Donovan, E. Dreyfuss, and B. Friedberg · 2022
Earlier work this paper cites.
Content moderation as systems thinking
Evelyn Douek · 2022
Earlier work this paper cites.
Risk, resilience and reward: Impacts of shifting to digital sex work
Vaughn Hamilton, Hanna Barakat, and Elissa M. Redmiles · 2022
Earlier work this paper cites.
Cloud Empires: How Digital Platforms Are Overtaking the State and How We Can Regain Control
V. Lehdonvirta · 2022
Earlier work this paper cites.
Ethical challenges in ai approaches to eating disorders
Gemma Sharp, John Torous, and Madeline L. West · 2022
Cited alongside, same era.
A sandbox approach to regulating high-risk artificial intelligence applications
Jon Truby, Rafael Dean Brown, Imad Antoine Ibrahim, and Oriol Caudevilla Parellada · 2022
Cited alongside, same era.
Evaluating the rail license family, November 2022
Luis Villa · 2022
Cited alongside, same era.
The ethical implications of generative audio models: A systematic literature review
Julia Barnett · 2023
Cited alongside, same era.
Considerations for governing open foundation models
Rishi Bommasani, Sayash Kapoor, Kevin Klyman, Shayne Longpre, Ashwin Ramaswami, Daniel Zhang, Marietje Schaake, Daniel E. Ho, Arvind Narayanan, and Percy Liang · 2023
Cited alongside, same era.
Ecosystem graphs: The social footprint of foundation models, 2023
Reducing risks posed by synthetic content, April 2024
Bilva Chandra, George Awad, Yooyoung Lee, Peter Fontana, Razvan Amironesei, Mark Przybocki, Kamie Roberts, Elham Tabassi, Mat Heyman, and Jesse Dunietz · 2024
Closest in time.
Ai foundation models: Technical update report
Competition and Markets Authority · 2024
Closest in time.
From fitting participation to forging relationships: The art of participatory ml
Ned Cooper and Alexandra Zafiroglu · 2024
Closest in time.
Interim measures for the management of generative artificial intelligence services
Cyberspace Administration of China · 2024
Closest in time.
Large legal fictions: Profiling legal hallucinations in large language models
Matthew Dahl, Varun Magesh, Mirac Suzgun, and Daniel E Ho · 2024
Closest in time.
The meta oversight board and the empty promise of legitimacy
Evelyn Douek · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rishi Bommasani, Dilara Soylu, Thomas I. Liao, Kathleen A. Creel, and Percy Liang · 2023
Cited alongside, same era.
The foundation model transparency index, 2023
Bommasani and Klyman, Shayne Longpre, Sayash Kapoor, Nestor Maslej, Betty Xiong, Daniel Zhang, and Percy Liang · 2023
Cited alongside, same era.
Rt-2: Vision-language-action models transfer web knowledge to robotic control, 2023
Anthony Brohan, Noah Brown, Justice Carbajal, Yevgen Chebotar, Xi Chen, Krzysztof Choromanski, Tianli Ding, Danny Driess, Avinava Dubey, Chelsea Finn, Pete Florence, Chuyuan Fu, Montse Gonzalez Arenas, Keerthana Gopalakrishnan, Kehang Han, Karol Hausman, Alexander Herzog, Jasmine Hsu, Brian Ichter, Alex Irpan, Nikhil Joshi, Ryan Julian, Dmitry Kalashnikov, Yuheng Kuang, Isabel Leal, Lisa Lee, Tsang-Wei Edward Lee, Sergey Levine, Yao Lu, Henryk Michalewski, Igor Mordatch, Karl Pertsch, Kanishka Rao, Krista Reymann, Michael Ryoo, Grecia Salazar, Pannag Sanketi, Pierre Sermanet, Jaspiar Singh, Anikait Singh, Radu Soricut, Huong Tran, Vincent Vanhoucke, Quan Vuong, Ayzaan Wahid, Stefan Welker, Paul Wohlhart, Jialin Wu, Fei Xia, Ted Xiao, Peng Xu, Sichun Xu, Tianhe Yu, and Brianna Zitkovich · 2023
Cited alongside, same era.
Ai supply chains, April 2023
Sarah Huiyi Cen, Aspen Hopkins, Andrew Ilyas, Aleksander Madry, Isabella Struckman, and Luis Videgaray Caso · 2023
Cited alongside, same era.
Deep reinforcement learning from human preferences, 2023
Paul Christiano, Jan Leike, Tom B. Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2023
Cited alongside, same era.
Riley reid on ai: ’i don’t want porn to get left behind’, Oct 2023
Samantha Cole · 2023
Cited alongside, same era.
Ai foundation models initial report
Competition and Markets Authority · 2023
Cited alongside, same era.
Closest in time.
Risks and opportunities of open-source generative ai, 2024
Francisco Eiras, Aleksandar Petrov, Bertie Vidgen, Christian Schroeder, Fabio Pizzati, Katherine Elkins, Supratik Mukhopadhyay, Adel Bibi, Aaron Purewal, Csaba Botos, Fabro Steibel, Fazel Keshtkar, Fazl Barez, Genevieve Smith, Gianluca Guadagni, Jon Chun, Jordi Cabot, Joseph Imperial, Juan Arturo Nolazco, Lori Landay, Matthew Jackson, Phillip H. S. Torr, Trevor Darrell, Yong Lee, and Jakob Foerster · 2024
Closest in time.
Reconfiguring participatory design to resist ai realism
Aakash Gautam · 2024
Closest in time.
Github acceptable use policies, 2024
GitHub · 2024
Closest in time.
How persuasive is AI-generated propaganda?
Josh A Goldstein, Jason Chao, Shelby Grossman, Alex Stamos, and Michael Tomz · 2024
Closest in time.
Policy guidelines for the gemini app, 2024
Google · 2024
Closest in time.
Moderating model marketplaces: Platform governance puzzles for ai intermediaries, 2024
Robert Gorwa and Michael Veale · 2024
Closest in time.
Reclaiming artificial intelligence accounts: A plea for a participatory turn in artificial intelligence inquiries
Pauline Gourlet, Donato Ricci, and Maxime Crépel · 2024
Closest in time.
Risks from language models for automated mental healthcare: Ethics and structure for implementation
Declan Grabb, Max Lamparth, and Nina Vasan · 2024
Closest in time.
Regulating gatekeeper artificial intelligence and data: Transparency, access and fairness under the digital markets act, the general data protection regulation and beyond
Philipp Hacker, Johann Cordes, and Janina Rochon · 2024
Closest in time.
Safety risks from customizing foundation models via fine-tuning
Peter Henderson, Xiangyu Qi, Yi Zeng, Tinghao Xie, Pin-Yu Chen, Ruoxi Jia, and Prateek Mittal · 2024
Closest in time.
Exclusive: Ftc seeking details on amazon deal with ai startup adept, source says
Krystal Hu, Greg Bensinger, and Jody Godoy · 2024
Closest in time.
Collective constitutional ai: Aligning a language model with public input
Saffron Huang, Divya Siddarth, Liane Lovitt, Thomas I. Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli · 2024
Closest in time.
On the societal impact of open foundation models, 2024
Sayash Kapoor, Rishi Bommasani, Kevin Klyman, Shayne Longpre, Ashwin Ramaswami, Peter Cihon, Aspen Hopkins, Kevin Bankston, Stella Biderman, Miranda Bogen, Rumman Chowdhury, Alex Engler, Peter Henderson, Yacine Jernite, Seth Lazar, Stefano Maffulli, Alondra Nelson, Joelle Pineau, Aviya Skowron, Dawn Song, Victor Storchan, Daniel Zhang, Daniel E. Ho, Percy Liang, and Arvind Narayanan · 2024
Closest in time.
Rishabh Kaushal, Jacob van de Kerkhof, Catalina Goanta, Gerasimos Spanakis, and Adriana Iamnitchi · 2024
Closest in time.
Openvla: An open-source vision-language-action model, 2024
Moo Jin Kim, Karl Pertsch, Siddharth Karamcheti, Ted Xiao, Ashwin Balakrishna, Suraj Nair, Rafael Rafailov, Ethan Foster, Grace Lam, Pannag Sanketi, Quan Vuong, Thomas Kollar, Benjamin Burchfiel, Russ Tedrake, Dorsa Sadigh, Sergey Levine, Percy Liang, and Chelsea Finn · 2024
Closest in time.
Position: Evolving AI collectives enhance human diversity and enable self-regulation
Shiyang Lai, Yujin Potter, Junsol Kim, Richard Zhuang, Dawn Song, and James Evans · 2024
Closest in time.
Llama 3.1 405b, meta’s ai strategy, and the new, open frontier model ecosystem
Nathan Lambert · 2024
Closest in time.
Talkin’ ’bout ai generation: Copyright and the generative-ai supply chain, 2024
Katherine Lee, A. Feder Cooper, and James Grimmelmann · 2024
Closest in time.
What’s documented in ai? systematic analysis of 32k ai model cards, 2024
Weixin Liang, Nazneen Rajani, Xinyu Yang, Ezinwanne Ozoani, Eric Wu, Yiqun Chen, Daniel Scott Smith, and James Zou · 2024
Closest in time.
The responsible foundation model development cheatsheet: A review of tools & resources, 2024
Shayne Longpre, Stella Biderman, Alon Albalak, Hailey Schoelkopf, Daniel McDuff, Sayash Kapoor, Kevin Klyman, Kyle Lo, Gabriel Ilharco, Nay San, Maribeth Rauh, Aviya Skowron, Bertie Vidgen, Laura Weidinger, Arvind Narayanan, Victor Sanh, David Adelani, Percy Liang, Rishi Bommasani, Peter Henderson, Sasha Luccioni, Yacine Jernite, and Luca Soldaini · 2024
Closest in time.
A safe harbor for ai evaluation and red teaming, 2024
Shayne Longpre, Sayash Kapoor, Kevin Klyman, Ashwin Ramaswami, Rishi Bommasani, Borhane Blili-Hamelin, Yangsibo Huang, Aviya Skowron, Zheng-Xin Yong, Suhas Kotha, Yi Zeng, Weiyan Shi, Xianjun Yang, Reid Southen, Alexander Robey, Patrick Chao, Diyi Yang, Ruoxi Jia, Daniel Kang, Sandy Pentland, Arvind Narayanan, Percy Liang, and Peter Henderson · 2024
Closest in time.
Auditing gpt’s content moderation guardrails: Can chatgpt write your favorite tv show?
Yaaseen Mahomed, Charlie M. Crawford, Sanjana Gautam, Sorelle A. Friedler, and Danaë Metaxa · 2024
Closest in time.
Generative ai misuse: A taxonomy of tactics and insights from real-world data, 2024
Nahema Marchal, Rachel Xu, Rasmi Elasmar, Iason Gabriel, Beth Goldberg, and William Isaac · 2024
Closest in time.
The unbearably high cost of cutting trust & safety corners, 2024
Glenn Ellingson Matt Motyl · 2024
Closest in time.
Daniel McDuff, Tim Korjakow, Scott Cambo, Jesse Josua Benjamin, Jenny Lee, Yacine Jernite, Carlos Muñoz Ferrandis, Aaron Gokaslan, Alek Tarkowski, Joseph Lindley, A. Feder Cooper, and Danish Contractor · 2024
Closest in time.
Understanding liability risk from using health care artificial intelligence tools
Michelle M. Mello and Neel Guha · 2024
Closest in time.
Adobe’s ‘ethical’ firefly ai was trained on midjourney images
Rachel Metz and Brody Ford · 2024
Closest in time.
The Operational Risks of AI in Large-Scale Biological Attacks: Results of a Red-Team Study
Christopher A. Mouton, Caleb Lucas, and Ella Guest · 2024
Closest in time.
Basic safety requirements for generative artificial intelligence services, April 2024
National Technical Committee 260 on Cybersecurity of Standardization Administration of China (SAC/TC260) · 2024
Closest in time.
The Open Source AI Definition – draft v. 0.0.8, 2024
Open Source Initiative · 2024
Closest in time.
Model spec, May 2024
OpenAI · 2024
Closest in time.
Addressing computer-generated child sex abuse imagery: Legal framework and policy implications
Riana Pfefferkorn · 2024
Closest in time.
Computational politeness in natural language processing: A survey
Priyanshu Priya, Mauajama Firdaus, and Asif Ekbal · 2024
Closest in time.
Open problems in technical ai governance, 2024
Anka Reuel, Ben Bucknall, Stephen Casper, Tim Fist, Lisa Soder, Onni Aarne, Lewis Hammond, Lujain Ibrahim, Alan Chan, Peter Wills, Markus Anderljung, Ben Garfinkel, Lennart Heim, Andrew Trask, Gabriel Mukobi, Rylan Schaeffer, Mauricio Baker, Sara Hooker, Irene Solaiman, Alexandra Sasha Luccioni, Nitarshan Rajkumar, Nicolas Moës, Jeffrey Ladish, Neel Guha, Jessica Newman, Yoshua Bengio, Tobin South, Alex Pentland, Sanmi Koyejo, Mykel J. Kochenderfer, and Robert Trager · 2024
Closest in time.
Exploring human-llm conversations: Mental models and the originator of toxicity, 2024
Johannes Schneider, Arianna Casanova Flores, and Anne-Catherine Kranz · 2024
Closest in time.
Generative ai should be developed and deployed responsibly at every level for everyone, February 1 2024
Megan Shahi, Adam Conner, and Nicole Alvarez · 2024
Closest in time.
Ai-powered autonomous weapons risk geopolitical instability and threaten ai research, 2024
Riley Simmons-Edler, Ryan Badman, Shayne Longpre, and Kanaka Rajan · 2024
Closest in time.
Evaluating the social impact of generative ai systems in systems and society, 2024
Irene Solaiman, Zeerak Talat, William Agnew, Lama Ahmad, Dylan Baker, Su Lin Blodgett, Canyu Chen, Hal Daumé III au2, Jesse Dodge, Isabella Duan, Ellie Evans, Felix Friedrich, Avijit Ghosh, Usman Gohar, Sara Hooker, Yacine Jernite, Ria Kalluri, Alberto Lusoli, Alina Leidinger, Michelle Lin, Xiuzhu Lin, Sasha Luccioni, Jennifer Mickel, Margaret Mitchell, Jessica Newman, Anaelia Ovalle, Marie-Therese Png, Shubham Singh, Andrew Strait, Lukas Struppek, and Arjun Subramonian · 2024
Closest in time.
Risk mitigation strategies for the open foundation model value chain, July 11 2024
Madhulika Srikumar, Jiyoo Chang, and Kasia Chmielinski · 2024
Closest in time.
Netchoice, llc v. paxton
Supreme Court of the United States · 2024
Closest in time.
Participation in the age of foundation models
Harini Suresh, Emily Tseng, Meg Young, Mary Gray, Emma Pierson, and Karen Levy · 2024
Closest in time.
Safety by design for generative ai: Preventing child sexual abuse, 2024
Thorn · 2024
Closest in time.
Proposal for a regulation of the european parliament and of the council laying down harmonised rules on artificial intelligence (artificial intelligence act) and amending certain union legislative acts, 2024
European Union · 2024
Closest in time.
United states et al. v. google llc, Aug 2024
United States District Court for the District of Columbia · 2024
Closest in time.
Joint statement on competition in generative ai foundation models and ai products, Jul 2024
Margrethe Vestager, Sarah Cardell, Jonathan Kanter, and Lina M. Khan · 2024
Closest in time.
Decodingtrust: A comprehensive assessment of trustworthiness in gpt models, 2024
Boxin Wang, Weixin Chen, Hengzhi Pei, Chulin Xie, Mintong Kang, Chenhui Zhang, Chejian Xu, Zidi Xiong, Ritik Dutta, Rylan Schaeffer, Sang T. Truong, Simran Arora, Mantas Mazeika, Dan Hendrycks, Zinan Lin, Yu Cheng, Sanmi Koyejo, Dawn Song, and Bo Li · 2024
Closest in time.
Mixture-of-agents enhances large language model capabilities, 2024
Junlin Wang, Jue Wang, Ben Athiwaratkun, Ce Zhang, and James Zou · 2024
Closest in time.
Star: Sociotechnical approach to red teaming language models, 2024
Laura Weidinger, John Mellor, Bernat Guillen Pegueroles, Nahema Marchal, Ravin Kumar, Kristian Lum, Canfer Akbulut, Mark Diaz, Stevie Bergman, Mikel Rodriguez, Verena Rieser, and William Isaac · 2024
Closest in time.
Matt White, Ibrahim Haddad, Cailean Osborne, Xiao-Yang Liu Yanglet, Ahmed Abdelmonsef, and Sachin Varghese · 2024
Closest in time.
Epistemic power in ai ethics labor: Legitimizing located complaints
David Gray Widder · 2024
Closest in time.
Air-bench 2024: A safety benchmark based on risk categories from regulations and policies, 2024
Yi Zeng, Yu Yang, Andy Zhou, Jeffrey Ziwei Tan, Yuheng Tu, Yifan Mai, Kevin Klyman, Minzhou Pan, Ruoxi Jia, Dawn Song, Percy Liang, and Bo Li · 2024
Closest in time.
Ai risk categorization decoded (air 2024): From government regulations to corporate policies, 2024
Zeng and Klyman, Andy Zhou, Yu Yang, Minzhou Pan, Ruoxi Jia, Dawn Song, Percy Liang, and Bo Li · 2024
Closest in time.
The promise and perils of china’s regulation of artificial intelligence
Angela Huyue Zhang · 2024
Closest in time.
Wildchat: 1m chatgpt interaction logs in the wild, 2024
Wenting Zhao, Xiang Ren, Jack Hessel, Claire Cardie, Yejin Choi, and Yuntian Deng · 2024
Closest in time.