Corruption perception index 2022
Transparency International. 2023 · 2022
Later among the works it cites.
Taxonomy of risks posed by language models
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, John Mellor, Amelia Glaese, Myra Cheng, Borja Balle, Atoosa Kasirzadeh, Courtney Biles, Sasha Brown, Zac Kenton, Will Hawkins, Tom Stepleton, Abeba Birhane, Lisa Anne Hendricks, Laura Rimell, William Isaac, Julia Haas, Sean Legassick, Geoffrey Irving, and Iason Gabriel. 2022 · 2022
Later among the works it cites.
The ai index 2022 annual report. ai index steering committee
Daniel Zhang, Nestor Maslej, Erik Brynjolfsson, John Etchemendy, Terah Lyons, James Manyika, Helen Ngo, Juan Carlos Niebles, Michael Sellitto, Ellie Sakhaee, et al. 2022a · 2022
Later among the works it cites.
Socially meaningful transparency in data-based systems: reflections and proposals from practice
J. Bates, H. Kennedy, and I. et al. Medina Perea. 2023 · 2023
Closest in time.
The Quest for AI Sovereignty, Transparency and Accountability
Luca Belli and Walter Gaspar. 2023 · 2023
Closest in time.
Evaluation for change
Rishi Bommasani. 2023 · 2023
Closest in time.
Expert explainer: Allocating accountability in ai supply chains
Ian Brown. 2023 · 2023
Closest in time.
Open problems and fundamental limitations of reinforcement learning from human feedback
Original
Stephen Casper, Xander Davies, Claudia Shi, Thomas Krendl Gilbert, Jérémy Scheurer, Javier Rando, Rachel Freedman, Tomasz Korbak, David Lindner, Pedro Freire, Tony Wang, Samuel Marks, Charbel-Raphaël Segerie, Micah Carroll, Andi Peng, Phillip Christoffersen, Mehul Damani, Stewart Slocum, Usman Anwar, Anand Siththaranjan, Max Nadeau, Eric J. Michaud, Jacob Pfau, Dmitrii Krasheninnikov, Xin Chen, Lauro Langosco, Peter Hase, Erdem Bıyık, Anca Dragan, David Krueger, Dorsa Sadigh, and Dylan Hadfield-Menell. 2023 · 2023
Closest in time.
Ai supply chains and why they matter
Sarah H. Cen, Aspen Hopkins, Andrew Ilyas, Aleksander Madry, Isabella Struckman, and Luis Videgaray. 2023 · 2023
Closest in time.
Understanding accountability in algorithmic supply chains
Jennifer Cobbe, Michael Veale, and Jatinder Singh. 2023 · 2023
Closest in time.
Ftc finalizes order requiring fortnite maker epic games to pay 245 million for tricking users into making unwanted charges
Federal Trade Commission. 2023 · 2023
Closest in time.
Machine learning and artificial intelligence: Legal concepts
A. Feder Cooper, David Mimno, Madiha Choksi, and Katherine Lee. 2023 · 2023
Closest in time.
Interrogating the t in facct
Eric Corbett and Emily Denton. 2023 · 2023
Closest in time.
Ai is a lot of work: As the technology becomes ubiquitous, a vast tasker underclass is emerging — and not going anywhere
Josh Dzieza. 2023 · 2023
Closest in time.
Gpts are gpts: An early look at the labor market impact potential of large language models
Original
Tyna Eloundou, Sam Manning, Pamela Mishkin, and Daniel Rock. 2023 · 2023
Closest in time.
A comprehensive and distributed approach to ai regulation
Alex Engler. 2023 · 2023
Closest in time.
Smart corruption: Satirical strategies for gaming accountability
Ritwick Ghosh and Hilary Oliva Faxon. 2023 · 2023
Closest in time.
The need for transparency in artificial intelligence
Sam Gregory. 2023 · 2023
Closest in time.
Regulating chatgpt and other large generative ai models
Philipp Hacker, Andreas Engel, and Marco Mauer. 2023 · 2023
Closest in time.
Cleaning up chatgpt takes heavy toll on human workers
Karen Hao and Deepa Seetharaman. 2023 · 2023
Closest in time.
Oversight of a.i.: Legislating on artificial intelligence
Woodrow Hartzog. 2023 · 2023
Closest in time.
Version control for ml models: Why you need it, what it is, how to implement it
Ahmed Hashesh. 2023 · 2023
Closest in time.
It’s high time for more ai transparency
Melissa Heikkilä. 2023 · 2023
Closest in time.
Foundation models and fair use
Original
Peter Henderson, Xuechen Li, Dan Jurafsky, Tatsunori Hashimoto, Mark A Lemley, and Percy Liang. 2023 · 2023
Closest in time.
Catastrophic jailbreak of open-source llms via exploiting generation
Original
Yangsibo Huang, Samyak Gupta, Mengzhou Xia, Kai Li, and Danqi Chen. 2023 · 2023
Closest in time.
Explainer: What is a foundation model?
Elliot Jones. 2023 · 2023
Closest in time.
Reforms: Reporting standards for machine learning based science
Original
Sayash Kapoor, Emily Cantrell, Kenny Peng, Thanh Hien Pham, Christopher A. Bail, Odd Erik Gundersen, Jake M. Hofman, Jessica Hullman, Michael A. Lones, Momin M. Malik, Priyanka Nanayakkara, Russell A. Poldrack, Inioluwa Deborah Raji, Michael Roberts, Matthew J. Salganik, Marta Serra-Garcia, Brandon M. Stewart, Gilles Vandewiele, and Arvind Narayanan. 2023 · 2023
Closest in time.
Leakage and the reproducibility crisis in machine-learning-based science
Sayash Kapoor and Arvind Narayanan. 2023 · 2023
Closest in time.
Propile: Probing privacy leakage in large language models
Original
Siwon Kim, Sangdoo Yun, Hwaran Lee, Martin Gubri, Sungroh Yoon, and Seong Joon Oh. 2023 · 2023
Closest in time.
A watermark for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein. 2023 · 2023
Closest in time.
How to promote responsible open foundation models
Kevin Klyman. 2023 · 2023
Closest in time.
Robust distortion-free watermarks for language models
Original
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto, and Percy Liang. 2023 · 2023
Closest in time.
Geo-bench: Toward foundation models for earth monitoring
Original
Alexandre Lacoste, Nils Lehmann, Pau Rodríguez López, Evan D. Sherwin, Hannah Rae Kerner, Bjorn Lutjens, Jeremy A. Irvin, David Dao, Hamed Alemohammad, Alexandre Drouin, Mehmet Gunturkun, Gabriel Huang, David Vázquez, Dava Newman, Yoshua Bengio, Stefano Ermon, and Xiao Xiang Zhu. 2023 · 2023
Closest in time.
Towards a transparent ai future: The call for less regulatory hurdles on open-source ai in europe
LAION. 2023 · 2023
Closest in time.
Holistic evaluation of language models
Percy Liang, Rishi Bommasani, Tony Lee, Dimitris Tsipras, Dilara Soylu, Michihiro Yasunaga, Yian Zhang, Deepak Narayanan, Yuhuai Wu, Ananya Kumar, Benjamin Newman, Binhang Yuan, Bobby Yan, Ce Zhang, Christian Alexander Cosgrove, Christopher D Manning, Christopher Re, Diana Acosta-Navas, Drew Arad Hudson, Eric Zelikman, Esin Durmus, Faisal Ladhak, Frieda Rong, Hongyu Ren, Huaxiu Yao, Jue WANG, Keshav Santhanam, Laurel Orr, Lucia Zheng, Mert Yuksekgonul, Mirac Suzgun, Nathan Kim, Neel Guha, Niladri S. Chatterji, Omar Khattab, Peter Henderson, Qian Huang, Ryan Andrew Chi, Sang Michael Xie, Shibani Santurkar, Surya Ganguli, Tatsunori Hashimoto, Thomas Icard, Tianyi Zhang, Vishrav Chaudhary, William Wang, Xuechen Li, Yifan Mai, Yuhui Zhang, and Yuta Koreeda. 2023 · 2023
Closest in time.
Ai transparency in the age of llms: A human-centered research roadmap
Original
Qingzi Vera Liao and Jennifer Wortman Vaughan. 2023 · 2023
Closest in time.
Opening up chatgpt: Tracking openness, transparency, and accountability in instruction-tuned text generators
Andreas Liesenfeld, Alianda Lopez, and Mark Dingemanse. 2023 · 2023
Closest in time.
Evolutionary-scale prediction of atomic-level protein structure with a language model
Zeming Lin, Halil Akin, Roshan Rao, Brian Hie, Zhongkai Zhu, Wenting Lu, Nikita Smetanin, Robert Verkuil, Ori Kabeli, Yaniv Shmueli, Allan dos Santos Costa, Maryam Fazel-Zarandi, Tom Sercu, Salvatore Candido, and Alexander Rives. 2023 · 2023
Closest in time.
Building a competitive advantage based on transparency: When and why does transparency matter for corporate social responsibility?
Yeyi Liu, Martin Heinberg, Xuan Huang, and Andreas B. Eisingerich. 2023 · 2023
Closest in time.
A pretrainer’s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity
Original
Shayne Longpre, Gregory Yauney, Emily Reif, Katherine Lee, Adam Roberts, Barret Zoph, Denny Zhou, Jason Wei, Kevin Robinson, David Mimno, et al. 2023 · 2023
Closest in time.
Counting carbon: A survey of factors influencing the emissions of machine learning
Original
Alexandra Sasha Luccioni and Alex Hernández-García. 2023 · 2023
Closest in time.
Analyzing leakage of personally identifiable information in language models
Original
Nils Lukas, Ahmed Salem, Robert Sim, Shruti Tople, Lukas Wutschitz, and Santiago Zanella-Béguelin. 2023 · 2023
Closest in time.
The ai index 2023 annual report
Nestor Maslej, Loredana Fattorini, Erik Brynjolfsson, John Etchemendy, Katrina Ligett, Terah Lyons, James Manyika, Helen Ngo, Juan Carlos Niebles, Vanessa Parli, et al. 2023 · 2023
Closest in time.
Meta platform terms
Meta. 2023 · 2023
Closest in time.
Generative ai companies must publish transparency reports
Arvind Narayanan and Sayash Kapoor. 2023 · 2023
Closest in time.
Cheaply evaluating inference efficiency metrics for autoregressive transformer APIs
Deepak Narayanan, Keshav Santhanam, Peter Henderson, Rishi Bommasani, Tony Lee, and Percy Liang. 2023 · 2023
Closest in time.
Astrollama: Towards specialized foundation models in astronomy
Original
Tuan Dung Nguyen, Yuan-Sen Ting, Ioana Ciucă, Charlie O’Neill, Ze-Chang Sun, Maja Jablo’nska, Sandor Kruk, Ernest Perkowski, Jack W. Miller, Jason Li, Josh Peek, Kartheik Iyer, Tomasz R’o.za’nski, Pranav Khetarpal, Sharaf Zaman, David Brodrick, Sergio J. Rodr’iguez M’endez, Thang Bui, Alyssa Goodman, Alberto Accomazzi, Jill P. Naiman, Jesse Cranney, Kevin Schawinski, and UniverseTBD. 2023 · 2023
Closest in time.
Open X-Embodiment: Robotic learning datasets and RT-X models
Open X-Embodiment Collaboration, Abhishek Padalkar, Acorn Pooley, Ajinkya Jain, Alex Bewley, Alex Herzog, Alex Irpan, Alexander Khazatsky, Anant Rai, Anikait Singh, Anthony Brohan, Antonin Raffin, Ayzaan Wahid, Ben Burgess-Limerick, Beomjoon Kim, Bernhard Schölkopf, Brian Ichter, Cewu Lu, Charles Xu, Chelsea Finn, Chenfeng Xu, Cheng Chi, Chenguang Huang, Christine Chan, Chuer Pan, Chuyuan Fu, Coline Devin, Danny Driess, Deepak Pathak, Dhruv Shah, Dieter Büchler, Dmitry Kalashnikov, Dorsa Sadigh, Edward Johns, Federico Ceola, Fei Xia, Freek Stulp, Gaoyue Zhou, Gaurav S. Sukhatme, Gautam Salhotra, Ge Yan, Giulio Schiavi, Hao Su, Hao-Shu Fang, Haochen Shi, Heni Ben Amor, Henrik I Christensen, Hiroki Furuta, Homer Walke, Hongjie Fang, Igor Mordatch, Ilija Radosavovic, Isabel Leal, Jacky Liang, Jaehyung Kim, Jan Schneider, Jasmine Hsu, Jeannette Bohg, Jeffrey Bingham, Jiajun Wu, Jialin Wu, Jianlan Luo, Jiayuan Gu, Jie Tan, Jihoon Oh, Jitendra Malik, Jonathan Tompson, Jonathan Yang, Joseph J. Lim, João Silvério, Junhyek Han, Kanishka Rao, Karl Pertsch, Karol Hausman, Keegan Go, Keerthana Gopalakrishnan, Ken Goldberg, Kendra Byrne, Kenneth Oslund, Kento Kawaharazuka, Kevin Zhang, Keyvan Majd, Krishan Rana, Krishnan Srinivasan, Lawrence Yunliang Chen, Lerrel Pinto, Liam Tan, Lionel Ott, Lisa Lee, Masayoshi Tomizuka, Maximilian Du, Michael Ahn, Mingtong Zhang, Mingyu Ding, Mohan Kumar Srirama, Mohit Sharma, Moo Jin Kim, Naoaki Kanazawa, Nicklas Hansen, Nicolas Heess, Nikhil J Joshi, Niko Suenderhauf, Norman Di Palo, Nur Muhammad Mahi Shafiullah, Oier Mees, Oliver Kroemer, Pannag R Sanketi, Paul Wohlhart, Peng Xu, Pierre Sermanet, Priya Sundaresan, Quan Vuong, Rafael Rafailov, Ran Tian, Ria Doshi, Roberto Martín-Martín, Russell Mendonca, Rutav Shah, Ryan Hoque, Ryan Julian, Samuel Bustamante, Sean Kirmani, Sergey Levine, Sherry Moore, Shikhar Bahl, Shivin Dass, Shuran Song, Sichun Xu, Siddhant Haldar, Simeon Adebola, Simon Guist, Soroush Nasiriany, Stefan Schaal, Stefan Welker, Stephen Tian, Sudeep Dasari, Suneel Belkhale, Takayuki Osa, Tatsuya Harada, Tatsuya Matsushima, Ted Xiao, Tianhe Yu, Tianli Ding, Todor Davchev, Tony Z. Zhao, Travis Armstrong, Trevor Darrell, Vidhi Jain, Vincent Vanhoucke, Wei Zhan, Wenxuan Zhou, Wolfram Burgard, Xi Chen, Xiaolong Wang, Xinghao Zhu, Xuanlin Li, Yao Lu, Yevgen Chebotar, Yifan Zhou, Yifeng Zhu, Ying Xu, Yixuan Wang, Yonatan Bisk, Yoonyoung Cho, Youngwoon Lee, Yuchen Cui, Yueh hua Wu, Yujin Tang, Yuke Zhu, Yunzhu Li, Yusuke Iwasawa, Yutaka Matsuo, Zhuo Xu, and Zichen Jeff Cui. 2023 · 2023
Closest in time.
Gpt-4 technical report
Original
OpenAI. 2023 · 2023
Closest in time.
The roots search tool: Data transparency for llms
Original
Aleksandra Piktus, Christopher Akiki, Paulo Villegas, Hugo Laurençon, Gérard Dupont, Alexandra Sasha Luccioni, Yacine Jernite, and Anna Rogers. 2023 · 2023
Closest in time.
Stronger together: on the articulation of ethical charters, legal tools, and technical documentation in ML
Giada Pistilli, Carlos Muñoz Ferrandis, Yacine Jernite, and Margaret Mitchell. 2023 · 2023
Closest in time.
Fine-tuning aligned language models compromises safety, even when users do not intend to!
Original
Xiangyu Qi, Yi Zeng, Tinghao Xie, Pin-Yu Chen, Ruoxi Jia, Prateek Mittal, and Peter Henderson. 2023 · 2023
Closest in time.
I’m afraid i can’t do that: Predicting prompt refusal in black-box generative language models
Original
Max Reuter and William Schulze. 2023 · 2023
Closest in time.
Paul tremblay, mona awad vs. openai, inc., et al
Joseph R. Saveri, Cadio Zirpoli, Christopher K.L. Young, and Kathleen J. McMahon. 2023 · 2023
Closest in time.
Open-Sourcing Highly Capable Foundation Models: An Evaluation of Risks, Benefits, and Alternative Methods for Pursuing Open-Source Objectives
Elizabeth Seger, Noemi Dreksler, Richard Moulange, Emily Dardaman, Jonas Schuett, K. Wei, Christoph Winter, Mackenzie Arnold, Seán Ó hÉigeartaigh, Anton Korinek, Markus Anderljung, Ben Bucknall, Alan Chan, Eoghan Stafford, Leonie Koessler, Aviv Ovadya, Ben Garfinkel, Emma Bluemke, Michael Aird, Patrick Levermore, Julian Hazell, and Abhishek Gupta. 2023 · 2023
Closest in time.
Eleutherai’s thoughts on the eu ai act
Aviya Skowron and Stella Biderman. 2023 · 2023
Closest in time.
The gradient of generative ai release: Methods and considerations
Irene Solaiman. 2023 · 2023
Closest in time.
Evaluating the social impact of generative ai systems in systems and society
Original
Irene Solaiman, Zeerak Talat, William Agnew, Lama Ahmad, Dylan Baker, Su Lin Blodgett, Hal Daumé III au2, Jesse Dodge, Ellie Evans, Sara Hooker, Yacine Jernite, Alexandra Sasha Luccioni, Alberto Lusoli, Margaret Mitchell, Jessica Newman, Marie-Therese Png, Andrew Strait, and Apostol Vassilev. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Original
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.
Market concentration implications of foundation models: The invisible hand of chatgpt
Jai Vipra and Anton Korinek. 2023 · 2023
Closest in time.
Computational power and ai
Jai Vipra and Sarah Myers West. 2023 · 2023
Closest in time.
Comments on the business practices of cloud computing providers
Jai Vipra and Sarah Myers West. 2023 · 2023
Closest in time.
Designing responsible ai: Adaptations of ux practice to meet responsible ai challenges
Qiaosi Wang, Michael Madaio, Shaun Kane, Shivani Kapania, Michael Terry, and Lauren Wilcox. 2023b · 2023
Closest in time.
The presidio recommendations on responsible generative ai
WEF. 2023 · 2023
Closest in time.
Sociotechnical safety evaluation of generative ai systems
Original
Laura Weidinger, Maribeth Rauh, Nahema Marchal, Arianna Manzini, Lisa Anne Hendricks, Juan Mateos-Garcia, Stevie Bergman, Jackie Kay, Conor Griffin, Ben Bariach, Iason Gabriel, Verena Rieser, and William S. Isaac. 2023 · 2023
Closest in time.
Open (for business): Big tech, concentrated power, and the political economy of open ai
David Gray Widder, Sarah West, and Meredith Whittaker. 2023 · 2023
Closest in time.
Thinking upstream: Ethics and policy opportunities in ai supply chains
Original
David Gray Widder and Richmond Wong. 2023 · 2023
Closest in time.
Loose-lipped large language modells spill your secrets: The privacy implications of large language models
Amy Winograd. 2023 · 2023
Closest in time.
How data protection authorities are de facto regulating generative ai
Gabriela Zanfir-Fortuna. 2023 · 2023
Closest in time.
Representation engineering: A top-down approach to ai transparency
Original
Andy Zou, Long Phan, Sarah Chen, James Campbell, Phillip Guo, Richard Ren, Alexander Pan, Xuwang Yin, Mantas Mazeika, Ann-Kathrin Dombrowski, Shashwat Goel, Nathaniel Li, Michael J. Byun, Zifan Wang, Alex Mallen, Steven Basart, Sanmi Koyejo, Dawn Song, Matt Fredrikson, J. Zico Kolter, and Dan Hendrycks. 2023 · 2023
Closest in time.
Privacy as contextual integrity
Helen Nissenbaum. 2024 · 2024
Closest in time.