Fetching the paper…
Reading the bibliography…
Applications of Generative AI (Gen AI) are expected to revolutionize a number of different areas, ranging from science & medicine to education.
Biological weapons convention
United Nations. 1975 · 1975
Earlier work this paper cites.
Reusing open-source software and practices: The impact of open-source on commercial vendors
Alan W Brown and Grady Booch. 2002 · 2002
Earlier work this paper cites.
An empirical study of open-source and closed-source software products
James W Paulson, Giancarlo Succi, and Armin Eberlein. 2004 · 2004
Earlier work this paper cites.
United nations security council resolution 1540 (2004)
United Nations. 2004 · 2004
Earlier work this paper cites.
Open source drives innovation
Christof Ebert. 2007 · 2007
Earlier work this paper cites.
The open source software phenomenon: Characteristics that promote research
Georg Von Krogh and Sebastian Spaeth. 2007 · 2007
Earlier work this paper cites.
The impact of open source software on the strategic choices of firms developing proprietary software
Jeevan Jaisingh, Eric WK See-To, and Kar Yan Tam. 2008 · 2008
Earlier work this paper cites.
Open source vs. closed source software: towards measuring security
Guido Schryen and Rouven Kadura. 2009 · 2009
Earlier work this paper cites.
Adoption of free/libre open source software in public organizations: factors of impact
Bruno Rossi, Barbara Russo, and Giancarlo Succi. 2012 · 2012
Earlier work this paper cites.
The wealthiest mafia in the world is undergoing a schism and it could get ugly
Business Insider. 2015 · 2015
Earlier work this paper cites.
Got $90,000? a windows 0-day could be yours
Krebs. 2016 · 2016
Earlier work this paper cites.
Rand study examines 200 real-world ’zero-day’ software vulnerabilities
RAND. 2017 · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Scalable agent alignment via reward modeling: a research direction
Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini, and Shane Legg. 2018 · 2018
Earlier work this paper cites.
On the measure of intelligence
François Chollet. 2019 · 2019
Earlier work this paper cites.
Incomplete contracting and ai alignment
Dylan Hadfield-Menell and Gillian K Hadfield. 2019 · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et al. 2019 · 2019
Earlier work this paper cites.
Explaining explanations in AI
Brent Mittelstadt, Chris Russell, and Sandra Wachter. 2019 · 2019
Earlier work this paper cites.
Nih guidelines for research involving recombinant or synthetic nucleic acid molecules
National Institutes of Health. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Earlier work this paper cites.
Release strategies and the social impacts of language models
Irene Solaiman, Miles Brundage, Jack Clark, Amanda Askell, Ariel Herbert-Voss, Jeff Wu, Alec Radford, Gretchen Krueger, Jong Wook Kim, Sarah Kreps, et al. 2019 · 2019
Earlier work this paper cites.
Energy and policy considerations for deep learning in NLP
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 2020
Earlier work this paper cites.
Biosafety in microbiological and biomedical laboratories (BMBL) 6th edition
Centers for Disease Control and Prevention. 2020 · 2020
Earlier work this paper cites.
Artificial intelligence, values, and alignment
Iason Gabriel. 2020 · 2020
Earlier work this paper cites.
The Pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al. 2020 · 2020
Earlier work this paper cites.
Improving biosecurity with first international standard for biorisk management iso 35001
International Organization for Standardization. 2020 · 2020
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the NLP world
Pratik Joshi, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2020
Earlier work this paper cites.
Green ai
Roy Schwartz, Jesse Dodge, Noah A. Smith, and Oren Etzioni. 2020 · 2020
Earlier work this paper cites.
Classification of global catastrophic risks connected with artificial intelligence
Alexey Turchin and David Denkenberger. 2020 · 2020
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2020 · 2020
Earlier work this paper cites.
AppleNeuralHash2ONNX: Reverse-engineered Apple NeuralHash, in ONNX and Python
AsuharietYgvar. 2021 · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
The impact of Open Source Software and Hardware on technological independence, competitiveness and innovation in the EU economy
Knut Blind, Mirko Böhm, Paula Grzegorzewska, Andrew Katz, Sachiko Muto, Sivan Pätsch, and Torben Schubert. 2021 · 2021
Earlier work this paper cites.
On the Opportunities and Risks of Foundation Models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Earlier work this paper cites.
Cis controls v8
Critical Security Controls. 2021 · 2021
Earlier work this paper cites.
Artificial Intelligence Act
European Parliament. 2021 · 2021
Earlier work this paper cites.
AI For Public Good: Open-Source Is Not Enough
Futurium. 2021 · 2021
Earlier work this paper cites.
Datasheets for datasets
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé Iii, and Kate Crawford. 2021 · 2021
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
Alignment of language agents
Zachary Kenton, Tom Everitt, Laura Weidinger, Iason Gabriel, Vladimir Mikulik, and Geoffrey Irving. 2021 · 2021
Earlier work this paper cites.
Global health security capacity and capability measurement framework within the biological threat reduction program
Nino Kharaishvili, Jane Blake, Douglas Gorsline, and Lance Brooks. 2021 · 2021
Earlier work this paper cites.
WHO laboratory biosafety manual (lbm) 4th edition
World Health Organization. 2021 · 2021
Earlier work this paper cites.
Large image datasets: A pyrrhic win for computer vision?
Vinay Uday Prabhu and Abeba Birhane. 2021 · 2021
Earlier work this paper cites.
The digital economy runs on open source. here’s how to protect it
Harvard Business Review. 2021 · 2021
Earlier work this paper cites.
Two contrasting data annotation paradigms for subjective nlp tasks
Paul Röttger, Bertie Vidgen, Dirk Hovy, and Janet B Pierrehumbert. 2021 · 2021
Earlier work this paper cites.
Beyond “release” vs.“not release”
Girish Sastry. 2021 · 2021
Earlier work this paper cites.
Functorial language models
Alexis Toumi and Alex Koziell-Pipe. 2021 · 2021
Earlier work this paper cites.
GPT-J-6B: A 6 billion parameter autoregressive language model
Ben Wang and Aran Komatsuzaki. 2021 · 2021
Earlier work this paper cites.
Civil society can help ensure ai benefits us all. here’s how
World Economic Forum. 2021 · 2021
Earlier work this paper cites.
ACT-1: Transformer for actions
AdeptTeam. 2022 · 2022
Earlier work this paper cites.
Quantifying memorization across neural language models
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee, Florian Tramer, and Chiyuan Zhang. 2022 · 2022
Earlier work this paper cites.
The transformative potential of artificial intelligence
Ross Gruetzemacher and Jess Whittlestone. 2022 · 2022
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Earlier work this paper cites.
Iso/iec 27001:2022 information security, cybersecurity and privacy protection
International Organization for Standardization. 2022 · 2022
Earlier work this paper cites.
The time is now to develop community norms for the release of foundation models
Percy Liang, Rishi Bommasani, Kathleen Creel, and Rob Reich. 2022 · 2022
Earlier work this paper cites.
Mitigating covertly unsafe text within natural language systems
Alex Mei, Anisha Kabir, Sharon Levy, Melanie Subbiah, Emily Allaway, John Judge, Desmond Patton, Bruce Bimber, Kathleen McKeown, and William Yang Wang. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Earlier work this paper cites.
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Jack W. Rae, Sebastian Borgeaud, Trevor Cai, Katie Millican, Jordan Hoffmann, Francis Song, John Aslanides, Sarah Henderson, Roman Ring, Susannah Young, et al. 2022 · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Earlier work this paper cites.
The Digital Government Authority issues free and open-source government software licenses to 6 government agencies
Saudi Arabia Digital Government Authority. 2022 · 2022
Earlier work this paper cites.
LAION-5B: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, et al. 2022 · 2022
Earlier work this paper cites.
Using deepspeed and megatron to train megatron-turing nlg 530b, a large-scale generative language model
Shaden Smith, Mostofa Patwary, Brandon Norick, Patrick LeGresley, Samyam Rajbhandari, Jared Casper, Zhun Liu, Shrimai Prabhumoye, George Zerveas, Vijay Korthikanti, et al. 2022 · 2022
Earlier work this paper cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, et al. 2022 · 2022
Earlier work this paper cites.
The impact of artificial intelligence on the future of workforces in the european union and the united states of america
The White House. 2022 · 2022
Earlier work this paper cites.
Sustainable AI: Environmental implications, challenges and opportunities
Carole-Jean Wu, Ramya Raghavendra, Udit Gupta, Bilge Acun, Newsha Ardalani, Kiwan Maeng, Gloria Chang, Fiona Aga, Jinshi Huang, Charles Bai, et al. 2022 · 2022
Earlier work this paper cites.
A systematic evaluation of large language models of code
Frank F. Xu, Uri Alon, Graham Neubig, and Vincent J. Hellendoorn. 2022 · 2022
Earlier work this paper cites.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2022 · 2022
Earlier work this paper cites.
GPT-4 technical report
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Earlier work this paper cites.
Do all languages cost the same? Tokenization in the era of commercial language models
Orevaoghene Ahia, Sachin Kumar, Hila Gonen, Jungo Kasai, David R Mortensen, Noah A Smith, and Yulia Tsvetkov. 2023 · 2023
Earlier work this paper cites.
The impact of large language models on scientific discovery: a preliminary study using GPT-4
Microsoft Research AI4Science and Microsoft Azure Quantum. 2023 · 2023
Earlier work this paper cites.
Coordinated pausing: An evaluation-based coordination scheme for frontier AI developers
Jide Alaga and Jonas Schuett. 2023 · 2023
Earlier work this paper cites.
Potential impact of large language models on academic writing
Fares Alahdab. 2023 · 2023
Earlier work this paper cites.
The falcon series of open language models
Ebtesam Almazrouei, Hamza Alobeidli, Abdulaziz Alshamsi, Alessandro Cappelli, Ruxandra Cojocaru, Mérouane Debbah, Étienne Goffinet, Daniel Hesslow, Julien Launay, Quentin Malartic, et al. 2023 · 2023
Earlier work this paper cites.
AWS expands Amazon Bedrock with additional foundation models, new model provider, and advanced capability to help customers build generative AI applications
Amazon. 2023 · 2023
Earlier work this paper cites.
Palm 2 technical report
Rohan Anil, Andrew M Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Siamak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, et al. 2023 · 2023
Earlier work this paper cites.
Anthropic’s responsible scaling policy
Anthropic. 2023 · 2023
Earlier work this paper cites.
DICES Dataset: Diversity in conversational AI evaluation for safety
Lora Aroyo, Alex S. Taylor, Mark Diaz, Christopher M. Homan, Alicia Parrish, Greg Serapio-Garcia, Vinodkumar Prabhakaran, and Ding Wang. 2023 · 2023
Earlier work this paper cites.
Ibm, meta form “ai alliance” with 50 organizations to promote open source ai
ArsTechnica. 2023 · 2023
Earlier work this paper cites.
The social impact of generative ai: An analysis on chatgpt
Maria Teresa Baldassarre, Danilo Caivano, Berenice Fernandez Nieto, Domenico Gigante, and Azzurra Ragone. 2023 · 2023
Cited alongside, same era.
The Belebele benchmark: a parallel reading comprehension dataset in 122 language variants
Lucas Bandarkar, Davis Liang, Benjamin Muller, Mikel Artetxe, Satya Narayan Shukla, Donald Husa, Naman Goyal, Abhinandan Krishnan, Luke Zettlemoyer, and Madian Khabsa. 2023 · 2023
Cited alongside, same era.
Safety-Tuned LLaMAs: Lessons from improving the safety of large language models that follow instructions
Federico Bianchi, Mirac Suzgun, Giuseppe Attanasio, Paul Röttger, Dan Jurafsky, Tatsunori Hashimoto, and James Zou. 2023 · 2023
Cited alongside, same era.
Science in the age of large language models
Abeba Birhane, Atoosa Kasirzadeh, David Leslie, and Sandra Wachter. 2023 · 2023
Cited alongside, same era.
Humans are biased. generative ai is even worse
Bloomberg. 2023 · 2023
Cited alongside, same era.
Nlp evaluation in trouble: On the need to measure llm data contamination for each benchmark
Oscar Sainz, Jon Ander Campos, Iker García-Ferrero, Julen Etxaniz, Oier Lopez de Lacalle, and Eneko Agirre. 2023 · 2023
Later among the works it cites.
Generative ai meets copyright
Pamela Samuelson. 2023 · 2023
Later among the works it cites.
AI ethics principles
Saudi Data and AI Authority. 2023 · 2023
Later among the works it cites.
LLM litigation
Joseph Saveri and Matthew Butterick. 2023 · 2023
Later among the works it cites.
Safe Latent Diffusion: Mitigating inappropriate degeneration in diffusion models
Patrick Schramowski, Manuel Brack, Björn Deiseroth, and Kristian Kersting. 2023 · 2023
Later among the works it cites.
A computer scientist breaks down generative AI’s hefty carbon footprint
Scientific American. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mitigating inappropriateness in image generation: Can there be value in reflecting the world’s ugliness?
Manuel Brack, Felix Friedrich, Patrick Schramowski, and Kristian Kersting. 2023 · 2023
Cited alongside, same era.
Chemcrow: Augmenting large-language models with chemistry tools
Andres M Bran, Sam Cox, Oliver Schilter, Carlo Baldassari, Andrew D White, and Philippe Schwaller. 2023 · 2023
Cited alongside, same era.
Generative AI at work
Erik Brynjolfsson, Danielle Li, and Lindsey R Raymond. 2023 · 2023
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al. 2023 · 2023
Cited alongside, same era.
China’s AI Regulations and How They Get Made
Carnegie Endowment for International Peace. 2023 · 2023
Cited alongside, same era.
Hazards from Increasingly Accessible Fine-Tuning of Downloadable Foundation Models
Alan Chan, Ben Bucknall, Herbie Bradley, and David Krueger. 2023 · 2023
Cited alongside, same era.
Jailbreaking black box large language models in twenty queries
Patrick Chao, Alexander Robey, Edgar Dobriban, Hamed Hassani, George J Pappas, and Eric Wong. 2023 · 2023
Cited alongside, same era.
Elizabeth Seger, Noemi Dreksler, Richard Moulange, Emily Dardaman, Jonas Schuett, K Wei, Christoph Winter, Mackenzie Arnold, Seán Ó hÉigeartaigh, Anton Korinek, et al. 2023 · 2023
Later among the works it cites.
Google “we have no moat, and neither does openai.”
SemiAnalysis. 2023 · 2023
Later among the works it cites.
The bias amplification paradox in text-to-image generation
Preethi Seshadri, Sameer Singh, and Yanai Elazar. 2023 · 2023
Later among the works it cites.
Model evaluation for extreme risks
Toby Shevlane, Sebastian Farquhar, Ben Garfinkel, Mary Phuong, Jess Whittlestone, Jade Leung, Daniel Kokotajlo, Nahema Marchal, Markus Anderljung, Noam Kolt, et al. 2023 · 2023
Later among the works it cites.
Building open-source AI
Yash Raj Shrestha, Georg von Krogh, and Stefan Feuerriegel. 2023 · 2023
Later among the works it cites.
Evaluating the social impact of generative ai systems in systems and society
Irene Solaiman, Zeerak Talat, William Agnew, Lama Ahmad, Dylan Baker, Su Lin Blodgett, Hal Daumé III, Jesse Dodge, Ellie Evans, Sara Hooker, et al. 2023 · 2023
Later among the works it cites.
Why open-source generative ai models are an ethical way forward for science
Arthur Spirling. 2023 · 2023
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R. Brown, et al. 2023 · 2023
Later among the works it cites.
The Janus effect of generative ai: Charting the path for responsible conduct of scholarly activities in information systems
Anjana Susarla, Ram Gopal, Jason Bennett Thatcher, and Suprateek Sarker. 2023 · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al. 2023 · 2023
Later among the works it cites.
Elon Musk’s new AI bot will help you make cocaine which proves it’s ‘based’ and ‘rebellious’
The Independent. 2023 · 2023
Later among the works it cites.
Meta’s free AI isn’t cheap to use, companies say
The Information. 2023 · 2023
Later among the works it cites.
A.i. poses “risk of extinction,” industry leaders warn
The New York Times. 2023 · 2023
Later among the works it cites.
UAE Strategy for Artificial Intelligence
The UAE Government. 2023 · 2023
Later among the works it cites.
The Bletchley Declaration by Countries Attending the AI Safety Summit, 1-2 November 2023
The UK Government. 2023a · 2023
Later among the works it cites.
FACT SHEET: President Biden Issues Executive Order on Safe, Secure, and Trustworthy Artificial Intelligence
The White House. 2023 · 2023
Later among the works it cites.
RedPajama: an Open Dataset for Training Large Language Models
Together Computer. 2023 · 2023
Later among the works it cites.
Ethical implications of large language models a multidimensional exploration of societal, economic, and technical concerns
Kassym-Jomart Tokayev. 2023 · 2023
Later among the works it cites.
Hype vs. reality: AI in the cybercriminal underground
Trendmicro. 2023 · 2023
Later among the works it cites.
U.s. oversight of laboratory biosafety and biosecurity: Current policies, recommended reforms, and options for congress
US Congress. 2023 · 2023
Later among the works it cites.
A systematic review of Green AI
Roberto Verdecchia, June Sallou, and Luís Cruz. 2023 · 2023
Later among the works it cites.
SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Bertie Vidgen, Hannah Rose Kirk, Rebecca Qian, Nino Scherrer, Anand Kannappan, Scott A Hale, and Paul Röttger. 2023 · 2023
Later among the works it cites.
Sociotechnical Safety Evaluation of Generative AI Systems
Laura Weidinger, Maribeth Rauh, Nahema Marchal, Arianna Manzini, Lisa Anne Hendricks, Juan Mateos-Garcia, Stevie Bergman, Jackie Kay, Conor Griffin, Ben Bariach, et al. 2023 · 2023
Later among the works it cites.
Open (for business): Big tech, concentrated power, and the political economy of open ai
David Gray Widder, Sarah West, and Meredith Whittaker. 2023 · 2023
Later among the works it cites.
Fundamental limitations of alignment in large language models
Yotam Wolf, Noam Wies, Yoav Levine, and Amnon Shashua. 2023 · 2023
Later among the works it cites.
LLMDet: A third party large language models generated text detection tool
Kangxi Wu, Liang Pang, Huawei Shen, Xueqi Cheng, and Tat-Seng Chua. 2023 · 2023
Later among the works it cites.
Low-resource languages jailbreak GPT-4
Zheng-Xin Yong, Cristina Menghini, and Stephen H Bach. 2023 · 2023
Later among the works it cites.
GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
Youliang Yuan, Wenxiang Jiao, Wenxuan Wang, Jen tse Huang, Pinjia He, Shuming Shi, and Zhaopeng Tu. 2023 · 2023
Later among the works it cites.
To generate or not? safety-driven unlearned diffusion models are still easy to generate unsafe images… for now
Yimeng Zhang, Jinghan Jia, Xin Chen, Aochuan Chen, Yihua Zhang, Jiancheng Liu, Ke Ding, and Sijia Liu. 2023 · 2023
Later among the works it cites.
Universal and transferable adversarial attacks on aligned language models
Andy Zou, Zifan Wang, J Zico Kolter, and Matt Fredrikson. 2023 · 2023
Later among the works it cites.
Mesa de diálogo “Inteligencia Artificial: oportunidades y desafíos de una estrategia nacional”
Agencia de Gobierno Uruguay. 2024 · 2024
Closest in time.
China’s Emerging Approach to Regulating General-Purpose Artificial Intelligence: Balancing Innovation and Control
Asia Society. 2024 · 2024
Closest in time.
First of its kind Generative AI Evaluation Sandbox for Trusted AI by AI Verify Foundation and IMDA
Infocomm Media Development Authority. 2024 · 2024
Closest in time.
Are models biased on text without gender-related language?
Catarina G Belém, Preethi Seshadri, Yasaman Razeghi, and Sameer Singh. 2024 · 2024
Closest in time.
Aligning robot and human representations
Andreea Bobu, Andi Peng, Pulkit Agrawal, Julie Shah, and Anca D Dragan. 2024 · 2024
Closest in time.
Deep utopia: Life and meaning in a solved world
Nick Bostrom. 2024 · 2024
Closest in time.
Packaging and transport of nuclear substances
Canada Government. 2024 · 2024
Closest in time.
Black-box access is insufficient for rigorous AI audits
Stephen Casper, Carson Ezell, Charlotte Siegmann, Noam Kolt, Taylor Lynn Curtis, Benjamin Bucknall, Andreas Haupt, Kevin Wei, Jérémy Scheurer, Marius Hobbhahn, et al. 2024 · 2024
Closest in time.
Visibility into AI agents
Alan Chan, Carson Ezell, Max Kaufmann, Kevin Wei, Lewis Hammond, Herbie Bradley, Emma Bluemke, Nitarshan Rajkumar, David Krueger, Noam Kolt, Lennart Heim, and Markus Anderljung. 2024 · 2024
Closest in time.
Chatgpt’s one-year anniversary: Are open-source large language models catching up?
Hailin Chen, Fangkai Jiao, Xingxuan Li, Chengwei Qin, Mathieu Ravaut, Ruochen Zhao, Caiming Xiong, and Shafiq Joty. 2024 · 2024
Closest in time.
Artículo: Ministerio De Ciencia Abre Consulta Ciudadana Para Actualizar Política Nacional De Inteligencia Artificial
Conocimiento e Innovación Chile Ministerio de Ciencia, Tecnología. 2024 · 2024
Closest in time.
Informed ai regulation: Comparing the ethical frameworks of leading llm chatbots using an ethics-based audit to assess moral reasoning and normative values
Jon Chun and Katherine Elkins. 2024 · 2024
Closest in time.
Guidelines for use of generative artificial intelligence in Courts and Tribunals
Courts of New Zealand. 2024 · 2024
Closest in time.
Biosecurity risk assessment for the use of artificial intelligence in synthetic biology
Leyma P De Haro. 2024 · 2024
Closest in time.
Multilingual jailbreak challenges in large language models
Yue Deng, Wenxuan Zhang, Sinno Jialin Pan, and Lidong Bing. 2024 · 2024
Closest in time.
Do membership inference attacks work on large language models?
Michael Duan, Anshuman Suri, Niloofar Mireshghallah, Sewon Min, Weijia Shi, Luke Zettlemoyer, Yulia Tsvetkov, Yejin Choi, David Evans, and Hannaneh Hajishirzi. 2024 · 2024
Closest in time.
Near to mid-term risks and opportunities of open source generative ai
Francisco Eiras, Aleksandar Petrov, Bertie Vidgen, Christian Schroeder de Witt, Fabio Pizzati, Katherine Elkins, Supratik Mukhopadhyay, Adel Bibi, Botos Csaba, Fabro Steibel, et al. 2024 · 2024
Closest in time.
EU AI Act
EU Parliament. 2023 · 2024
Closest in time.
GenAI against humanity: Nefarious applications of generative artificial intelligence and large language models
Emilio Ferrara. 2024 · 2024
Closest in time.
AI will transform the global economy. Let’s make sure it benefits humanity
International Monetary Fund. 2024 · 2024
Closest in time.
Mechanistically analyzing the effects of fine-tuning on procedurally defined tasks
Samyak Jain, Robert Kirk, Ekdeep Singh Lubana, Robert P. Dick, Hidenori Tanaka, Edward Grefenstette, Tim Rocktäschel, and David Scott Krueger. 2024 · 2024
Closest in time.
Ai regulation in india: Current state and future perspectives
Rahul Kapoor, Shokoh H Yaghoubi, and Theresa T Kalathil. 2024 · 2024
Closest in time.
Real-time zero-day intrusion detection system for automotive controller area network on fpgas
Shashwat Khandelwal Khandelwal and Shreejith Shanker. 2024 · 2024
Closest in time.
A call to protect open source AI in europe
LAION.ai. 2023 · 2024
Closest in time.
Holistic evaluation of language models
Percy Liang, Rishi Bommasani, Tony Lee, Dimitris Tsipras, Dilara Soylu, Michihiro Yasunaga, Yian Zhang, Deepak Narayanan, Yuhuai Wu, Ananya Kumar, et al. 2024 · 2024
Closest in time.
MAS Partners Industry to Develop Generative AI Risk Framework for the Financial Sector
Monetary Authority of Singapore. 2024 · 2024
Closest in time.
Nist cybersecurity framework
National Institute of Standards and Technology. 2024 · 2024
Closest in time.
OECD’s live repository of AI strategies & policies
OECD. 2024 · 2024
Closest in time.
Building an early warning system for llm-aided biological threat creation
OpenAI. 2024 · 2024
Closest in time.
Prompting a pretrained transformer can be a universal approximator
Aleksandar Petrov, Philip HS Torr, and Adel Bibi. 2024 · 2024
Closest in time.
Overview of ‘The Executive Order on the Safe, Secure, and Trustworthy Development and Use of Artificial Intelligence’
PwC. 2024 · 2024
Closest in time.
The societal impacts of generative artificial intelligence: A balanced perspective
Rajiv Sabherwal and Varun Grover. 2024 · 2024
Closest in time.
SDAIA launches ALLAM AI application for Arabic chat
Saudi Gazette. 2024 · 2024
Closest in time.
The Gradient of Generative AI Release: Methods and Considerations
Irene Solaiman. 2024 · 2024
Closest in time.
Trustllm: Trustworthiness in large language models
Lichao Sun, Yue Huang, Haoran Wang, Siyuan Wu, Qihui Zhang, Chujie Gao, Yixin Huang, Wenhan Lyu, Yixuan Zhang, Xiner Li, et al. 2024 · 2024
Closest in time.
Openai says new york times lawsuit against it is “without merit”
The New York Times. 2024 · 2024
Closest in time.
Nrc regulations title 10, code of federal regulations
The US Government. 2024 · 2024
Closest in time.
Introducing v0. 5 of the ai safety benchmark from mlcommons
Bertie Vidgen, Adarsh Agrawal, Ahmed M Ahmed, Victor Akinwande, Namir Al-Nuaimi, Najla Alfaraj, Elie Alhajjar, Lora Aroyo, Trupti Bavalatti, Borhane Blili-Hamelin, et al. 2024 · 2024
Closest in time.
Talking existential risk into being: a habermasian critical discourse perspective to ai hype
Salla Westerstrand, Rauli Westerstrand, and Jani Koskinen. 2024 · 2024
Closest in time.
On the out-of-distribution generalization of multimodal large language models
Xingxuan Zhang, Jiansheng Li Li, Wenjing Chu, Junjia Hai, Renzhe Xu, Yuqing Yang, Shikai Guan, Jiazheng Xu, and Peng Cui. 2024 · 2024
Closest in time.
Lima: Less is more for alignment
Chunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer, Jiao Sun, Yuning Mao, Xuezhe Ma, Avia Efrat, Ping Yu, Lili Yu, et al. 2024 · 2024
Closest in time.
Safety and security risks of generative artificial intelligence to 2025
The UK Government. 2023b · 2025
Closest in time.