Fetching the paper…
Reading the bibliography…
The rapid integration of Large Language Models (LLMs) across diverse sectors has marked a transformative era, showcasing remarkable capabilities in text generation and problem-solving tasks.
“Harmonized threat and risk assessment (TRA) methodology .”
Communications(Canada) Police · 2007
Earlier work this paper cites.
“NIST SP 800-30 Rev. 1 Guide for Conducting Risk Assessments”
NIST · 2012
Earlier work this paper cites.
“NIST Risk Management Framework (RMF)”
NIST · 2016
Earlier work this paper cites.
“Deep reinforcement learning from human preferences”
Paul Christiano et al · 2017
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“An architectural risk analysis of machine learning systems: Toward more secure machine learning”
Gary McGraw, Harold Figueroa, Victor Shepardson and Richie Bonett · 2020
Earlier work this paper cites.
“A threat analysis methodology for security requirements elicitation in machine learning based systems”
Carl Wilhjelm and Awad Younis · 2020
Earlier work this paper cites.
“Lora: Low-rank adaptation of large language models”
Edward Hu et al · 2021
Earlier work this paper cites.
“Training a helpful and harmless assistant with reinforcement learning from human feedback”
Yuntao Bai et al · 2022
Earlier work this paper cites.
“Modeling threats to AI-ML systems using STRIDE”
Lara Mauri and Ernesto Damiani · 2022
Earlier work this paper cites.
“Training language models to follow instructions with human feedback”
Long Ouyang et al · 2022
Earlier work this paper cites.
“Attacks on ML Systems: From Security Analysis to Attack Mitigation”
Qingtian Zou et al · 2022
Cited alongside, same era.
“ISO 27001:2022 Information security management systems”
ISO · 2022
Cited alongside, same era.
“Jailbreaker: Automated jailbreak across multiple large language model chatbots”
Gelei Deng et al · 2023
Cited alongside, same era.
“Bias and fairness in large language models: A survey”
Isabel Gallegos et al · 2023
Cited alongside, same era.
“More than you’ve asked for: A Comprehensive Analysis of Novel Prompt Injection Threats to Application-Integrated Large Language Models”
Kai Greshake et al · 2023
Cited alongside, same era.
“Samsung bans chatgpt among employees after sensitive code leak”
Siladitya Ray · 2023
Later among the works it cites.
“Amazon’s Q has “severe hallucinations” and leaks confidential data in public preview, employees warn”
Schiffer · 2023
Later among the works it cites.
“Achieving Code Execution in MathGPT via Prompt Injection”
Ludwig-Ferdinand Stumpp · 2023
Later among the works it cites.
“Alpaca: A strong, replicable instruction-following model”
Rohan Taori et al · 2023
Later among the works it cites.
“Llama 2: Open foundation and fine-tuned chat models”
Hugo Touvron et al · 2023
Later among the works it cites.
“Fundamental limitations of alignment in large language models”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yangsibo Huang et al · 2023
Cited alongside, same era.
Albert Jiang et al · 2023
Cited alongside, same era.
“Multi-step jailbreaking privacy attacks on chatgpt”
Haoran Li et al · 2023
Cited alongside, same era.
“ATLAS Machine Learning Threat Matrix”
MITRE · 2023
Cited alongside, same era.
“OWASP Top 10 for Large Language Model Applications”
OWASP · 2023
Cited alongside, same era.
“ENISA Risk Assessment”
ENISA
Cited in the paper.
“OWASP Risk Rating Methodology”
OWASP
Cited in the paper.
Yotam Wolf, Noam Wies, Yoav Levine and Amnon Shashua · 2023
Later among the works it cites.
Jiashu Xu et al · 2023
Later among the works it cites.
“Siren’s song in the ai ocean: A survey on hallucination in large language models”
Yue Zhang et al · 2023
Later among the works it cites.
“Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems”
Tianyu Cui et al · 2024
Closest in time.
“On the Societal Impact of Open Foundation Models”
Kapoor and Bommasani et al · 2024
Closest in time.