Fetching the paper…
Reading the bibliography…
Significant advancements have recently been made in large language models represented by GPT models.
On the Importance of Securing Your Bins: The Garbage-man-in-the-middle Attack
Marc Joye and Jean-Jacques Quisquater · 1997
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
The Enron Corpus: A New Dataset for Email Classification Research
Bryan Klimt and Yiming Yang · 2004
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with High Levels of Correlation with Human Judgments
Alon Lavie and Abhaya Agarwal · 2007
Earlier work this paper cites.
MultiUN: A Multilingual Corpus from United Nation Documents
Andreas Eisele and Yu Chen · 2010
Earlier work this paper cites.
Wiretapping via Mimicry: Short Voice Imitation Man-in-the-Middle Attacks on Crypto Phones
Maliheh Shirvanian and Nitesh Saxena · 2014
Earlier work this paper cites.
SQuAD: 100, 000+ Questions for Machine Comprehension of Text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang · 2016
Earlier work this paper cites.
Membership Inference Attacks Against Machine Learning Models
Reza Shokri, Marco Stronati, Congzheng Song, and Vitaly Shmatikov · 2017
Earlier work this paper cites.
Hierarchical Neural Story Generation
Angela Fan, Mike Lewis, and Yann Dauphin · 2018
Earlier work this paper cites.
Privacy- and Utility-Preserving Textual Analysis via Calibrated Multivariate Perturbations
Oluwaseyi Feyisetan, Borja Balle, Thomas Drake, and Tom Diethe · 2019
Earlier work this paper cites.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Earlier work this paper cites.
CodeSearchNet Challenge: Evaluating the State of Semantic Code Search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt · 2020
Earlier work this paper cites.
AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts
Taylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace, and Sameer Singh · 2020
Earlier work this paper cites.
Cache-in-the-Middle (CITM) Attacks: Manipulating Sensitive Data in Isolated Execution Environments
Jie Wang, Kun Sun, Lingguang Lei, Shengye Wan, Yuewu Wang, and Jiwu Jing · 2020
Earlier work this paper cites.
MedDialog: Large-scale Medical Dialogue Datasets
Guangtao Zeng, Wenmian Yang, Zeqian Ju, Yue Yang, Sicheng Wang, Ruisi Zhang, Meng Zhou, Jiaqi Zeng, Xiangyu Dong, Ruoyu Zhang, Hongchao Fang, Penghui Zhu, Shu Chen, and Pengtao Xie · 2020
Earlier work this paper cites.
Extracting Training Data from Large Language Models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, Dawn Song, Úlfar Erlingsson, Alina Oprea, and Colin Raffel · 2021
Earlier work this paper cites.
TEM: High Utility Metric Differential Privacy on Text
Ricardo Silva Carvalho, Theodore Vasiloudis, and Oluwaseyi Feyisetan · 2021
Earlier work this paper cites.
How BPE Affects Memorization in Transformers
Eugene Kharitonov, Marco Baroni, and Dieuwke Hupkes · 2021
Cited alongside, same era.
Synthetic Data Generation for Grammatical Error Correction with Tagged Corruption Models
Felix Stahlberg and Shankar Kumar · 2021
Cited alongside, same era.
Membership Inference Attacks from First Principles
Nicholas Carlini, Steve Chien, Milad Nasr, Shuang Song, Andreas Terzis, and Florian Tramèr · 2022
Cited alongside, same era.
The Limits of Word Level Differential Privacy
Justus Mattern, Benjamin Weggenmann, and Florian Kerschbaum · 2022
Cited alongside, same era.
Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2022
How Much Do Language Models Copy from Their Training Data? Evaluating Linguistic Novelty in Text Generation Using RAVEN
R. Thomas McCoy, Paul Smolensky, Tal Linzen, Jianfeng Gao, and Asli Celikyilmaz · 2023
Later among the works it cites.
https://www.llama.com/llama2/ , 2023
Meta · 2023
Later among the works it cites.
Niloofar Mireshghallah, Hyunwoo Kim, Xuhui Zhou, Yulia Tsvetkov, Maarten Sap, Reza Shokri, and Yejin Choi · 2023
Later among the works it cites.
Membership Inference Attacks With Token-Level Deduplication on Korean Language Models
Myung Gyo Oh, Leo Hyun Park, Jaeuk Kim, Jaewoo Park, and Taekyoung Kwon · 2023
Later among the works it cites.
OpenAI · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
An Empirical Analysis of Memorization in Fine-tuned Autoregressive Language Models
Fatemehsadat Mireshghallah, Archit Uniyal, Tianhao Wang, David Evans, and Taylor Berg-Kirkpatrick · 2022
Cited alongside, same era.
Ignore Previous Prompt: Attack Techniques For Language Models
Fábio Perez and Ian Ribeiro · 2022
Cited alongside, same era.
Memorization Without Overfitting: Analyzing the Training Dynamics of Large Language Models
Kushal Tirumala, Aram H. Markosyan, Luke Zettlemoyer, and Armen Aghajanyan · 2022
Cited alongside, same era.
Quantifying Memorization Across Neural Language Models
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee, Florian Tramèr, and Chiyuan Zhang · 2023
Cited alongside, same era.
Data Curation Alone Can Stabilize In-context Learning
Ting-Yun Chang and Robin Jia · 2023
Cited alongside, same era.
On the Privacy Risk of In-context Learning
Haonan Duan, Adam Dziedzic, Mohammad Yaghini, Nicolas Papernot, and Franziska Boenisch · 2023
Cited alongside, same era.
Man-in-the-Middle Attacks Without Rogue AP: When WPAs Meet ICMP Redirects
Xuewei Feng, Qi Li, Kun Sun, Yuxiang Yang, and Ke Xu · 2023
Cited alongside, same era.
Differentially Private In-Context Learning
Ashwinee Panda, Tong Wu, Jiachen T. Wang, and Prateek Mittal · 2023
Later among the works it cites.
LLaMA: Open and Efficient Foundation Language Models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurélien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample · 2023
Later among the works it cites.
Locally Differentially Private Document Generation Using Zero Shot Prompting
Saiteja Utpala, Sara Hooker, and Pin-Yu Chen · 2023
Later among the works it cites.
Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Zeming Wei, Yifei Wang, Ang Li, Yichuan Mo, and Yisen Wang · 2023
Later among the works it cites.
Defending ChatGPT against jailbreak attack via self-reminders
Yueqi Xie, Jingwei Yi, Jiawei Shao, Justin Curl, Lingjuan Lyu, Qifeng Chen, Xing Xie, and Fangzhao Wu · 2023
Later among the works it cites.
Counterfactual Memorization in Neural Language Models
Chiyuan Zhang, Daphne Ippolito, Katherine Lee, Matthew Jagielski, Florian Tramèr, and Nicholas Carlini · 2023
Later among the works it cites.
https://www.anthropic.com/news/claude-3-haiku/ , 2024
Anthropic · 2024
Closest in time.
Comprehensive Assessment of Jailbreak Attacks Against LLMs
Junjie Chu, Yugeng Liu, Ziqing Yang, Xinyue Shen, Michael Backes, and Yang Zhang · 2024
Closest in time.
https://github.com/meta-llama/llama3/ , 2024
Meta · 2024
Closest in time.
https://openai.com/blog/introducing-gpts , 2024
OpenAI · 2024
Closest in time.
https://openai.com/api/ , 2024
OpenAI · 2024
Closest in time.
Do Anything Now: Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
Xinyue Shen, Zeyuan Chen, Michael Backes, Yun Shen, and Yang Zhang · 2024
Closest in time.