Fetching the paper…
Reading the bibliography…
As large language models (LLMs) generate texts with increasing fluency and realism, there is a growing need to identify the source of texts to prevent the abuse of LLMs.
A Review of Digital Watermarking Techniques for Text Documents
Zunera Jalil and Anwar M Mirza · 2009
Earlier work this paper cites.
Watermarking the Outputs of Structured Prediction with an Application in Statistical Machine Translation
Ashish Venugopal, Jakob Uszkoreit, David Talbot, Franz Josef Och, and Juri Ganitkevitch · 2011
Earlier work this paper cites.
Real or Fake? Learning to Discriminate Machine from Human-Generated Text
Anton Bakhtin, Sam Gross, Myle Ott, Yuntian Deng, Marc’Aurelio Ranzato, and Arthur Szlam · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Roberta: A Robustly Optimized Bert Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Earlier work this paper cites.
GPT-2: 1.5B Release
OpenAI · 2019
Earlier work this paper cites.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Earlier work this paper cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2019
Earlier work this paper cites.
DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 2019
Earlier work this paper cites.
Release Strategies and the Social Impacts of Language Models
Irene Solaiman, Miles Brundage, Jack Clark, Amanda Askell, Ariel Herbert-Voss, Jeff Wu, Alec Radford, Gretchen Krueger, Jong Wook Kim, Sarah Kreps, et al · 2019
Cited alongside, same era.
Provable Robust Watermarking for AI-Generated Text
Xuandong Zhao, Prabhanjan Vijendra Ananth, Lei Li, and Yu-Xiang Wang · 2019
Cited alongside, same era.
Automatic Detection of Machine-Generated Text: A Critical Survey
Ganesh Jawahar, Muhammad Abdul-Mageed, and Laks VS Lakshmanan · 2020
Cited alongside, same era.
Adversarial Watermarking Transformer: Towards Tracing Text Provenance with Data Hiding
Sahar Abdelnabi and Mario Fritz · 2021
Cited alongside, same era.
ChatGPT: Optimizing Language Models for Dialogue
OpenAI · 2022
Cited alongside, same era.
Who Wrote this Code? Watermarking for Code Generation
Taehyun Lee, Seokhee Hong, Jaewoo Ahn, Ilgee Hong, Hwaran Lee, Sangdoo Yun, Jamin Shin, and Gunhee Kim · 2023
Closest in time.
Detectgpt: Zero-Shot Machine-Generated Text Detection using Probability Curvature
Eric Mitchell, Yoonho Lee, Alexander Khazatsky, Christopher D Manning, and Chelsea Finn · 2023
Closest in time.
Sandra Mitrović, Davide Andreoletti, and Omran Ayoub · 2023
Closest in time.
New AI Classifier for Indicating AI-Written Text, 2023
OpenAI · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tracing Text Provenance via Context-Aware Lexical Substitution
Xi Yang, Jie Zhang, Kejiang Chen, Weiming Zhang, Zehua Ma, Feng Wang, and Nenghai Yu · 2022
Cited alongside, same era.
OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, Todor Mihaylov, Myle Ott, Sam Shleifer, Kurt Shuster, Daniel Simig, Punit Singh Koura, Anjali Sridhar, Tianlu Wang, and Luke Zettlemoyer · 2022
Cited alongside, same era.
Three Bricks to Consolidate Watermarks for Large Language Models
Pierre Fernandez, Antoine Chaffin, Karim Tit, Vivien Chappelier, and Teddy Furon · 2023
Cited alongside, same era.
Ryuto Koike, Masahiro Kaneko, and Naoaki Okazaki · 2023
Cited alongside, same era.
A Watermark for Large Language Models
John Kirchen., Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein
Cited in the paper.
On the Reliability of Watermarks for Large Language Models
John Kirchen., Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum, and Tom Goldstein
Cited in the paper.
Is ChatGPT Involved in Texts? Measure the Polish Ratio to Detect ChatGPT-Generated Text
Lingyi Yang, Feng Jiang, and Haizhou Li
Cited in the paper.
Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian, Wenxiao Wang, and Soheil Feizi · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
LLMDet: A Large Language Models Detection Tool
Kangxi Wu, Liang Pang, Huawei Shen, Xueqi Cheng, and Tat-Seng Chua · 2023
Closest in time.
Robust Multi-bit Natural Language Watermarking through Invariant Features
KiYoon Yoo, Wonhyuk Ahn, Jiho Jang, and Nojun Kwak · 2092
Closest in time.