Fetching the paper…
Reading the bibliography…
The proliferation of large language models (LLMs) in generating content raises concerns about text copyright.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Natural language watermarking: Design, analysis, and a proof-of-concept implementation
Mikhail J Atallah, Victor Raskin, Michael Crogan, Christian Hempelmann, Florian Kerschbaum, Dina Mohamed, and Sanket Naik. 2001 · 2001
Earlier work this paper cites.
Natural language watermarking via morphosyntactic alterations
Hasan Mesut Meral, Bülent Sankur, A Sumru Özsoy, Tunga Güngör, and Emre Sevinç. 2009 · 2009
Earlier work this paper cites.
An overview of attacks against digital watermarking and their respective countermeasures
Maryam Tanha, Seyed Dawood Sajjadi Torshizi, Mohd Taufik Abdullah, and Fazirulhisyam Hashim. 2012 · 2012
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Earlier work this paper cites.
Adversarial watermarking transformer: Towards tracing text provenance with data hiding
Sahar Abdelnabi and Mario Fritz. 2020 · 2021
Earlier work this paper cites.
Dissimilar: Towards fake news detection using information hiding, signal processing and machine learning
David Megías, Minoru Kuribayashi, Andrea Rosales, and Wojciech Mazurczyk. 2021 · 2021
Earlier work this paper cites.
Coprotector: Protect open-source code against unauthorized training usage with data poisoning
Zhensu Sun, Xiaoning Du, Fu Song, Mingze Ni, and Li Li. 2022 · 2022
Earlier work this paper cites.
Tracing text provenance via context-aware lexical substitution
Xi Yang, Jie Zhang, Kejiang Chen, Weiming Zhang, Zehua Ma, Feng Wang, and Nenghai Yu. 2022 · 2022
Earlier work this paper cites.
Opt: Open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Earlier work this paper cites.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al. 2023 · 2023
Cited alongside, same era.
Undetectable watermarks for language models
Miranda Christ, Sam Gunn, and Or Zamir. 2023 · 2023
Cited alongside, same era.
Unbiased watermark for large language models
Zhengmian Hu, Lichang Chen, Xidong Wu, Yihan Wu, Hongyang Zhang, and Heng Huang. 2023 · 2023
Cited alongside, same era.
Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense
Kalpesh Krishna, Yixiao Song, Marzena Karpinska, John Wieting, and Mohit Iyyer. 2023 · 2023
Cited alongside, same era.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Waterbench: Towards holistic evaluation of watermarks for large language models
Shangqing Tu, Yuliang Sun, Yushi Bai, Jifan Yu, Lei Hou, and Juanzi Li. 2023 · 2023
Later among the works it cites.
Advancing beyond identification: Multi-bit watermark for language models
KiYoon Yoo, Wonhyuk Ahn, and Nojun Kwak. 2023 · 2023
Later among the works it cites.
Provable robust watermarking for ai-generated text
Xuandong Zhao, Prabhanjan Ananth, Lei Li, and Yu-Xiang Wang. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto, and Percy Liang. 2023 · 2023
Cited alongside, same era.
Adversarial attack for robust watermark protection against inpainting-based and blind watermark removers
Mingzhi Lyu, Yi Huang, and Adams Wai-Kin Kong. 2023 · 2023
Cited alongside, same era.
Travis Munyer and Xin Zhong. 2023 · 2023
Cited alongside, same era.
Embarrassingly simple text watermarks
Ryoma Sato, Yuki Takezawa, Han Bao, Kenta Niwa, and Makoto Yamada. 2023 · 2023
Cited alongside, same era.
Codemark: Imperceptible watermarking for code datasets against neural code completion models
Zhensu Sun, Xiaoning Du, Fu Song, and Li Li. 2023 · 2023
Cited alongside, same era.
Did you train on my dataset? towards public dataset protection with cleanlabel backdoor watermarking
Ruixiang Tang, Qizhang Feng, Ninghao Liu, Fan Yang, and Xia Hu. 2023 · 2023
Cited alongside, same era.
A watermark for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein. 2023a
Cited in the paper.
On the reliability of watermarks for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum, and Tom Goldstein. 2023b
Cited in the paper.
Massieh Kordi Boroujeny, Ya Jiang, Kai Zeng, and Brian Mark. 2024 · 2024
Closest in time.
Gumbelsoft: Diversified language model watermarking via the gumbelmax-trick
Jiayi Fu, Xuandong Zhao, Ruihan Yang, Yuansen Zhang, Jiangjie Chen, and Yanghua Xiao. 2024 · 2024
Closest in time.
Zhiwei He, Binglin Zhou, Hongkun Hao, Aiwei Liu, Xing Wang, Zhaopeng Tu, Zhuosheng Zhang, and Rui Wang. 2024 · 2024
Closest in time.
Zero-shot generative linguistic steganography
Ke Lin, Yiyang Luo, Zijian Zhang, and Luo Ping. 2024 · 2024
Closest in time.
A survey of text watermarking in the era of large language models
Aiwei Liu, Leyi Pan, Yijian Lu, Jingjing Li, Xuming Hu, Xi Zhang, Lijie Wen, Irwin King, Hui Xiong, and Philip S. Yu. 2024 · 2024
Closest in time.
Duwak: Dual watermarks in large language models
Chaoyi Zhu, Jeroen Galjaard, Pin-Yu Chen, and Lydia Y Chen. 2024 · 2024
Closest in time.