Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated remarkable capabilities of generating texts resembling human language.
On a class of error correcting binary group codes
Raj Chandra Bose and Dwijendra K Chaudhuri · 1960
Earlier work this paper cites.
Polynomial codes over certain finite fields
Irving S Reed and Gustave Solomon · 1960
Earlier work this paper cites.
Maximum distance q-nary codes
Richard Singleton · 1964
Earlier work this paper cites.
How to prove yourself: Practical solutions to identification and signature problems
Amos Fiat and Adi Shamir · 1986
Earlier work this paper cites.
Tutorial on reed-solomon error correction coding
William A Geisel · 1990
Earlier work this paper cites.
Random oracles are practical: a paradigm for designing efficient protocols
Mihir Bellare and Phillip Rogaway · 1993
Earlier work this paper cites.
Digital image watermarking: an overview
Nikos Nikolaidis and Ioannis Pitas · 1999
Earlier work this paper cites.
Information hiding: steganography and watermarking-attacks and countermeasures: steganography and watermarking: attacks and countermeasures
Neil F Johnson, Zoran Duric, and Sushil Jajodia · 2001
Earlier work this paper cites.
A guided tour to approximate string matching
Gonzalo Navarro · 2001
Earlier work this paper cites.
Natural language watermarking and tamperproofing
Mikhail J Atallah, Victor Raskin, Christian F Hempelmann, Mercan Karahan, Radu Sion, Umut Topkara, and Katrina E Triezenberg · 2002
Earlier work this paper cites.
Digital watermarking
Ingemar Cox, Matthew Miller, Jeffrey Bloom, and Chris Honsinger · 2002
Earlier work this paper cites.
Computing science: The easiest hard problem
Brian Hayes · 2002
Earlier work this paper cites.
Robust content-dependent high-fidelity watermark for tracking in digital cinema
Jeffrey Lubin, Jeffrey A Bloom, and Hui Cheng · 2003
Earlier work this paper cites.
The easiest hard problem: Number partitioning
Stephan Mertens · 2003
Earlier work this paper cites.
A survey of digital image watermarking techniques
Vidyasagar M Potdar, Song Han, and Elizabeth Chang · 2005
Earlier work this paper cites.
The hiding virtues of ambiguity: quantifiably resilient watermarking of natural language text through synonym substitutions
Umut Topkara, Mercan Topkara, and Mikhail J Atallah · 2006
Earlier work this paper cites.
An underwater optical communication system implementing reed-solomon channel coding
William C Cox, Jim A Simpson, Carlo P Domizioli, John F Muth, and Brian L Hughes · 2008
Earlier work this paper cites.
An introduction to galois fields and reed-solomon coding
James Westall and James Martin · 2010
Earlier work this paper cites.
Error-correction coding for digital communications
Jr Clark, C George, and J Bibb Cain · 2013
Earlier work this paper cites.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Ligero: Lightweight sublinear arguments without a trusted setup
Scott Ames, Carmit Hazay, Yuval Ishai, and Muthuramakrishnan Venkitasubramaniam · 2017
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi · 2019
Earlier work this paper cites.
Generating sentiment-preserving fake online reviews using neural language models and their human-and machine-based detection
David Ifeoluwa Adelani, Haotian Mai, Fuming Fang, Huy H Nguyen, Junichi Yamagishi, and Isao Echizen · 2020
Earlier work this paper cites.
Token-level adaptive training for neural machine translation
Shuhao Gu, Jinchao Zhang, Fandong Meng, Yang Feng, Wanying Xie, Jie Zhou, and Dong Yu · 2020
Cited alongside, same era.
Distortion agnostic deep watermarking
Xiyang Luo, Ruohan Zhan, Huiwen Chang, Feng Yang, and Peyman Milanfar · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Cited alongside, same era.
Adversarial watermarking transformer: Towards tracing text provenance with data hiding
Sahar Abdelnabi and Mario Fritz · 2021
Cited alongside, same era.
Statistical inference
George Casella and Roger L Berger · 2021
Cited alongside, same era.
Fundamentals of Classical and Modern Error-Correcting Codes
A robust semantics-based watermark for large language model against paraphrasing
Jie Ren, Han Xu, Yiding Liu, Yingqian Cui, Shuaiqiang Wang, Dawei Yin, and Jiliang Tang · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Assessing the importance of frequency versus compositionality for subword-based tokenization in nmt
Benoist Wolleb, Romain Silvestri, Giorgos Vernikos, Ljiljana Dolamic, and Andrei Popescu-Belis · 2023
Later among the works it cites.
Robust multi-bit natural language watermarking through invariant features
KiYoon Yoo, Wonhyuk Ahn, Jiho Jang, and Nojun Kwak · 2023
Later among the works it cites.
Multi-bit distortion-free watermarking for large language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shu Lin and Juane Li · 2021
Cited alongside, same era.
Codes for distributed storage
Vinayak Ramkumar, Myna Vajha, Srinivasan Babu Balaji, M Nikhil Krishnan, Birenjith Sasidharan, and P Vijay Kumar · 2021
Cited alongside, same era.
Generating fake cyber threat intelligence using transformer-based models
Priyanka Ranade, Aritran Piplai, Sudip Mittal, Anupam Joshi, and Tim Finin · 2021
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models, 2021
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2021
Cited alongside, same era.
Protecting intellectual property of language generation apis with lexical watermark
Xuanli He, Qiongkai Xu, Lingjuan Lyu, Fangzhao Wu, and Chenguang Wang · 2022
Cited alongside, same era.
Cater: Intellectual property protection on text generation apis via conditional watermarks
Xuanli He, Qiongkai Xu, Yi Zeng, Lingjuan Lyu, Fangzhao Wu, Jiwei Li, and Ruoxi Jia · 2022
Cited alongside, same era.
Targeted phishing campaigns using large scale language models
Rabimba Karanjai · 2022
Cited alongside, same era.
Massieh Kordi Boroujeny, Ya Jiang, Kai Zeng, and Brian Mark · 2024
Closest in time.
Watermarking language models with error correcting codes
Patrick Chao, Edgar Dobriban, and Hamed Hassani · 2024
Closest in time.
Undetectable watermarks for language models
Miranda Christ, Sam Gunn, and Or Zamir · 2024
Closest in time.
Scalable watermarking for identifying large language model outputs
Sumanth Dathathri, Abigail See, Sumedh Ghaisas, Po-Sen Huang, Rob McAdam, Johannes Welbl, Vandana Bachani, Alex Kaskasoli, Robert Stanforth, Tatiana Matejovicova, et al · 2024
Closest in time.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer · 2024
Closest in time.
Watermax: breaking the llm watermark detectability-robustness-quality trade-off
Eva Giboulot and Teddy Furon · 2024
Closest in time.
IvyPanda
IvyPanda · 2024
Closest in time.
Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense
Kalpesh Krishna, Yixiao Song, Marzena Karpinska, John Wieting, and Mohit Iyyer · 2024
Closest in time.
Robust distortion-free watermarks for language models
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto, and Percy Liang · 2024
Closest in time.
Who wrote this code? watermarking for code generation
Taehyun Lee, Seokhee Hong, Jaewoo Ahn, Ilgee Hong, Hwaran Lee, Sangdoo Yun, Jamin Shin, and Gunhee Kim · 2024
Closest in time.
Resilient watermarking for llm-generated codes
Boquan Li, Mengdi Zhang, Peixin Zhang, Jun Sun, and Xingmei Wang · 2024
Closest in time.
New Bing
Microsoft · 2024
Closest in time.
https://huggingface.co/datasets/ChristophSchuhmann/essays-with-instructions , 2023
Christoph Schuhmann · 2024
Closest in time.
Towards codable watermarking for injecting multi-bits information to llms
Lean Wang, Wenkai Yang, Deli Chen, Hao Zhou, Yankai Lin, Fandong Meng, Jie Zhou, and Xu Sun · 2024
Closest in time.
Advancing beyond identification: Multi-bit watermark for large language models
KiYoon Yoo, Wonhyuk Ahn, and Nojun Kwak · 2024
Closest in time.
Excuse me, sir? your language model is leaking (information)
Or Zamir · 2024
Closest in time.
REMARK-LLM: A robust and efficient watermarking framework for generative large language models
Ruisi Zhang, Shehzeen Samarah Hussain, Paarth Neekhara, and Farinaz Koushanfar · 2024
Closest in time.
Provable robust watermarking for AI-generated text
Xuandong Zhao, Prabhanjan Vijendra Ananth, Lei Li, and Yu-Xiang Wang · 2024
Closest in time.
Watermarking Language Models for Many Adaptive Users
Aloni Cohen, Alexander Hoover, and Gabe Schoenbach · 2025
Closest in time.