Fetching the paper…
Reading the bibliography…
Text watermarking algorithms are crucial for protecting the copyright of textual content.
Large language models in medicine
Arun James Thirunavukarasu, Darren Shu Jeng Ting, Kabilan Elangovan, Laura Gutierrez, Ting Fang Tan, and Daniel Shu Wei Ting. 2023 · 1940
Earlier work this paper cites.
A robust digital watermarking algorithm for text document copyright protection based on feature coding. In
Muhammad Munwar Iqbal, Umair Khadam, Ki Jun Han, Jihun Han, and Sohail Jabbar. 2019 · 1945
Earlier work this paper cites.
Electronic marking and identification techniques to discourage document copying
Jack T Brassil, Steven Low, Nicholas F. Maxemchuk, and Lawrence O’Gorman. 1995 · 1995
Earlier work this paper cites.
WordNet: An electronic lexical database
Christiane Fellbaum. 1998 · 1998
Earlier work this paper cites.
Natural language watermarking: Design, analysis, and a proof-of-concept implementation. In
Mikhail J Atallah, Victor Raskin, Michael Crogan, Christian Hempelmann, Florian Kerschbaum, Dina Mohamed, and Sanket Naik. 2001 · 2001
Earlier work this paper cites.
The homograph attack
Evgeniy Gabrilovich and Alex Gontmakher. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation. In
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Bad characters: Imperceptible nlp attacks. In
Nicholas Boucher, Ilia Shumailov, Ross Anderson, and Nicolas Papernot. 2022 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Natural language watermarking via morphosyntactic alterations
Hasan Mesut Meral, Bülent Sankur, A Sumru Özsoy, Tunga Güngör, and Emre Sevinç. 2009 · 2009
Earlier work this paper cites.
UniSpaCh: A text-based data hiding method using Unicode space characters
Lip Yee Por, KokSheik Wong, and Kok Onn Chee. 2012 · 2012
Earlier work this paper cites.
Copyright for web content using invisible text watermarking
Nighat Mir. 2014 · 2014
Earlier work this paper cites.
Concise analysis of current text automation and watermarking approaches
Mohammed Hazim Alkawaz, Ghazali Sulong, Tanzila Saba, Abdulaziz S Almazyad, and Amjad Rehman. 2016 · 2016
Earlier work this paper cites.
Content-preserving text watermarking through unicode homoglyph substitution. In
Stefano Giovanni Rizzo, Flavio Bertini, and Danilo Montesi. 2016 · 2016
Earlier work this paper cites.
Categorical Reparameterization with Gumbel-Softmax. In
Eric Jang, Shixiang Gu, and Ben Poole. 2017 · 2017
Earlier work this paper cites.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Mandar Joshi, Eunsol Choi, Daniel S Weld, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Embedding watermarks into deep neural networks. In
Yusuke Uchida, Yuki Nagai, Shigeyuki Sakazawa, and Shin’ichi Satoh. 2017 · 2017
Earlier work this paper cites.
Daniel Cer, Yinfei Yang, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, et al · 2018
Earlier work this paper cites.
Understanding Back-Translation at Scale. In
Sergey Edunov, Myle Ott, Michael Auli, and David Grangier. 2018 · 2018
Earlier work this paper cites.
A review of text watermarking: theory, methods, and applications
Nurul Shamimi Kamaruddin, Amirrudin Kamsin, Lip Yee Por, and Hameedur Rahman. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
A comparative analysis of information hiding techniques for copyright protection of text documents
Milad Taleby Ahvanooey, Qianmu Li, Hiuk Jae Shim, and Yanyan Huang. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
ELI5: Long Form Question Answering. In
Angela Fan, Yacine Jernite, Ethan Perez, David Grangier, Jason Weston, and Michael Auli. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In
Nils Reimers and Iryna Gurevych. 2019 · 2019
Earlier work this paper cites.
Digital image watermarking techniques: a review
Mahbuba Begum and Mohammad Shorif Uddin. 2020 · 2020
Earlier work this paper cites.
The Curious Case of Neural Text Degeneration. In
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Earlier work this paper cites.
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension. In
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Earlier work this paper cites.
Unsupervised text generation by learning from search
Jingjing Li, Zichao Li, Lili Mou, Xin Jiang, Michael Lyu, and Irwin King. 2020 · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Earlier work this paper cites.
Incorporating bert into neural machine translation
Jinhua Zhu, Yingce Xia, Lijun Wu, Di He, Tao Qin, Wengang Zhou, Houqiang Li, and Tie-Yan Liu. 2020 · 2020
Earlier work this paper cites.
Adversarial watermarking transformer: Towards tracing text provenance with data hiding. In
Sahar Abdelnabi and Mario Fritz. 2021 · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde De Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Truthfulqa: Measuring how models mimic human falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans. 2021 · 2021
Earlier work this paper cites.
DISSIMILAR: Towards fake news detection using information hiding, signal processing and machine learning. In
David Megías, Minoru Kuribayashi, Andrea Rosales, and Wojciech Mazurczyk. 2021 · 2021
Earlier work this paper cites.
Review of the Literature on the Steganography Concept
Fatih Şahin, Taner Çevik, and Mustafa Takaoğlu. 2021 · 2021
Earlier work this paper cites.
Watermarking GPT outputs
S. Aaronson and H. Kirchner. 2022 · 2022
Earlier work this paper cites.
No language left behind: Scaling human-centered machine translation
Marta R Costa-jussà, James Cross, Onur Çelebi, Maha Elbayad, Kenneth Heafield, Kevin Heffernan, Elahe Kalbassi, Janice Lam, Daniel Licht, Jean Maillard, et al · 2022
Earlier work this paper cites.
Cater: Intellectual property protection on text generation apis via conditional watermarks
Xuanli He, Qiongkai Xu, Yi Zeng, Lingjuan Lyu, Fangzhao Wu, Jiwei Li, and Ruoxi Jia. 2022b · 2022
Earlier work this paper cites.
Text Revision by On-the-Fly Representation Optimization
Jingjing Li, Zichao Li, Tao Ge, Irwin King, and Michael R Lyu. 2022 · 2022
Earlier work this paper cites.
PanGu-Bot: Efficient Generative Dialogue Pre-training from Pre-trained Language Model
Fei Mi, Yitong Li, Yulong Zeng, Jingyan Zhou, Yasheng Wang, Chuanfei Xu, Lifeng Shang, Xin Jiang, Shiqi Zhao, and Qun Liu. 2022 · 2022
Cited alongside, same era.
Red Teaming Language Models with Language Models. In
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving. 2022 · 2022
Cited alongside, same era.
Language Models that Seek for Knowledge: Modular Search & Generation for Dialogue and Prompt Completion. In
Kurt Shuster, Mojtaba Komeili, Leonard Adolphs, Stephen Roller, Arthur Szlam, and Jason Weston. 2022 · 2022
Cited alongside, same era.
Coprotector: Protect open-source code against unauthorized training usage with data poisoning. In
Zhensu Sun, Xiaoning Du, Fu Song, Mingze Ni, and Li Li. 2022 · 2022
Cited alongside, same era.
Advancing Beyond Identification: Multi-bit Watermark for Language Models
KiYoon Yoo, Wonhyuk Ahn, and Nojun Kwak. 2023a · 2023
Closest in time.
REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
Ruisi Zhang, Shehzeen Samarah Hussain, Paarth Neekhara, and Farinaz Koushanfar. 2023 · 2023
Closest in time.
Synthetic Lies: Understanding AI-Generated Misinformation and Evaluating Algorithmic and Human Solutions. In
Jiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea G Parker, and Munmun De Choudhury. 2023 · 2023
Closest in time.
Can LLM-Generated Misinformation Be Detected?. In
Canyu Chen and Kai Shu. 2024 · 2024
Closest in time.
Undetectable watermarks for language models. In
Miranda Christ, Sam Gunn, and Or Zamir. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Romal Thoppilan, Daniel De Freitas, Jamie Hall, Noam Shazeer, Apoorv Kulshreshtha, Heng-Tze Cheng, Alicia Jin, Taylor Bos, Leslie Baker, Yu Du, et al · 2022
Cited alongside, same era.
Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models. In
Priyan Vaithilingam, Tianyi Zhang, and Elena L Glassman. 2022 · 2022
Cited alongside, same era.
Paraphrastic Representations at Scale. In
John Wieting, Kevin Gimpel, Graham Neubig, and Taylor Berg-Kirkpatrick. 2022 · 2022
Cited alongside, same era.
A systematic evaluation of large language models of code. In
Frank F Xu, Uri Alon, Graham Neubig, and Vincent Josua Hellendoorn. 2022 · 2022
Cited alongside, same era.
Tracing text provenance via context-aware lexical substitution. In
Xi Yang, Jie Zhang, Kejiang Chen, Weiming Zhang, Zehua Ma, Feng Wang, and Nenghai Yu. 2022 · 2022
Cited alongside, same era.
Opt: Open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al · 2022
Cited alongside, same era.
Model leeching: An extraction attack targeting llms
Lewis Birch, William Hackett, Stefan Trawicki, Neeraj Suri, and Peter Garraghan. 2023 · 2023
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al · 2023
Cited alongside, same era.
Watermarking conditional text generation for ai detection: Unveiling challenges and a semantic-aware watermark remedy. In
Yu Fu, Deyi Xiong, and Yue Dong. 2024 · 2024
Closest in time.
WaterMax: breaking the LLM watermark detectability-robustness-quality trade-off
Eva Giboulot and Furon Teddy. 2024 · 2024
Closest in time.
On the Learnability of Watermarks for Language Models. In
Chenchen Gu, Xiang Lisa Li, Percy Liang, and Tatsunori Hashimoto. 2024 · 2024
Closest in time.
Codeip: A grammar-guided multi-bit watermark for large language models of code
Batu Guan, Yao Wan, Zhangqian Bi, Zheng Wang, Hongyu Zhang, Yulei Sui, Pan Zhou, and Lichao Sun. 2024 · 2024
Closest in time.
Zhiwei He, Binglin Zhou, Hongkun Hao, Aiwei Liu, Xing Wang, Zhaopeng Tu, Zhuosheng Zhang, and Rui Wang. 2024 · 2024
Closest in time.
SemStamp: A Semantic Watermark with Paraphrastic Robustness for Text Generation. In
Abe Hou, Jingyu Zhang, Tianxing He, Yichen Wang, Yung-Sung Chuang, Hongwei Wang, Lingfeng Shen, Benjamin Van Durme, Daniel Khashabi, and Yulia Tsvetkov. 2024a · 2024
Closest in time.
k-SemStamp: A Clustering-Based Semantic Watermark for Detection of Machine-Generated Text
Abe Bohan Hou, Jingyu Zhang, Yichen Wang, Daniel Khashabi, and Tianxing He. 2024b · 2024
Closest in time.
Unbiased Watermark for Large Language Models. In
Zhengmian Hu, Lichang Chen, Xidong Wu, Yihan Wu, Hongyang Zhang, and Heng Huang. 2024 · 2024
Closest in time.
Watermark stealing in large language models
Nikola Jovanović, Robin Staab, and Martin Vechev. 2024 · 2024
Closest in time.
On the Reliability of Watermarks for Large Language Models. In
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum, and Tom Goldstein. 2024 · 2024
Closest in time.
Robust Distortion-free Watermarks for Language Models
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto, and Percy Liang. 2024 · 2024
Closest in time.
Academic integrity and artificial intelligence: An overview
Rahul Kumar, Sarah Elaine Eaton, Michael Mindzak, and Ryan Morrison. 2024 · 2024
Closest in time.
Adaptive ensembles of fine-tuned transformers for llm-generated text detection
Zhixin Lai, Xuesheng Zhang, and Suiyao Chen. 2024 · 2024
Closest in time.
Waterfall: Framework for Robust and Scalable Text Watermarking
Gregory Kang Ruey Lau, Xinyuan Niu, Hieu Dao, Jiangwei Chen, Chuan-Sheng Foo, and Bryan Kian Hsiang Low. 2024 · 2024
Closest in time.
WatME: Towards Lossless Watermarking Through Lexical Redundancy. In
CHEN Liang, Yatao Bian, Yang Deng, Deng Cai, Shuaiyi Li, Peilin Zhao, and Kam-Fai Wong. 2024 · 2024
Closest in time.
Adaptive Text Watermark for Large Language Models. In
Yepeng Liu and Yuheng Bu. 2024 · 2024
Closest in time.
An Entropy-based Text Watermarking Detection Method
Yijian Lu, Aiwei Liu, Dianzhi Yu, Jingjing Li, and Irwin King. 2024 · 2024
Closest in time.
Lost in Overlap: Exploring Watermark Collision in LLMs
Yiyang Luo, Ke Lin, and Chao Gu. 2024 · 2024
Closest in time.
WaterJudge: Quality-Detection Trade-off when Watermarking Large Language Models. In
Piotr Molenda, Adian Liusie, and Mark Gales. 2024 · 2024
Closest in time.
MarkLLM: An Open-Source Toolkit for LLM Watermarking
Leyi Pan, Aiwei Liu, Zhiwei He, Zitian Gao, Xuandong Zhao, Yijian Lu, Binglin Zhou, Shuliang Liu, Xuming Hu, Lijie Wen, et al · 2024
Closest in time.
Can Sensitive Information Be Deleted From LLMs? Objectives for Defending Against Extraction Attacks. In
Vaidehi Patil, Peter Hase, and Mohit Bansal. 2024 · 2024
Closest in time.
A Robust Semantics-based Watermark for Large Language Model against Paraphrasing. In
Jie Ren, Han Xu, Yiding Liu, Yingqian Cui, Shuaiqiang Wang, Dawei Yin, and Jiliang Tang. 2024 · 2024
Closest in time.
Watermarking Makes Language Models Radioactive
Tom Sander, Pierre Fernandez, Alain Durmus, Matthijs Douze, and Teddy Furon. 2024 · 2024
Closest in time.
Is Watermarking LLM-Generated Code Robust?. In
Tarun Suresh, Shubham Ugare, Gagandeep Singh, and Sasa Misailovic. 2024 · 2024
Closest in time.
Towards Codable Watermarking for Injecting Multi-Bits Information to LLMs. In
Lean Wang, Wenkai Yang, Deli Chen, Hao Zhou, Yankai Lin, Fandong Meng, Jie Zhou, and Xu Sun. 2024 · 2024
Closest in time.
Bypassing LLM Watermarks with Color-Aware Substitutions
Qilong Wu and Varun Chandrasekaran. 2024 · 2024
Closest in time.
Distortion-free Watermarks are not Truly Distortion-free under Watermark Key Collisions
Yihan Wu, Ruibo Chen, Zhengmian Hu, Yanshuo Chen, Junfeng Guo, Hongyang Zhang, and Heng Huang. 2024 · 2024
Closest in time.
Hengyuan Xu, Liyao Xiang, Xingjun Ma, Borui Yang, and Baochun Li. 2024a · 2024
Closest in time.
Learning to watermark llm-generated text via reinforcement learning
Xiaojun Xu, Yuanshun Yao, and Yang Liu. 2024b · 2024
Closest in time.
A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Yifan Yao, Jinhao Duan, Kaidi Xu, Yuanfang Cai, Zhibo Sun, and Yue Zhang. 2024 · 2024
Closest in time.
Benchmarking Large Language Models for News Summarization
Tianyi Zhang, Faisal Ladhak, Esin Durmus, Percy Liang, Kathleen Mckeown, and Tatsunori B Hashimoto. 2024a · 2024
Closest in time.
Large Language Model Watermark Stealing With Mixed Integer Programming
Zhaoxi Zhang, Xiaomei Zhang, Yanjun Zhang, Leo Yu Zhang, Chao Chen, Shengshan Hu, Asif Gill, and Shirui Pan. 2024b · 2024
Closest in time.
Provable Robust Watermarking for AI-Generated Text. In
Xuandong Zhao, Prabhanjan Vijendra Ananth, Lei Li, and Yu-Xiang Wang. 2024 · 2024
Closest in time.
Generative AI security: challenges and countermeasures
Banghua Zhu, Norman Mu, Jiantao Jiao, and David Wagner. 2024 · 2024
Closest in time.
Robust multi-bit natural language watermarking through invariant features. In
KiYoon Yoo, Wonhyuk Ahn, Jiho Jang, and Nojun Kwak. 2023b · 2092
Closest in time.