Fetching the paper…
Reading the bibliography…
The advancement of Large Language Models (LLMs) has led to increasing concerns about the misuse of AI-generated text, and watermarking for LLM-generated text has emerged as a potential solution.
A mathematical theory of communication
Shannon, C. E · 1948
Earlier work this paper cites.
Electronic marking and identification techniques to discourage document copying
Brassil, J., Low, S., Maxemchuk, N., and O’Gorman, L · 1995
Earlier work this paper cites.
Natural language watermarking: Design, analysis, and a proof-of-concept implementation
Atallah, M. J., Raskin, V., Crogan, M., Hempelmann, C., Kerschbaum, F., Mohamed, D., and Naik, S · 2001
Earlier work this paper cites.
Words are not enough: Sentence level natural language watermarking
Topkara, M., Topkara, U., and Atallah, M. J · 2006
Earlier work this paper cites.
Natural language watermarking via morphosyntactic alterations
Meral, H. M., Sankur, B., Sumru Özsoy, A., Güngör, T., and Sevinç, E · 2008
Earlier work this paper cites.
Unispach: A text-based data hiding method using unicode space characters
Por, L. Y., Wong, K., and Chee, K. O · 2012
Earlier work this paper cites.
Content-preserving text watermarking through unicode homoglyph substitution
Rizzo, S. G., Bertini, F., and Montesi, D · 2016
Earlier work this paper cites.
Eli5: Long form question answering
Fan, A., Jernite, Y., Perez, E., Grangier, D., Weston, J., and Auli, M · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Reimers, N. and Gurevych, I · 2019
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2020
Earlier work this paper cites.
Adversarial watermarking transformer: Towards tracing text provenance with data hiding
Abdelnabi, S. and Fritz, M · 2021
Earlier work this paper cites.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
Wang, B. and Komatsuzaki, A · 2021
Earlier work this paper cites.
Tracing text provenance via context-aware lexical substitution
Yang, X., Zhang, J., Chen, K., Zhang, W., Ma, Z., Wang, F., and Yu, N · 2022
Cited alongside, same era.
Opt: Open pre-trained transformer language models
Zhang, S., Roller, S., Goyal, N., Artetxe, M., Chen, M., Chen, S., Dewan, C., Diab, M., Li, X., Lin, X. V., et al · 2022
Cited alongside, same era.
Undetectable watermarks for language models
Christ, M., Gunn, S., and Zamir, O · 2023
Cited alongside, same era.
Publicly detectable watermarking for language models
Fairoze, J., Garg, S., Jha, S., Mahloujifar, S., Mahmoody, M., and Wang, M · 2023
Cited alongside, same era.
Semstamp: A semantic watermark with paraphrastic robustness for text generation
Can ai-generated text be reliably detected?
Sadasivan, V. S., Kumar, A., Balasubramanian, S., Wang, W., and Feizi, S · 2023
Later among the works it cites.
Embarrassingly simple text watermarks
Sato, R., Takezawa, Y., Bao, H., Niwa, K., and Yamada, M · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models, 2023
Touvron, H., Martin, L., Stone, K., Albert, P., Almahairi, A., Babaei, Y., Bashlykov, N., Batra, S., Bhargava, P., Bhosale, S., Bikel, D., Blecher, L., Ferrer, C. C., Chen, M., Cucurull, G., Esiobu, D., Fernandes, J., Fu, J., Fu, W., Fuller, B., Gao, C., Goswami, V., Goyal, N., Hartshorn, A., Hosseini, S., Hou, R., Inan, H., Kardas, M., Kerkez, V., Khabsa, M., Kloumann, I., Korenev, A., Koura, P. S., Lachaux, M.-A., Lavril, T., Lee, J., Liskovich, D., Lu, Y., Mao, Y., Martinet, X., Mihaylov, T., Mishra, P., Molybog, I., Nie, Y., Poulton, A., Reizenstein, J., Rungta, R., Saladi, K., Schelten, A., Silva, R., Smith, E. M., Subramanian, R., Tan, X. E., Tang, B., Taylor, R., Williams, A., Kuan, J. X., Xu, P., Yan, Z., Zarov, I., Zhang, Y., Fan, A., Kambadur, M., Narang, S., Rodriguez, A., Stojnic, R., Edunov, S., and Scialom, T · 2023
Later among the works it cites.
Waterbench: Towards holistic evaluation of watermarks for large language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hou, A. B., Zhang, J., He, T., Wang, Y., Chuang, Y.-S., Wang, H., Shen, L., Van Durme, B., Khashabi, D., and Tsvetkov, Y · 2023
Cited alongside, same era.
Unbiased watermark for large language models
Hu, Z., Chen, L., Wu, X., Wu, Y., Zhang, H., and Huang, H · 2023
Cited alongside, same era.
Jiang, A. Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D. S., Casas, D. d. l., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., et al · 2023
Cited alongside, same era.
Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense
Krishna, K., Song, Y., Karpinska, M., Wieting, J., and Iyyer, M · 2023
Cited alongside, same era.
Robust distortion-free watermarks for language models
Kuditipudi, R., Thickstun, J., Hashimoto, T., and Liang, P · 2023
Cited alongside, same era.
Who wrote this code? watermarking for code generation
Lee, T., Hong, S., Ahn, J., Hong, I., Lee, H., Yun, S., Shin, J., and Kim, G · 2023
Cited alongside, same era.
Munyer, T. and Zhong, X · 2023
Cited alongside, same era.
A robust semantics-based watermark for large language model against paraphrasing
Ren, J., Xu, H., Liu, Y., Cui, Y., Wang, S., Yin, D., and Tang, J · 2023
Cited alongside, same era.
Tu, S., Sun, Y., Bai, Y., Yu, J., Hou, L., and Li, J · 2023
Later among the works it cites.
Towards codable text watermarking for large language models
Wang, L., Yang, W., Chen, D., Zhou, H., Lin, Y., Meng, F., Zhou, J., and Sun, X · 2023
Later among the works it cites.
Dipmark: A stealthy, efficient and resilient watermark for large language models
Wu, Y., Hu, Z., Zhang, H., and Huang, H · 2023
Later among the works it cites.
Watermarking text generated by black-box language models
Yang, X., Chen, K., Zhang, W., Liu, C., Qi, Y., Zhang, J., Fang, H., and Yu, N · 2023
Later among the works it cites.
Robust multi-bit natural language watermarking through invariant features
Yoo, K., Ahn, W., Jang, J., and Kwak, N · 2023
Later among the works it cites.
Provable robust watermarking for ai-generated text
Zhao, X., Ananth, P., Li, L., and Wang, Y.-X · 2023
Later among the works it cites.
Gumbelsoft: Diversified language model watermarking via the gumbelmax-trick
Fu, J., Zhao, X., Yang, R., Zhang, Y., Chen, J., and Xiao, Y · 2024
Closest in time.
Huo, M., Somayajula, S. A., Liang, Y., Zhang, R., Koushanfar, F., and Xie, P · 2024
Closest in time.
Permute-and-flip: An optimally robust and watermarkable decoder for llms
Zhao, X., Li, L., and Wang, Y.-X · 2024
Closest in time.