Fetching the paper…
Reading the bibliography…
Watermarking approaches are proposed to identify if text being circulated is human or large language model (LLM) generated.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Tutorial on large deviations for the binomial distribution
Richard Arratia and Louis Gordon. 1989 · 1989
Earlier work this paper cites.
Probability inequalities for sums of bounded random variables
Wassily Hoeffding. 1994 · 1994
Earlier work this paper cites.
The de moivre-laplace central limit theorem
Steven R Dunbar. 2011 · 2011
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Earlier work this paper cites.
Academic dishonesty or academic integrity? using natural language processing (nlp) techniques to investigate positive integrity in academic integrity research
Thomas Lancaster. 2021 · 2021
Earlier work this paper cites.
Illustrating reinforcement learning from human feedback (rlhf)
Nathan Lambert, Louis Castricato, Leandro von Werra, and Alex Havrilla. 2022 · 2022
Earlier work this paper cites.
Can llm-generated misinformation be detected?
Canyu Chen and Kai Shu. 2023 · 2023
Earlier work this paper cites.
Undetectable watermarks for language models
Miranda Christ, Sam Gunn, and Or Zamir. 2023 · 2023
Earlier work this paper cites.
Large language models for software engineering: Survey and open problems
Angela Fan, Beliz Gokkaya, Mark Harman, Mitya Lyubarskiy, Shubho Sengupta, Shin Yoo, and Jie M Zhang. 2023 · 2023
Cited alongside, same era.
Towards possibilities & impossibilities of ai-generated text detection: A survey
Soumya Suvra Ghosal, Souradip Chakraborty, Jonas Geiping, Furong Huang, Dinesh Manocha, and Amrit Singh Bedi. 2023 · 2023
Cited alongside, same era.
Large language models: a comprehensive survey of its applications, challenges, limitations, and future prospects
Muhammad Usman Hadi, Rizwan Qureshi, Abbas Shah, Muhammad Irfan, Anas Zafar, Muhammad Bilal Shaikh, Naveed Akhtar, Jia Wu, Seyedali Mirjalili, et al. 2023 · 2023
Cited alongside, same era.
Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense
Kalpesh Krishna, Yixiao Song, Marzena Karpinska, John Wieting, and Mohit Iyyer. 2023 · 2023
Cited alongside, same era.
Red teaming language model detectors with language models
Zhouxing Shi, Yihan Wang, Fan Yin, Xiangning Chen, Kai-Wei Chang, and Cho-Jui Hsieh. 2023 · 2023
Later among the works it cites.
Baselines for identifying watermarked large language models
Leonard Tang, Gavin Uberti, and Tom Shlomi. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models, 2023
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
A survey on detection of llms-generated content
Xianjun Yang, Liangming Pan, Xuandong Zhao, Haifeng Chen, Linda Petzold, William Yang Wang, and Wei Cheng. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto, and Percy Liang. 2023 · 2023
Cited alongside, same era.
Large language models can be guided to evade ai-generated text detection
Ning Lu, Shengcai Liu, Rui He, and Ke Tang. 2023 · 2023
Cited alongside, same era.
Mark my words: Analyzing and evaluating language model watermarks
Julien Piet, Chawin Sitawarin, Vivian Fang, Norman Mu, and David Wagner. 2023 · 2023
Cited alongside, same era.
Can ai-generated text be reliably detected?
Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian, Wenxiao Wang, and Soheil Feizi. 2023 · 2023
Cited alongside, same era.
A watermark for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein. 2023a
Cited in the paper.
On the reliability of watermarks for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum, and Tom Goldstein. 2023b
Cited in the paper.
Hanlin Zhang, Benjamin L Edelman, Danilo Francati, Daniele Venturi, Giuseppe Ateniese, and Boaz Barak. 2023 · 2023
Later among the works it cites.
Provable robust watermarking for ai-generated text
Xuandong Zhao, Prabhanjan Ananth, Lei Li, and Yu-Xiang Wang. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al. 2023 · 2023
Later among the works it cites.
My ai safety lecture for ut effective altruism
Scott Aaronson. 2022 · 2024
Closest in time.