Fetching the paper…
Reading the bibliography…
The emergence of large language models (LLMs) has resulted in the production of LLM-generated texts that is highly sophisticated and almost indistinguishable from texts written by humans.
An estimate of an upper bound for the entropy of English
Peter F Brown, Stephen A Della Pietra, Vincent J Della Pietra, Jennifer C Lai, and Robert L Mercer. 1992 · 1992
Earlier work this paper cites.
Electronic marking and identification techniques to discourage document copying
Jack T Brassil, Steven Low, Nicholas F. Maxemchuk, and Lawrence O’Gorman. 1995 · 1995
Earlier work this paper cites.
Natural language watermarking: Design, analysis, and a proof-of-concept implementation. In Information Hiding: 4th International Workshop, IH 2001 Pittsburgh, PA, USA, April 25–27, 2001 Proceedings 4 . Springer, Pittsburgh, 185–200
Mikhail J Atallah, Victor Raskin, Michael Crogan, Christian Hempelmann, Florian Kerschbaum, Dina Mohamed, and Sanket Naik. 2001 · 2001
Earlier work this paper cites.
Word reordering and a dynamic programming beam search algorithm for statistical machine translation
Christoph Tillmann and Hermann Ney. 2003 · 2003
Earlier work this paper cites.
The hiding virtues of ambiguity: quantifiably resilient watermarking of natural language text through synonym substitutions. In Proceedings of the 8th workshop on Multimedia and security . Association for Computing Machinery, Geneva, Switzerland, 164–174
Umut Topkara, Mercan Topkara, and Mikhail J Atallah. 2006 · 2006
Earlier work this paper cites.
A review of digital watermarking techniques for text documents. In 2009 International Conference on Information and Multimedia Technology . IEEE, IEEE Computer Society, Washington DC, United States, 230–234
Zunera Jalil and Anwar M Mirza. 2009 · 2009
Earlier work this paper cites.
Zipf’s word frequency law in natural language: A critical review and future directions
Steven T Piantadosi. 2014 · 2014
Earlier work this paper cites.
Better malware ground truth: Techniques for weighting anti-virus vendor labels. In Proceedings of the 8th ACM Workshop on Artificial Intelligence and Security . Association for Computing Machinery, Denver, Colorado, USA, 45–56
Alex Kantchelian, Michael Carl Tschantz, Sadia Afroz, Brad Miller, Vaishaal Shankar, Rekha Bachwani, Anthony D Joseph, and J Doug Tygar. 2015 · 2015
Earlier work this paper cites.
TURINGBENCH: A Benchmark Environment for Turing Test in the Age of Neural Text Generation. In Findings of the Association for Computational Linguistics: EMNLP 2021 . 2001–2016
Adaku Uchendu, Zeyu Ma, Thai Le, Rui Zhang, and Dongwon Lee. 2021 · 2016
Earlier work this paper cites.
ELI5: Long Form Question Answering. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy, 3558–3567
Angela Fan, Yacine Jernite, Ethan Perez, David Grangier, Jason Weston, and Michael Auli. 2019 · 2019
Earlier work this paper cites.
GLTR: Statistical Detection and Visualization of Generated Text. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics: System Demonstrations . 111–116
Sebastian Gehrmann, Hendrik Strobelt, and Alexander M Rush. 2019 · 2019
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Earlier work this paper cites.
Release strategies and the social impacts of language models
Irene Solaiman, Miles Brundage, Jack Clark, Amanda Askell, Ariel Herbert-Voss, Jeff Wu, Alec Radford, Gretchen Krueger, Jong Wook Kim, Sarah Kreps, et al · 2019
Cited alongside, same era.
Defending against neural fake news
Rowan Zellers, Ari Holtzman, Hannah Rashkin, Yonatan Bisk, Ali Farhadi, Franziska Roesner, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
RoFT: A Tool for Evaluating Human Detection of Machine-Generated Text. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations . Association for Computational Linguistics, online, 189–196
Liam Dugan, Daphne Ippolito, Arun Kirubarajan, and Chris Callison-Burch. 2020 · 2020
Cited alongside, same era.
GPT-3: Its nature, scope, limits, and consequences
Luciano Floridi and Massimo Chiriatti. 2020 · 2020
Cited alongside, same era.
Automatic Detection of Generated Text is Easiest when Humans are Fooled. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 1808–1822
Fake news detection and fact verification using knowledge graphs and machine learning
Danish Shakeel and Nitin Jain. 2021 · 2021
Later among the works it cites.
Adversarial Robustness of Neural-Statistical Features in Detection of Generative Transformers. In 2022 International Joint Conference on Neural Networks (IJCNN) . Institute of Electrical and Electronics Engineers, Padua, Italy, 1–8
Evan Crothers, Nathalie Japkowicz, Herna Viktor, and Paula Branco. 2022 · 2022
Later among the works it cites.
Cross-Domain Detection of GPT-2-Generated Technical Text. In Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Seattle, United States, 1213–1233
Juan Rodriguez, Todd Hay, David Gros, Zain Shamsi, and Ravi Srinivasan. 2022 · 2022
Later among the works it cites.
Real or Fake Text? Investigating Human Ability to Detect Boundaries Between Human-Written and Machine-Generated Text. In The 37th AAAI Conference on Artificial Intelligence (AAAI 2023) . Association for the Advancement of Artificial Intelligence, Washington, USA, 104979
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Daphne Ippolito, Daniel Duckworth, Chris Callison-Burch, and Douglas Eck. 2020 · 2020
Cited alongside, same era.
How Decoding Strategies Affect the Verifiability of Generated Text. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, Online, 223–235
Luca Massarelli, Fabio Petroni, Aleksandra Piktus, Myle Ott, Tim Rocktäschel, Vassilis Plachouras, Fabrizio Silvestri, and Sebastian Riedel. 2020 · 2020
Cited alongside, same era.
Neural Deepfake Detection with Factual Structure of Text. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Minneapolis, Minnesota, 2461–2470
Wanjun Zhong, Duyu Tang, Zenan Xu, Ruize Wang, Nan Duan, Ming Zhou, Jiahai Wang, and Jian Yin. 2020 · 2020
Cited alongside, same era.
Adversarial watermarking transformer: Towards tracing text provenance with data hiding. In 2021 IEEE Symposium on Security and Privacy (SP) . IEEE, Institute of Electrical and Electronics Engineers, Online, 121–140
Sahar Abdelnabi and Mario Fritz. 2021 · 2021
Cited alongside, same era.
Detecting Bot-Generated Text by Characterizing Linguistic Accommodation in Human-Bot Interactions. In Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 . Association for Computational Linguistics, Bangkok, Thailand, 3235–3247
Paras Bhatt and Anthony Rios. 2021 · 2021
Cited alongside, same era.
All That’s ‘Human’Is Not Gold: Evaluating Human Evaluation of Generated Text. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Bangkok, Thailand, 7282–7296
Elizabeth Clark, Tal August, Sofia Serrano, Nikita Haduong, Suchin Gururangan, and Noah A Smith. 2021 · 2021
Cited alongside, same era.
TweepFake: About detecting deepfake tweets
Tiziano Fagni, Fabrizio Falchi, Margherita Gambini, Antonio Martella, and Maurizio Tesconi. 2021 · 2021
Cited alongside, same era.
Feature-based detection of automated language models: tackling GPT-2, GPT-3 and Grover
Leon Fröhling and Arkaitz Zubiaga. 2021 · 2021
Cited alongside, same era.
Liam Dugan, Daphne Ippolito, Arun Kirubarajan, Sherry Shi, Chris Callison-Burch, Pei Zhou, Andrew Zhu, Jennifer Hu, Jay Pujara, Xiang Ren, et al · 2023
Closest in time.
College instructor put on blast for accusing students of using ChatGPT on final assignments
Uwa Ede-Osifo. 2023 · 2023
Closest in time.
NYC education department blocks ChatGPT on school devices, networks
Michael Elsen-Rooney. 2023 · 2023
Closest in time.
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection
Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, and Yupeng Wu. 2023 · 2023
Closest in time.
A Watermark for Large Language Models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein. 2023 · 2023
Closest in time.
gpt-2-output-dataset
OpenAI. 2021 · 2023
Closest in time.
ChatGPT passes MBA exam given by a Wharton professor
Kalhan Rosenblatt. 2023 · 2023
Closest in time.
Can AI-Generated Text be Reliably Detected?
Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian, Wenxiao Wang, and Soheil Feizi. 2023 · 2023
Closest in time.
Meta’s powerful AI language model has leaked online — what happens now?
James Vincent. 2023 · 2023
Closest in time.