Fetching the paper…
Reading the bibliography…
It has been shown that finetuned transformers and other supervised detectors effectively distinguish between human and machine-generated text in some situations arXiv:2305.13242, but we find that even simple classifiers on top of n-gram and part-of-speech features can achieve very robust performance on both in- and out-of-domain data.
Greedy function approximation: a gradient boosting machine
Jerome H Friedman. 2001 · 2001
Earlier work this paper cites.
Longformer: The long-document transformer
Iz Beltagy, Matthew E. Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay. 2011 · 2011
Earlier work this paper cites.
Gltr: Statistical detection and visualization of generated text
Sebastian Gehrmann, Hendrik Strobelt, and Alexander M. Rush. 2019 · 2019
Earlier work this paper cites.
All That’s ‘Human’ Is Not Gold: Evaluating Human Evaluation of Generated Text
Elizabeth Clark, Tal August, Sofia Serrano, Nikita Haduong, Suchin Gururangan, and Noah A. Smith. 2021 · 2021
Earlier work this paper cites.
Learning universal authorship representations
Rafael A. Rivera Soto, Olivia Elizabeth Miano, Juanita Ordoñez, Barry Y. Chen, Aleem Khan, Marcus Bishop, and Nicholas Andrews. 2021 · 2021
Earlier work this paper cites.
Feedback prize - predicting effective arguments
Alex Franklin, Maggie, Meg Benner, Natalie Rambis, Perpetual Baffour, Ryan Holbrook, Scott Crossley, and ulrichboser. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke E. Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Francis Christiano, Jan Leike, and Ryan J. Lowe. 2022 · 2022
Earlier work this paper cites.
ChatGPT Is Making Universities Rethink Plagiarism
Sofia Barnett. 2023 · 2023
Cited alongside, same era.
Evade chatgpt detectors via a single space
Shuyang Cai and Wanyun Cui. 2023 · 2023
Cited alongside, same era.
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection
Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, and Yupeng Wu. 2023 · 2023
Cited alongside, same era.
A watermark for large language models
John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, and Tom Goldstein. 2023 · 2023
Cited alongside, same era.
Ryuto Koike, Masahiro Kaneko, and Naoaki Okazaki. 2023 · 2023
Cited alongside, same era.
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Stefano Ermon, Christopher D. Manning, and Chelsea Finn. 2023 · 2023
Later among the works it cites.
Ghostbuster: Detecting text ghostwritten by large language models
Vivek Kumar Verma, Eve Fleisig, Nicholas Tomlin, and Dan Klein. 2023 · 2023
Later among the works it cites.
Educators Battle Plagiarism As 89% Of Students Admit To Using OpenAI’s ChatGPT For Homework
Chris Westfall. 2023 · 2023
Later among the works it cites.
Authorship attribution methods, challenges, and future research directions: A comprehensive survey
Xie He, Arash Habibi Lashkari, Nikhill Vombatkere, and Dilli Prasad Sharma. 2024 · 2024
Closest in time.
Petkaz at semeval-2024 task 8: Can linguistics capture the specifics of llm-generated text?
Kseniia Petukhova, Roman Kazakov, and Ekaterina Kochmar. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kalpesh Krishna, Yixiao Song, Marzena Karpinska, John Wieting, and Mohit Iyyer. 2023 · 2023
Cited alongside, same era.
Deepfake Text Detection in the Wild
Yafu Li, Qintong Li, Leyang Cui, Wei Bi, Longyue Wang, Linyi Yang, Shuming Shi, and Yue Zhang. 2023 · 2023
Cited alongside, same era.
DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability Curvature
Eric Mitchell, Yoonho Lee, Alexander Khazatsky, Christopher D. Manning, and Chelsea Finn. 2023 · 2023
Cited alongside, same era.
M4: Multi-generator, multi-domain, and multi-lingual black-box machine-generated text detection
Yuxia Wang, Jonibek Mansurov, Petar Ivanov, Jinyan Su, Artem Shelmanov, Akim Tsvigun, Chenxi Whitehouse, Osama Mohammed Afzal, Tarek Mahmoud, Toru Sasaki, Thomas Arnold, Alham Aji, Nizar Habash, Iryna Gurevych, and Preslav Nakov. 2024b
Cited in the paper.
Abel Salinas and Fred Morstatter. 2024 · 2024
Closest in time.
SemEval-2024 task 8: Multigenerator, multidomain, and multilingual black-box machine-generated text detection
Yuxia Wang, Jonibek Mansurov, Petar Ivanov, Jinyan Su, Artem Shelmanov, Akim Tsvigun, Chenxi Whitehouse, Osama Mohammed Afzal, Tarek Mahmoud, Giovanni Puccetti, Thomas Arnold, Alham Fikri Aji, Nizar Habash, Iryna Gurevych, and Preslav Nakov. 2024a · 2024
Closest in time.
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis
Wataru Zaitsu and Mingzhe Jin. 2023 · 2024
Closest in time.