Fetching the paper…
Reading the bibliography…
Due to the rapid development of large language models, people increasingly often encounter texts that may start as written by a human but continue as machine-generated.
Liii. on lines and planes of closest fit to systems of points in space
Karl Pearson F.R.S · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach, 2019
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
Attribution and obfuscation of neural text authorship: A data mining perspective
Adaku Uchendu, Thai Le, and Dongwon Lee · 1931
Earlier work this paper cites.
Growth rates of euclidean minimal spanning trees with power weighted edges
J Michael Steele · 1988
Earlier work this paper cites.
The framed morse complex and its invariants
Serguei Barannikov · 1994
Earlier work this paper cites.
Greedy function approximation: A gradient boosting machine
Jerome H. Friedman · 2000
Earlier work this paper cites.
Greedy function approximation: A gradient boosting machine
Jerome H. Friedman · 2001
Earlier work this paper cites.
Stochastic gradient boosting
Jerome H. Friedman · 2002
Earlier work this paper cites.
Sample complexity of testing the manifold hypothesis
Hariharan Narayanan and Sanjoy K. Mitter · 2010
Earlier work this paper cites.
Towards automatic boundary detection for human-ai hybrid essay in education
Zijie Zeng, Lele Sha, Yuheng Li, Kaixun Yang, Dragan Gašević, and Guanliang Chen · 2010
Earlier work this paper cites.
Fast global alignment kernels
Marco Cuturi · 2011
Earlier work this paper cites.
Randomly weighted d − d- complexes: Minimal spanning acycles and persistence diagrams
Primoz Skraba, Gugan Thoppe, and D Yogeshwaran · 2017
Earlier work this paper cites.
Intrinsic Dimensionality Estimation within Tight Localities
Laurent Amsaleg, Oussama Chelly, Michael E Houle, Ken-Ichi Kawarabayashi, Miloš Radovanović, and Weeris Treeratanajaru · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
An introduction to a new text classification and visualization for natural language processing using topological data analysis, 2019
Naiereh Elyasi and Mehdi Hosseini Moghadam · 2019
Earlier work this paper cites.
CTRL - A Conditional Transformer Language Model for Controllable Generation
Nitish Shirish Keskar, Bryan McCann, Lav Varshney, Caiming Xiong, and Richard Socher · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych · 2019
Earlier work this paper cites.
Release strategies and the social impacts of language models
Irene Solaiman, Miles Brundage, Jack Clark, Amanda Askell, Ariel Herbert-Voss, Jeff Wu, Alec Radford, Gretchen Krueger, Jong Wook Kim, Sarah Kreps, et al · 2019
Cited alongside, same era.
A fractal dimension for measures via persistent homology
Henry Adams, Manuchehr Aminian, Elin Farnell, Michael Kirby, Joshua Mirth, Rachel Neville, Chris Peterson, and Clayton Shonkwiler · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Cited alongside, same era.
Persistent homology and applied homotopy theory
Gunnar Carlsson · 2020
Cited alongside, same era.
Electra: Pre-training text encoders as discriminators rather than generators
Ensemble pre-trained transformer models for writing style change detection
Tzu-Mi Lin, Chao-Yi Chen, Yu-Wen Tzeng, and Lung-Hao Lee · 2022
Later among the works it cites.
Exploring the use of topological data analysis to automatically detect data quality faults
M. Eduard Tudoreanu · 2022
Later among the works it cites.
Phi-2: The surprising power of small language models
Marah Abdin, Jyoti Aneja, Sebastien Bubeck, Caio César Teodoro Mendes, Weizhu Chen, Allie Del Giorno, Ronen Eldan, Sivakanth Gopi, Suriya Gunasekar, Mojan Javaheripi, Piero Kauffmann, Yin Tat Lee, Yuanzhi Li, Anh Nguyen, Gustavo de Rosa, Olli Saarikivi, Adil Salim, Shital Shah, Michael Santacroce, Harkirat Singh Behl, Adam Taumann Kalai, Xin Wang, Rachel Ward, Philipp Witte, Cyril Zhang, and Yi Zhang · 2023
Closest in time.
Real or fake text? investigating human ability to detect boundaries between human-written and machine-generated text
Liam Dugan, Daphne Ippolito, Arun Kirubarajan, Sherry Shi, and Chris Callison-Burch · 2023
Closest in time.
Textbooks are all you need, June 2023
Suriya Gunasekar, Yi Zhang, Jyoti Aneja, Caio Cesar, Teodoro Mendes, Allie Del Giorno, Sivakanth Gopi, Mojan Javaheripi, Piero Kauffmann, Gustavo de Rosa, Olli Saarikivi, Adil Salim, Shital Shah, Harkirat Singh Behl, Xin Wang, Sébastien Bubeck, Ronen Eldan, Adam Tauman Kalai, Yin Tat Lee, and Yuanzhi Li · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning · 2020
Cited alongside, same era.
RoFT: A tool for evaluating human detection of machine-generated text
Liam Dugan, Daphne Ippolito, Arun Kirubarajan, and Chris Callison-Burch · 2020
Cited alongside, same era.
Style change detection using bert
Aarish Iyer and Soroush Vosoughi · 2020
Cited alongside, same era.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Cited alongside, same era.
Fractal dimension and the persistent homology of random geometric complexes
Benjamin Schweinhart · 2020
Cited alongside, same era.
Intrinsic dimension, persistent homology and generalization in neural networks
Tolga Birdal, Aaron Lou, Leonidas J Guibas, and Umut Simsekli · 2021
Cited alongside, same era.
All that’s ‘human’ is not gold: Evaluating human evaluation of generated text
Elizabeth Clark, Tal August, Sofia Serrano, Nikita Haduong, Suchin Gururangan, and Noah A. Smith · 2021
Cited alongside, same era.
Closest in time.
Deep neural networks architectures from the perspective of manifold learning
German Magai · 2023
Closest in time.
Detectgpt: Zero-shot machine-generated text detection using probability curvature, 2023
Eric Mitchell, Yoonho Lee, Alexander Khazatsky, Christopher D. Manning, and Chelsea Finn · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models, 2023
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, Dan Bikel, Lukas Blecher, Cristian Canton Ferrer, Moya Chen, Guillem Cucurull, David Esiobu, Jude Fernandes, Jeremy Fu, Wenyin Fu, Brian Fuller, Cynthia Gao, Vedanuj Goswami, Naman Goyal, Anthony Hartshorn, Saghar Hosseini, Rui Hou, Hakan Inan, Marcin Kardas, Viktor Kerkez, Madian Khabsa, Isabel Kloumann, Artem Korenev, Punit Singh Koura, Marie-Anne Lachaux, Thibaut Lavril, Jenya Lee, Diana Liskovich, Yinghai Lu, Yuning Mao, Xavier Martinet, Todor Mihaylov, Pushkar Mishra, Igor Molybog, Yixin Nie, Andrew Poulton, Jeremy Reizenstein, Rashi Rungta, Kalyan Saladi, Alan Schelten, Ruan Silva, Eric Michael Smith, Ranjan Subramanian, Xiaoqing Ellen Tan, Binh Tang, Ross Taylor, Adina Williams, Jian Xiang Kuan, Puxin Xu, Zheng Yan, Iliyan Zarov, Yuchen Zhang, Angela Fan, Melanie Kambadur, Sharan Narang, Aurelien Rodriguez, Robert Stojnic, Sergey Edunov, and Thomas Scialom · 2023
Closest in time.
Topological Data Analysis for Speech Processing
Eduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva, Daniil Cherniavskii, Serguei Barannikov, Irina Piontkovskaya, Sergey Nikolenko, and Evgeny Burnaev · 2023
Closest in time.
Gpt-who: An information density-based machine-generated text detector
Saranya Venkatraman, Adaku Uchendu, and Dongwon Lee · 2023
Closest in time.
SeqXGPT: Sentence-level AI-generated text detection
Pengyu Wang, Linyang Li, Ke Ren, Botian Jiang, Dong Zhang, and Xipeng Qiu · 2023
Closest in time.
Testing of detection tools for ai-generated text
Debora Weber-Wulff, Alla Anohina-Naumeca, Sonja Bjelobaba, Tomáš Foltỳnek, Jean Guerrero-Dib, Olumide Popoola, Petr Šigut, and Lorna Waddington · 2023
Closest in time.
Llmdet: A third party large language models generated text detection tool
Kangxi Wu, Liang Pang, Huawei Shen, Xueqi Cheng, and Tat-Seng Chua · 2023
Closest in time.
A survey on detection of llms-generated content
Xianjun Yang, Liangming Pan, Xuandong Zhao, Haifeng Chen, Linda Petzold, William Yang Wang, and Wei Cheng · 2023
Closest in time.
Semeval-2024 task 8: Weighted layer averaging roberta for black-box machine-generated text detection, 2024
Ayan Datta, Aryan Chandramania, and Radhika Mamidi · 2024
Closest in time.
Rfbes at semeval-2024 task 8: Investigating syntactic and semantic features for distinguishing ai-generated and human-written texts, 2024
Mohammad Heydari Rad, Farhan Farsi, Shayan Bali, Romina Etezadi, and Mehrnoush Shamsfard · 2024
Closest in time.
Kinit at semeval-2024 task 8: Fine-tuned llms for multilingual machine-generated text detection, 2024
Michal Spiegel and Dominik Macko · 2024
Closest in time.