Fetching the paper…
Reading the bibliography…
The rise in malicious usage of large language models, such as fake content creation and academic plagiarism, has motivated the development of approaches that identify AI-generated text, including those based on watermarking or outlier detection.
A general method applicable to the search for similarities in the amino acid sequence of two proteins
S. B. Needleman and C. D. Wunsch · 1970
Earlier work this paper cites.
Okapi at trec-3
S. E. Robertson, S. Walker, S. Jones, M. M. Hancock-Beaulieu, M. Gatford, et al · 1995
Earlier work this paper cites.
Neural syntactic preordering for controlled paraphrase generation
T. Goyal and G. Durrett · 2005
Earlier work this paper cites.
Natural language watermarking
M. Topkara, C. M. Taskiran, and E. J. Delp III · 2005
Earlier work this paper cites.
Text and context in translation
J. House · 2006
Earlier work this paper cites.
Context sensitive paraphrasing with a global unsupervised classifier
M. Connor and D. Roth · 2007
Earlier work this paper cites.
Identifying real or fake articles: Towards better language modeling
S. Badaskar, S. Agarwal, and S. Arora · 2008
Earlier work this paper cites.
Detecting fake content with relative entropy scoring
T. Lavergne, T. Urvoy, and F. Yvon · 2008
Earlier work this paper cites.
Detection of artificial texts
E. Grechnikov, G. Gusev, A. Kustarev, and A. Raigorodsky · 2009
Earlier work this paper cites.
Sub-sentencial paraphrasing by contextual pivot translation
A. Max · 2009
Earlier work this paper cites.
Natural language watermarking via morphosyntactic alterations
H. M. Meral, B. Sankur, A. S. Özsoy, T. Güngör, and E. Sevinç · 2009
Earlier work this paper cites.
Squibs: What is a paraphrase?
R. Bhagat and E. Hovy · 2013
Earlier work this paper cites.
Creating a billion-scale searchable web archive
D. Gomes, M. Costa, D. Cruz, J. Miranda, and S. Fontes · 2013
Earlier work this paper cites.
Problems in current text simplification research: New data can help
W. Xu, C. Callison-Burch, and C. Napoles · 2015
Earlier work this paper cites.
SemEval-2016 task 1: Semantic textual similarity, monolingual and cross-lingual evaluation
E. Agirre, C. Banea, D. Cer, M. Diab, A. Gonzalez-Agirre, R. Mihalcea, G. Rigau, and J. Wiebe · 2016
Earlier work this paper cites.
Computer-generated text detection using machine learning: A systematic review
D. Beresneva · 2016
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang · 2016
Earlier work this paper cites.
Does neural machine translation benefit from larger context?
S. Jean, S. Lauly, O. Firat, and K. Cho · 2017
Earlier work this paper cites.
Pointer sentinel mixture models
S. Merity, C. Xiong, J. Bradbury, and R. Socher · 2017
Earlier work this paper cites.
Membership inference attacks against machine learning models
R. Shokri, M. Stronati, C. Song, and V. Shmatikov · 2017
Earlier work this paper cites.
Neural machine translation with extended context
J. Tiedemann and Y. Scherrer · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Exploiting cross-sentence context for neural machine translation
L. Wang, Z. Tu, A. Way, and Q. Liu · 2017
Earlier work this paper cites.
Contextual handling in neural machine translation: Look behind, ahead and on both sides
R. R. Agrawal, M. Turchi, and M. Negri · 2018
Earlier work this paper cites.
JAX: composable transformations of Python+NumPy programs, 2018
J. Bradbury, R. Frostig, P. Hawkins, M. J. Johnson, C. Leary, D. Maclaurin, G. Necula, A. Paszke, J. VanderPlas, S. Wanderman-Milne, and Q. Zhang · 2018
Earlier work this paper cites.
Adversarial example generation with syntactically controlled paraphrase networks
M. Iyyer, J. Wieting, K. Gimpel, and L. Zettlemoyer · 2018
Earlier work this paper cites.
Modeling coherence for neural machine translation with dynamic and topic caches
S. Kuang, D. Xiong, W. Luo, and G. Zhou · 2018
Earlier work this paper cites.
Paraphrase generation with deep reinforcement learning
Z. Li, X. Jiang, L. Shang, and H. Li · 2018
Earlier work this paper cites.
Document-level neural machine translation with hierarchical attention networks
L. Miculicich, D. Ram, N. Pappas, and J. Henderson · 2018
Earlier work this paper cites.
Context-aware neural machine translation learns anaphora resolution
E. Voita, P. Serdyukov, R. Sennrich, and I. Titov · 2018
Earlier work this paper cites.
ParaNMT-50M: Pushing the limits of paraphrastic sentence embeddings with millions of machine translations
J. Wieting and K. Gimpel · 2018
Earlier work this paper cites.
Improving the transformer translation model with document-level context
J. Zhang, H. Luan, M. Sun, F. Zhai, J. Xu, M. Zhang, and Y. Liu · 2018
Earlier work this paper cites.
Real or fake? learning to discriminate machine from human generated text
A. Bakhtin, S. Gross, M. Ott, Y. Deng, M. Ranzato, and A. Szlam · 2019
Earlier work this paper cites.
Controllable paraphrase generation with a syntactic exemplar
M. Chen, Q. Tang, S. Wiseman, and K. Gimpel · 2019
Earlier work this paper cites.
Eli5: Long form question answering
A. Fan, Y. Jernite, E. Perez, D. Grangier, J. Weston, and M. Auli · 2019
Earlier work this paper cites.
GLTR: Statistical detection and visualization of generated text
S. Gehrmann, H. Strobelt, and A. Rush · 2019
Earlier work this paper cites.
Parabank: Monolingual bitext generation and sentential paraphrasing via lexically-constrained neural machine translation
J. E. Hu, R. Rudinger, M. Post, and B. Van Durme · 2019
Earlier work this paper cites.
Fill in the blanks: Imputing missing sentences for larger-context neural machine translation
S. Jean, A. Bapna, and O. Firat · 2019
Cited alongside, same era.
Billion-scale similarity search with gpus
J. Johnson, M. Douze, and H. Jégou · 2019
Cited alongside, same era.
Microsoft translator at WMT 2019: Towards large-scale document-level neural machine translation
M. Junczys-Dowmunt · 2019
Cited alongside, same era.
Submodular optimization-based diverse paraphrasing and its effectiveness in data augmentation
A. Kumar, S. Bhattamishra, M. Bhandari, and P. Talukdar · 2019
Cited alongside, same era.
Decomposable neural paraphrase generation
Z. Li, X. Jiang, L. Shang, and Q. Liu · 2019
Cited alongside, same era.
Towards document-level paraphrase generation with sentence rewriting and reordering
Z. Lin, Y. Cai, and X. Wan · 2021
Later among the works it cites.
Capturing document context inside sentence-level neural machine translation models with self-training
E. Mansimov, G. Melis, and L. Yu · 2021
Later among the works it cites.
A survey on document-level neural machine translation: Methods and evaluation
S. Maruf, F. Saleh, and G. Haffari · 2021
Later among the works it cites.
ConRPG: Paraphrase generation using contexts as regularizer
Y. Meng, X. Ao, Q. He, X. Sun, Q. Han, F. Wu, C. Fan, and J. Li · 2021
Later among the works it cites.
Beir: A heterogenous benchmark for zero-shot evaluation of information retrieval models
N. Thakur, N. Reimers, A. Rücklé, A. Srivastava, and I. Gurevych · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Popescu-Belis, S. Loáiciga, C. Hardmeier, and D. Xiong, editors · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever · 2019
Cited alongside, same era.
Unsupervised paraphrasing without translation
A. Roy and D. Grangier · 2019
Cited alongside, same era.
Do massively pretrained language models make better storytellers?
A. See, A. Pappu, R. Saxena, A. Yerukola, and C. D. Manning · 2019
Cited alongside, same era.
Beyond BLEU:training neural machine translation with semantic similarity
J. Wieting, T. Berg-Kirkpatrick, K. Gimpel, and G. Neubig · 2019
Cited alongside, same era.
Paraphrasing with large language models
S. Witteveen and M. Andrews · 2019
Cited alongside, same era.
Modeling coherence for discourse neural machine translation
H. Xiong, Z. He, H. Wu, and H. Wang · 2019
Cited alongside, same era.
K. Yin, P. Fernandes, A. F. Martins, and G. Neubig · 2021
Later among the works it cites.
Quality controlled paraphrase generation
E. Bandel, R. Aharonov, M. Shmueli-Scheuer, I. Shnayderman, N. Slonim, and L. Ein-Dor · 2022
Later among the works it cites.
retriv: A user-friendly and efficient search engine in python., 2022
E. Bassani · 2022
Later among the works it cites.
The ethical need for watermarks in machine-generated language
A. Grinbaum and L. Adomaitis · 2022
Later among the works it cites.
Hierarchical sketch induction for paraphrase generation
T. Hosking, H. Tang, and M. Lapata · 2022
Later among the works it cites.
OpenAI Models - GPT3.5, 2022
OpenAI · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Later among the works it cites.
Scaling up models and data with T5X and seqio
A. Roberts, H. W. Chung, A. Levskaya, G. Mishra, J. Bradbury, D. Andor, S. Narang, B. Lester, C. Gaffney, A. Mohiuddin, et al · 2022
Later among the works it cites.
ChatGPT: Optimizing language models for dialogue, 2022
J. Schulman, B. Zoph, C. Kim, J. Hilton, J. Menick, J. Weng, J. Uribe, L. Fedus, L. Metz, et al · 2022
Later among the works it cites.
Ai bot chatgpt writes smart essays-should academics worry?
C. Stokel-Walker · 2022
Later among the works it cites.
Exploring document-level literary machine translation with parallel paragraphs from world literature
K. Thai, M. Karpinska, K. Krishna, W. Ray, M. Inghilleri, J. Wieting, and M. Iyyer · 2022
Later among the works it cites.
Label Studio: Data labeling software, 2020-2022
M. Tkachenko, M. Malyuk, A. Holmanyuk, and N. Liubimov · 2022
Later among the works it cites.
Paraphrastic representations at scale
J. Wieting, K. Gimpel, G. Neubig, and T. Berg-kirkpatrick · 2022
Later among the works it cites.
Multi-task learning for paraphrase generation with keyword and part-of-speech reconstruction
X. Xie, X. Lu, and B. Chen · 2022
Later among the works it cites.
Gcpg: A general framework for controllable paraphrase generation
K. Yang, D. Liu, W. Lei, B. Yang, H. Zhang, X. Zhao, W. Yao, and B. Chen · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
S. Zhang, S. Roller, N. Goyal, M. Artetxe, M. Chen, S. Chen, C. Dewan, M. Diab, X. Li, X. V. Lin, et al · 2022
Later among the works it cites.
On the possibilities of ai-generated text detection
S. Chakraborty, A. S. Bedi, S. Zhu, B. An, D. Manocha, and F. Huang · 2023
Closest in time.
How google search organizes information, 2023
Google · 2023
Closest in time.
R. Koike, M. Kaneko, and N. Okazaki · 2023
Closest in time.
How reliable are ai-generated-text detectors? an assessment framework using evasive soft prompts
T. Kumarage, P. Sheth, R. Moraffah, J. Garland, and H. Liu · 2023
Closest in time.
A semantic invariant robust watermark for large language models
A. Liu, L. Pan, X. Hu, S. Meng, and L. Wen · 2023
Closest in time.
Large language models can be guided to evade ai-generated text detection
N. Lu, S. Liu, R. He, and K. Tang · 2023
Closest in time.
Detectgpt: Zero-shot machine-generated text detection using probability curvature
E. Mitchell, Y. Lee, A. Khazatsky, C. D. Manning, and C. Finn · 2023
Closest in time.
Chatgpt faces ‘growing pains’ as website traffic drops for first time, 2023
NewYorkPost · 2023
Closest in time.
Can sensitive information be deleted from llms? objectives for defending against extraction attacks
V. Patil, P. Hase, and M. Bansal · 2023
Closest in time.
Can ai-generated text be reliably detected?
V. S. Sadasivan, A. Kumar, S. Balasubramanian, W. Wang, and S. Feizi · 2023
Closest in time.
Gptzero: An ai text detector, 2023
E. Tian · 2023
Closest in time.
Llama: Open and efficient foundation language models
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, et al · 2023
Closest in time.
Advancing beyond identification: Multi-bit watermark for language models
K. Yoo, W. Ahn, and N. Kwak · 2023
Closest in time.
Provable robust watermarking for ai-generated text
X. Zhao, P. Ananth, L. Li, and Y.-X. Wang · 2023
Closest in time.
The enemy in your own camp: How well can we detect statistically-generated fake reviews – an adversarial study
D. Hovy · 2057
Closest in time.