Fetching the paper…
Reading the bibliography…
Consider a scenario where an author-e.g., activist, whistle-blower, with many public writings wishes to write "anonymously" when attackers may have already built an authorship attribution (AA) model based off of public writings including those of the author.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
BERTScore: Evaluating Text Generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2019 · 1904
Earlier work this paper cites.
Introducing the Enron corpus.. In CEAS
Bryan Klimt and Yiming Yang. 2004 · 2004
Earlier work this paper cites.
A framework for authorship identification of online messages: Writing-style features and classification techniques
Rong Zheng, Jiexun Li, Hsinchun Chen, and Zan Huang. 2006 · 2006
Earlier work this paper cites.
Authorship attribution of e-mail: Comparing classifiers over a new corpus for evaluation. In Proceedings of the sixth International Conference on Language Resources and Evaluation (LREC’08)
Ben Allison and Louise Guthrie. 2008 · 2008
Earlier work this paper cites.
Authorship attribution
Patrick Juola et al · 2008
Earlier work this paper cites.
E-Mail authorship attribution applied to the Extended Enron Authorship Corpus (XEAC)
Hendrik Neumann and Martin Schnurrenberger. 2009 · 2009
Earlier work this paper cites.
Stylometric analysis for authorship attribution on twitter. In Big Data Analytics: Second International Conference, BDA 2013, Mysore, India, December 16-18, 2013, Proceedings 2 . Springer, 37–47
Mudit Bhargava, Pulkit Mehndiratta, and Krishna Asawa. 2013 · 2013
Earlier work this paper cites.
Authorship attribution with topic models
Yanir Seroussi, Ingrid Zukerman, and Fabian Bohnert. 2014 · 2014
Earlier work this paper cites.
Measuring individuals’ concerns over collective privacy on social networking sites
Haiyan Jia and Heng Xu. 2016 · 2016
Earlier work this paper cites.
Authorship attribution using stylometry and machine learning techniques
Hoshiladevi Ramnial, Shireen Panchoo, and Sameerchand Pudaruth. 2016 · 2016
Earlier work this paper cites.
Authorship attribution for social media forensics
Anderson Rocha, Walter J Scheirer, Christopher W Forstall, Thiago Cavalcante, Antonio Theophilo, Bingyu Shen, Ariadne RB Carvalho, and Efstathios Stamatatos. 2016 · 2016
Earlier work this paper cites.
Overview of the Author Obfuscation Task at PAN 2017: Safety Evaluation Revisited.. In CLEF (Working Notes)
Matthias Hagen, Martin Potthast, and Benno Stein. 2017 · 2017
Earlier work this paper cites.
Daniel Cer, Yinfei Yang, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, et al · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Clean-label backdoor attacks
Alexander Turner, Dimitris Tsipras, and Aleksander Madry. 2018 · 2018
Earlier work this paper cites.
Syntax encoding with application in authorship attribution. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . 2742–2753
Richong Zhang, Zhiyuan Hu, Hongyu Guo, and Yongyi Mao. 2018 · 2018
Earlier work this paper cites.
Authorship Attribution in Fan-Fictional Texts given variable length Character and Word N-Grams
Janek Amann. 2019 · 2019
Earlier work this paper cites.
Cross-Domain Authorship Attribution Combining Instance Based and Profile-Based Features.. In CLEF (Working Notes)
Andrea Bacciu, Massimo La Morgia, Alessandro Mei, Eugenio Nerio Nemmi, Valerio Neri, and Julinda Stefa. 2019 · 2019
Cited alongside, same era.
Explainable authorship verification in social media via attention-based similarity learning. In 2019 IEEE International Conference on Big Data (Big Data) . IEEE, 36–45
Benedikt Boenninghoff, Steffen Hessler, Dorothea Kolossa, and Robert M Nickel. 2019 · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Cited alongside, same era.
A Girl Has No Name: Automated Authorship Obfuscation using Mutant-X
Asad Mahmood, Faizan Ahmad, Zubair Shafiq, Padmini Srinivasan, and Fareed Zaffar. 2019 · 2019
Cited alongside, same era.
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
LowKey: Leveraging Adversarial Attacks to Protect Social Media Users from Facial Recognition
Valeriia Cherepanova, Micah Goldblum, Harrison Foley, Shiyuan Duan, John Dickerson, Gavin Taylor, and Tom Goldstein. 2021 · 2021
Later among the works it cites.
Adversarial stylometry in the wild: Transferable lexical substitution attacks on author profiling
Chris Emmery, Ákos Kádár, and Grzegorz Chrupała. 2021 · 2021
Later among the works it cites.
Adversarial examples make strong poisons
Liam Fowl, Micah Goldblum, Ping-yeh Chiang, Jonas Geiping, Wojtek Czaja, and Tom Goldstein. 2021 · 2021
Later among the works it cites.
Avengers Ensemble! Improving Transferability of Authorship Obfuscation
Muhammad Haroon, Fareed Zaffar, Padmini Srinivasan, and Zubair Shafiq. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 2019
Cited alongside, same era.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Wei Zhao, Maxime Peyrard, Fei Liu, Yang Gao, Christian M Meyer, and Steffen Eger. 2019 · 2019
Cited alongside, same era.
Poison attacks against text datasets with conditional adversarially regularized autoencoder
Alvin Chan, Yi Tay, Yew-Soon Ong, and Aston Zhang. 2020 · 2020
Cited alongside, same era.
Foggysight: A scheme for facial lookup privacy
Ivan Evtimov, Pascal Sturmfels, and Tadayoshi Kohno. 2020 · 2020
Cited alongside, same era.
BertAA: BERT fine-tuning for Authorship Attribution. In Proceedings of the 17th International Conference on Natural Language Processing (ICON) . 127–137
Maël Fabien, Esaú Villatoro-Tello, Petr Motlicek, and Shantipriya Parida. 2020 · 2020
Cited alongside, same era.
NeuSpell: A Neural Spelling Correction Toolkit. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations . Association for Computational Linguistics, Online, 158–164
Sai Muralidhar Jayanthi, Danish Pruthi, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Is bert really robust? a strong baseline for natural language attack on text classification and entailment. In Proceedings of the AAAI conference on artificial intelligence , Vol. 34. 8018–8025
Di Jin, Zhijing Jin, Joey Tianyi Zhou, and Peter Szolovits. 2020 · 2020
Cited alongside, same era.
Weight poisoning attacks on pre-trained models
Keita Kurita, Paul Michel, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Xuanli He, Lingjuan Lyu, Qiongkai Xu, and Lichao Sun. 2021 · 2021
Later among the works it cites.
Unlearnable examples: Making personal data unexploitable
Hanxun Huang, Xingjun Ma, Sarah Monazam Erfani, James Bailey, and Yisen Wang. 2021 · 2021
Later among the works it cites.
Uncovering the connections between adversarial transferability and knowledge transferability. In International Conference on Machine Learning . PMLR, 6577–6587
Kaizhao Liang, Jacky Y Zhang, Boxin Wang, Zhuolin Yang, Sanmi Koyejo, and Bo Li. 2021 · 2021
Later among the works it cites.
Exploring Data and Model Poisoning Attacks to Deep Learning-Based NLP Systems
Fiammetta Marulli, Laura Verde, and Lelio Campanile. 2021 · 2021
Later among the works it cites.
Authorship attribution of social media messages
Antonio Theophilo, Romain Giot, and Anderson Rocha. 2021 · 2021
Later among the works it cites.
Turingbench: A benchmark environment for turing test in the age of neural text generation
Adaku Uchendu, Zeyu Ma, Thai Le, Rui Zhang, and Dongwon Lee. 2021 · 2021
Later among the works it cites.
Wenkai Yang, Lei Li, Zhiyuan Zhang, Xuancheng Ren, Xu Sun, and Bin He. 2021 · 2021
Later among the works it cites.
Authorship Attribution with Temporal Data in Reddit. In XVIII Brazilian Symposium on Information Systems . 1–8
Guilherme Ramos Casimiro and Luciano Antonio Digiampietri. 2022 · 2022
Closest in time.
Dataset security for machine learning: Data poisoning, backdoor attacks, and defenses
Micah Goldblum, Dimitris Tsipras, Chulin Xie, Xinyun Chen, Avi Schwarzschild, Dawn Song, Aleksander Madry, Bo Li, and Tom Goldstein. 2022 · 2022
Closest in time.
Neuspell defense against adversarial character attacks
Sai Muralidhar Jayanthi, Danish Pruthi, and Graham Neubig. 2022 · 2022
Closest in time.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Closest in time.
Explainable Artificial Intelligence for Authorship Attribution on Social Media. In ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2909–2913
Antonio Theophilo, Rafael Padilha, Fernanda A Andaló, and Anderson Rocha. 2022 · 2022
Closest in time.
Deepfake Text Detection in the Wild
Yafu Li, Qintong Li, Leyang Cui, Wei Bi, Longyue Wang, Linyi Yang, Shuming Shi, and Yue Zhang. 2023 · 2023
Closest in time.
Attribution and Obfuscation of Neural Text Authorship: A Data Mining Perspective
Adaku Uchendu, Thai Le, and Dongwon Lee. 2023 · 2023
Closest in time.