Fetching the paper…
Reading the bibliography…
As large language models (LLMs) rapidly advance and integrate into daily life, the privacy risks they pose are attracting increasing attention.
The enron corpus: A new dataset for email classification research
Bryan Klimt and Yiming Yang · 2004
Earlier work this paper cites.
Effects of age and gender on blogging
Jonathan Schler, Moshe Koppel, Shlomo Engelson Argamon, and James W. Pennebaker · 2006
Earlier work this paper cites.
Measuring differentiability: Unmasking pseudonymous authors
Moshe Koppel, Jonathan Schler, and Elisheva Bonchek-Dokow · 2007
Earlier work this paper cites.
On anonymity in an electronic society: A survey of anonymous communication systems
Matthew Edman and Bülent Yener · 2009
Earlier work this paper cites.
A survey of modern authorship attribution methods
Efstathios Stamatatos · 2009
Earlier work this paper cites.
Permllm: Private inference of large language models within 3 seconds under wan
Fei Zheng, Chaochao Chen, Zhongxuan Han, and Xiaolin Zheng · 2009
Earlier work this paper cites.
The effect of anonymity in peer review
Matthew Coomber and Richard Silver · 2010
Earlier work this paper cites.
Authorship attribution with latent dirichlet allocation
Yanir Seroussi, Ingrid Zukerman, and Fabian Bohnert · 2011
Earlier work this paper cites.
On peer review in computer science: Analysis of its effectiveness and suggestions for improvement
Azzurra Ragone, Katsiaryna Mirylenka, Fabio Casati, and Maurizio Marchese · 2013
Earlier work this paper cites.
On the robustness of authorship attribution based on character n-gram features
Efstathios Stamatatos · 2013
Earlier work this paper cites.
Memorization and privacy risks in domain-specific large language models
Xinyu Yang, Zichen Wen, Wenjie Qu, Zhaorun Chen, Zhiying Xiang, Beidi Chen, and Huaxiu Yao · 2013
Earlier work this paper cites.
Anonize: A large-scale anonymous survey system
Susan Hohenberger, Steven Myers, Rafael Pass, et al · 2014
Earlier work this paper cites.
Author identification using multi-headed recurrent neural networks
Douglas Bagnall · 2015
Earlier work this paper cites.
Sebastian Ruder, Parsa Ghaffari, and John Gerard Breslin · 2016
Earlier work this paper cites.
Surveying stylometry techniques and applications
Tempestt Neal, Kalaivani Sundararajan, Aneez Fatima, Yiming Yan, Yingfei Xiang, and Damon Woodard · 2017
Earlier work this paper cites.
Authorship Attribution with Topic Models
Yanir Seroussi, Ingrid Zukerman, and Fabian Bohnert · 2017
Earlier work this paper cites.
Multi-classifier system for authorship verification task using word embeddings
Nacer Eddine Benzebouchi, Nabiha Azizi, Monther Aldwairi, and Nadir Farah · 2018
Earlier work this paper cites.
Abhay Sharma, Ananya Nandan, and Reetika Ralhan · 2018
Earlier work this paper cites.
What represents “style” in authorship attribution?
Kalaivani Sundararajan and D. Woodard · 2018
Cited alongside, same era.
An empirical review of anonymity effects in peer assessment, peer feedback, peer review, peer evaluation and peer grading
Ernesto Panadero and Maryam Alqassab · 2019
Cited alongside, same era.
Cross-domain authorship attribution using pre-trained language models
Georgios Barlas and Efstathios Stamatatos · 2020
Cited alongside, same era.
Bertaa : Bert fine-tuning for authorship attribution
Mael Fabien, Esaú Villatoro-Tello, Petr Motlícek, and Shantipriya Parida · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Cited alongside, same era.
Learning universal authorship representations
Scalable extraction of training data from (production) language models
Milad Nasr, Nicholas Carlini, Jonathan Hayase, Matthew Jagielski, A Feder Cooper, Daphne Ippolito, Christopher A Choquette-Choo, Eric Wallace, Florian Tramèr, and Katherine Lee · 2023
Later among the works it cites.
Anonymity at risk? assessing re-identification capabilities of large language models
Alex Nyffenegger, Matthias Stürmer, and Joel Niklaus · 2023
Later among the works it cites.
Beyond memorization: Violating privacy via inference with large language models
Robin Staab, Mark Vero, Mislav Balunović, and Martin Vechev · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rafael A. Rivera-Soto, Olivia Elizabeth Miano, Juanita Ordonez, Barry Y. Chen, Aleem Khan, Marcus Bishop, and Nicholas Andrews · 2021
Cited alongside, same era.
Part: Pre-trained authorship representation transformer
Javier Huertas-Tato, Álvaro Huertas-García, Alejandro Martín, and David Camacho · 2022
Cited alongside, same era.
Overview of the authorship verification task at PAN 2022
Efstathios Stamatatos, Mike Kestemont, Krzysztof Kredens, Piotr Pezik, Annina Heini, Janek Bevendorff, Benno Stein, and Martin Potthast · 2022
Cited alongside, same era.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Instruction-tuning aligns llms to the human brain
Khai Loong Aw, Syrielle Montariol, Badr AlKhamissi, Martin Schrimpf, and Antoine Bosselut · 2023
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Cited alongside, same era.
Rui Wen, Tianhao Wang, Michael Backes, Yang Zhang, and Ahmed Salem · 2023
Later among the works it cites.
Baichuan 2: Open large-scale language models
Aiyuan Yang, Bin Xiao, Bingning Wang, Borong Zhang, Ce Bian, Chao Yin, Chenxu Lv, Da Pan, Dian Wang, Dong Yan, et al · 2023
Later among the works it cites.
A comprehensive capability analysis of gpt-3 and gpt-3.5 series models
Junjie Ye, Xuanting Chen, Nuo Xu, Can Zu, Zekai Shao, Shichun Liu, Yuhan Cui, Zeyang Zhou, Chao Gong, Yang Shen, et al · 2023
Later among the works it cites.
Instruction tuning for large language models: A survey
Shengyu Zhang, Linfeng Dong, Xiaoya Li, Sen Zhang, Xiaofei Sun, Shuhe Wang, Jiwei Li, Runyi Hu, Tianwei Zhang, Fei Wu, et al · 2023
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
The claude 3 model family: Opus, sonnet, haiku
AI Anthropic · 2024
Closest in time.
Overview of PAN 2024: multi-author writing style analysis, multilingual text detoxification, oppositional thinking analysis, and generative ai authorship verification
Janek Bevendorff, Xavier Bonet Casals, Berta Chulvi, Daryna Dementieva, Ashaf Elnagar, Dayne Freitag, Maik Fröbe, Damir Korenčić, Maximilian Mayerl, Animesh Mukherjee, et al · 2024
Closest in time.
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, et al · 2024
Closest in time.
PANDORA: Detailed LLM jailbreaking via collaborated phishing agents with decomposed reasoning
Zhaorun Chen, Zhuokai Zhao, Wenjie Qu, Zichen Wen, Zhiguang Han, Zhihong Zhu, Jiaheng Zhang, and Huaxiu Yao · 2024
Closest in time.
Llm maybe longlm: Self-extend llm context window without tuning
Hongye Jin, Xiaotian Han, Jingfeng Yang, Zhimeng Jiang, Zirui Liu, Chia-Yuan Chang, Huiyuan Chen, and Xia Hu · 2024
Closest in time.
Propile: Probing privacy leakage in large language models
Siwon Kim, Sangdoo Yun, Hwaran Lee, Martin Gubri, Sungroh Yoon, and Seong Joon Oh · 2024
Closest in time.
Introducing meta llama 3: The most capable openly available llm to date
AI Meta · 2024
Closest in time.
Extending llms’ context window with 100 samples
Yikai Zhang, Junlong Li, and Pengfei Liu · 2024
Closest in time.
Retrieval-augmented generation for ai-generated content: A survey
Penghao Zhao, Hailin Zhang, Qinhan Yu, Zhengren Wang, Yunteng Geng, Fangcheng Fu, Ling Yang, Wentao Zhang, and Bin Cui · 2024
Closest in time.