Fetching the paper…
Reading the bibliography…
The widespread usage of online Large Language Models (LLMs) inference services has raised significant privacy concerns about the potential exposure of private information in user inputs to malicious eavesdroppers.
Bert rediscovers the classical nlp pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 1905
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Mechanism design via differential privacy
Frank McSherry and Kunal Talwar. 2007 · 2007
Earlier work this paper cites.
Privacy and the Internet Your Expectations and Rights Under the Law
Jasper M C. 2009 · 2009
Earlier work this paper cites.
A differentially private text perturbation method using a regularized mahalanobis metric
Zekun Xu, Abhinav Aggarwal, Oluwaseyi Feyisetan, and Nathanael Teissier. 2020 · 2010
Earlier work this paper cites.
Broadening the scope of differential privacy using metrics
Konstantinos Chatzikokolakis, Miguel E. Andrés, Nicolás Emilio Bordenabe, and Catuscia Palamidessi. 2013 · 2013
Earlier work this paper cites.
Local privacy and statistical minimax rates
John C. Duchi, Michael I. Jordan, and Martin J. Wainwright. 2013 · 2013
Earlier work this paper cites.
Privacy notices versus informational self-determination: Minding the gap
B Van Alsenoy, E Kosta, and J Dumortier. 2014 · 2014
Earlier work this paper cites.
Understanding intermediate layers using linear classifier probes
Guillaume Alain and Yoshua Bengio. 2016 · 2016
Earlier work this paper cites.
Calibrating noise to sensitivity in private data analysis
Cynthia Dwork, Frank McSherry, Kobbi Nissim, and Adam D. Smith. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Invited paper: Local differential privacy on metric spaces: Optimizing the trade-off with utility
Mário S. Alvim, Konstantinos Chatzikokolakis, Catuscia Palamidessi, and Anna Pazii. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Leveraging hierarchical representations for preserving privacy and utility in text
Oluwaseyi Feyisetan, Tom Diethe, and Thomas Drake. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Earlier work this paper cites.
Privacy-and utility-preserving textual analysis via calibrated multivariate perturbations
Oluwaseyi Feyisetan, Borja Balle, Thomas Drake, and Tom Diethe. 2020 · 2020
Earlier work this paper cites.
ER-AE: differentially private text generation for authorship anonymization
Haohan Bo, Steven H. H. Ding, Benjamin C. M. Fung, and Farkhund Iqbal. 2021 · 2021
Earlier work this paper cites.
Extracting training data from large language models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, Dawn Song, Úlfar Erlingsson, Alina Oprea, and Colin Raffel. 2021 · 2021
Cited alongside, same era.
Probing classifiers: Promises, shortcomings, and advances
Yonatan Belinkov. 2022 · 2022
Cited alongside, same era.
You don’t know my favorite color: Preventing dialogue representations from revealing speakers’ private personas
Haoran Li, Yangqiu Song, and Lixin Fan. 2022 · 2022
Cited alongside, same era.
Differentially private language models for secure data sharing
Justus Mattern, Zhijing Jin, Benjamin Weggenmann, Bernhard Schoelkopf, and Mrinmaya Sachan. 2022a · 2022
Cited alongside, same era.
The limits of word level differential privacy
Justus Mattern, Benjamin Weggenmann, and Florian Kerschbaum. 2022b · 2022
Cited alongside, same era.
Ddxplus: A new dataset for automatic medical diagnosis
Activation addition: Steering language models without optimization
Alexander Matt Turner, Lisa Thiergart, David Udell, Gavin Leech, Ulisse Mini, and Monte MacDiarmid. 2023 · 2023
Later among the works it cites.
Locally differentially private document generation using zero shot prompting
Saiteja Utpala, Sara Hooker, and Pin-Yu Chen. 2023b · 2023
Later among the works it cites.
Bloomberggpt: A large language model for finance
Shijie Wu, Ozan Irsoy, Steven Lu, Vadim Dabravolski, Mark Dredze, Sebastian Gehrmann, Prabhanjan Kambadur, David Rosenberg, and Gideon Mann. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica. 2023a · 2023
Later among the works it cites.
Primer: Fast private transformer inference on encrypted data
Mengxin Zheng, Qian Lou, and Lei Jiang. 2023b · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Arsène Fansi Tchango, Rishab Goel, Zhi Wen, Julien Martel, and Joumana Ghosn. 2022 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Cited alongside, same era.
Nlice: Synthetic medical record generation for effective primary healthcare differential diagnosis
Zaid Al-Ars, Obinna Agba, Zhuoran Guo, Christiaan Boerkamp, Ziyaad Jaber, and Tareq Jaber. 2023 · 2023
Cited alongside, same era.
Syllogistic reasoning for legal judgment analysis
Wentao Deng, Jiahuan Pei, Keyi Kong, Zhe Chen, Furu Wei, Yujun Li, Zhaochun Ren, Zhumin Chen, and Pengjie Ren. 2023 · 2023
Cited alongside, same era.
Sigma: Secure gpt inference with function secret sharing
Kanav Gupta, Neha Jawalkar, Ananta Mukherjee, Nishanth Chandran, Divya Gupta, Ashish Panwar, and Rahul Sharma. 2023 · 2023
Cited alongside, same era.
Inspecting and editing knowledge representations in language models
Evan Hernandez, Belinda Z. Li, and Jacob Andreas. 2023 · 2023
Cited alongside, same era.
Zhigang Kan, Linbo Qiao, Hao Yu, Liwen Peng, Yifu Gao, and Dongsheng Li. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
UnitedHealth says data of 100 million stolen in Change Healthcare breach
Lawrence Abrams. 2024 · 2024
Closest in time.
Truth forest: Toward multi-scale truthfulness in large language models through intervention without tuning
Zhongzhi Chen, Xingwu Sun, Xianfeng Jiao, Fengzong Lian, Zhanhui Kang, Di Wang, and Chengzhong Xu. 2024 · 2024
Closest in time.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Alex Castro-Ros, Marie Pellat, Kevin Robinson, Dasha Valter, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Y. Zhao, Yanping Huang, Andrew M. Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2024 · 2024
Closest in time.
Ahmed Frikha, Nassim Walha, Krishna Kanth Nakka, Ricardo Mendes, Xue Jiang, and Xuebing Zhou. 2024 · 2024
Closest in time.
ORPO: monolithic preference optimization without reference model
Jiwoo Hong, Noah Lee, and James Thorne. 2024 · 2024
Closest in time.
Merge: Fast private text generation
Zi Liang, Pinghui Wang, Ruofei Zhang, Nuo Xu, Shuo Zhang, Lifeng Xing, Haitao Bai, and Ziyang Zhou. 2024 · 2024
Closest in time.
Zero-shot event argument extraction by disentangling trigger from argument and role
Zhengdong Lu, Ziqian Zeng, Jianwei Wang, Hanlin Wang, Weikai Lu, and Huiping Zhuang. 2024 · 2024
Closest in time.
Large language models are advanced anonymizers
Robin Staab, Mark Vero, Mislav Balunovic, and Martin T. Vechev. 2024 · 2024
Closest in time.
Xuchen Suo. 2024 · 2024
Closest in time.
Anna Jaques Hospital ransomware breach exposed data of 300K patients
Bill Toulas. 2024 · 2024
Closest in time.
An Yang, Baosong Yang, Binyuan Hui, Bo Zheng, Bowen Yu, Chang Zhou, Chengpeng Li, Chengyuan Li, Dayiheng Liu, Fei Huang, Guanting Dong, Haoran Wei, Huan Lin, Jialong Tang, Jialin Wang, Jian Yang, Jianhong Tu, Jianwei Zhang, Jianxin Ma, Jin Xu, Jingren Zhou, Jinze Bai, Jinzheng He, Junyang Lin, Kai Dang, Keming Lu, Keqin Chen, Kexin Yang, Mei Li, Mingfeng Xue, Na Ni, Pei Zhang, Peng Wang, Ru Peng, Rui Men, Ruize Gao, Runji Lin, Shijie Wang, Shuai Bai, Sinan Tan, Tianhang Zhu, Tianhao Li, Tianyu Liu, Wenbin Ge, Xiaodong Deng, Xiaohuan Zhou, Xingzhang Ren, Xinyu Zhang, Xipin Wei, Xuancheng Ren, Yang Fan, Yang Yao, Yichang Zhang, Yu Wan, Yunfei Chu, Yuqiong Liu, Zeyu Cui, Zhenru Zhang, and Zhihao Fan. 2024 · 2024
Closest in time.
Truthx: Alleviating hallucinations by editing large language models in truthful space
Shaolei Zhang, Tian Yu, and Yang Feng. 2024 · 2024
Closest in time.