Fetching the paper…
Reading the bibliography…
Large language models (LLMs) are complex artificial intelligence systems capable of understanding, generating and translating human language.
O. Goldreich, “Secure multi-party computation,” Manuscript. Preliminary version , vol. 78, no. 110, pp. 1–108, 1998
1998
Earlier work this paper cites.
E. Boyle, N. Gilboa, and Y. Ishai, “Function secret sharing,” in Annual international conference on the theory and applications of cryptographic techniques . Springer, 2015, pp. 337–367
2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
R. Shokri, M. Stronati, C. Song, and V. Shmatikov, “Membership inference attacks against machine learning models,” in 2017 IEEE symposium on security and privacy (SP) . IEEE, 2017, pp. 3–18
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Acar, H. Aksu, A. S. Uluagac, and M. Conti, “A survey on homomorphic encryption schemes: Theory and implementation,” ACM Computing Surveys (Csur) , vol. 51, no. 4, pp. 1–35, 2018
2018
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Shvartzshnaider, Z. Pavlinovic, A. Balashankar, T. Wies, L. Subramanian, H. Nissenbaum, and P. Mittal, “Vaccine: Using contextual integrity for data leakage detection,” in The World Wide Web Conference , 2019, pp. 1702–1712
2019
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” 2020
2020
Earlier work this paper cites.
C. Song and A. Raghunathan, “Information leakage in embedding models,” in Proceedings of the 2020 ACM SIGSAC conference on computer and communications security , 2020, pp. 377–390
2020
Earlier work this paper cites.
X. Pan, M. Zhang, S. Ji, and M. Yang, “Privacy risks of general-purpose language models,” in 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 2020, pp. 1314–1331
2020
Earlier work this paper cites.
J. Zhu, R. Hou, X. Wang, W. Wang, J. Cao, B. Zhao, Z. Wang, Y. Zhang, J. Ying, L. Zhang et al. , “Enabling rack-scale confidential computing using heterogeneous trusted execution environment,” in 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 2020, pp. 1450–1465
2020
Earlier work this paper cites.
K. Liu, M. Yang, Z. Ling, H. Yan, Y. Zhang, X. Fu, and W. Zhao, “On manually reverse engineering communication protocols of linux-based iot systems,” IEEE Internet of Things Journal , vol. 8, no. 8, pp. 6815–6827, 2020
2020
Earlier work this paper cites.
B. Pearson, C. Zou, Y. Zhang, Z. Ling, and X. Fu, “Sic 2: Securing microcontroller based iot devices with low-cost crypto coprocessors,” in 2020 IEEE 26th International Conference on Parallel and Distributed Systems (ICPADS) . IEEE, 2020, pp. 372–381
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
H. Huang, W. Luo, G. Zeng, J. Weng, Y. Zhang, and A. Yang, “Damia: leveraging domain adaptation as a defense against membership inference attacks,” IEEE Transactions on Dependable and Secure Computing , vol. 19, no. 5, pp. 3183–3199, 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson et al. , “Extracting training data from large language models,” in 30th USENIX Security Symposium (USENIX Security 21) , 2021, pp. 2633–2650
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
J.-B. Truong, P. Maini, R. J. Walls, and N. Papernot, “Data-free model extraction,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 4771–4780
2021
Earlier work this paper cites.
S. Hoory, A. Feder, A. Tendler, S. Erell, A. Peled-Cohen, I. Laish, H. Nakhost, U. Stemmer, A. Benjamini, A. Hassidim et al. , “Learning and evaluating a differentially private pre-trained language model,” in Findings of the Association for Computational Linguistics: EMNLP 2021 , 2021, pp. 1178–1189
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
B. Li and D. Micciancio, “On the security of homomorphic encryption on approximate numbers,” in Advances in Cryptology–EUROCRYPT 2021: 40th Annual International Conference on the Theory and Applications of Cryptographic Techniques, Zagreb, Croatia, October 17–21, 2021, Proceedings, Part I 40 . Springer, 2021, pp. 648–677
2021
Earlier work this paper cites.
J. Duan, S. Yu, H. L. Tan, H. Zhu, and C. Tan, “A survey of embodied ai: From simulators to research tasks,” IEEE Transactions on Emerging Topics in Computational Intelligence , vol. 6, no. 2, pp. 230–244, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
F. Mireshghallah, A. Uniyal, T. Wang, D. K. Evans, and T. Berg-Kirkpatrick, “An empirical analysis of memorization in fine-tuned autoregressive language models,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 1816–1826
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, J. Schulman, J. Hilton, F. Kelton, L. Miller, M. Simens, A. Askell, P. Welinder, P. Christiano, J. Leike, and R. Lowe, “Training language models to follow instructions with human feedback,” 2022
2022
Earlier work this paper cites.
Y. Bai, A. Jones, K. Ndousse, A. Askell, A. Chen, N. DasSarma, D. Drain, S. Fort, D. Ganguli, T. Henighan, N. Joseph, S. Kadavath, J. Kernion, T. Conerly, S. El-Showk, N. Elhage, Z. Hatfield-Dodds, D. Hernandez, T. Hume, S. Johnston, S. Kravec, L. Lovitt, N. Nanda, C. Olsson, D. Amodei, T. Brown, J. Clark, S. McCandlish, C. Olah, B. Mann, and J. Kaplan, “Training a helpful and harmless assistant with reinforcement learning from human feedback,” 2022
2022
Earlier work this paper cites.
Y. Bai, S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, C. Chen, C. Olsson, C. Olah, D. Hernandez, D. Drain, D. Ganguli, D. Li, E. Tran-Johnson, E. Perez, J. Kerr, J. Mueller, J. Ladish, J. Landau, K. Ndousse, K. Lukosuite, L. Lovitt, M. Sellitto, N. Elhage, N. Schiefer, N. Mercado, N. DasSarma, R. Lasenby, R. Larson, S. Ringer, S. Johnston, S. Kravec, S. E. Showk, S. Fort, T. Lanham, T. Telleen-Lawton, T. Conerly, T. Henighan, T. Hume, S. R. Bowman, Z. Hatfield-Dodds, B. Mann, D. Amodei, N. Joseph, S. McCandlish, T. Brown, and J. Kaplan, “Constitutional ai: Harmlessness from ai feedback,” 2022
2022
Earlier work this paper cites.
N. Kandpal, E. Wallace, and C. Raffel, “Deduplicating training data mitigates privacy risks in language models,” 2022
2022
Earlier work this paper cites.
R. Behnia, M. R. Ebrahimi, J. Pacheco, and B. Padmanabhan, “Ew-tune: A framework for privately fine-tuning large language models with differential privacy,” in 2022 IEEE International Conference on Data Mining Workshops (ICDMW) . IEEE, 2022, pp. 560–566
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
X. Wu, L. Gong, and D. Xiong, “Adaptive differential privacy for language model training,” in Proceedings of the First Workshop on Federated Learning for Natural Language Processing (FL4NLP 2022) , 2022, pp. 21–26
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Jang, D. Yoon, S. Yang, S. Cha, M. Lee, L. Logeswaran, and M. Seo, “Knowledge unlearning for mitigating privacy risks in language models,” 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
M. Hao, H. Li, H. Chen, P. Xing, G. Xu, and T. Zhang, “Iron: Private inference on transformers,” Advances in Neural Information Processing Systems , vol. 35, pp. 15 718–15 731, 2022
2022
Earlier work this paper cites.
Y. Wang, G. E. Suh, W. Xiong, B. Lefaudeux, B. Knott, M. Annavaram, and H.-H. S. Lee, “Characterization of mpc-based private inference for transformer-based models,” in 2022 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) . IEEE, 2022, pp. 187–197
2022
Earlier work this paper cites.
X. Zhou, J. Lu, T. Gui, R. Ma, Z. Fei, Y. Wang, Y. Ding, Y. Cheung, Q. Zhang, and X.-J. Huang, “Textfusion: Privacy-preserving pre-trained model inference via token fusion,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 8360–8371
2022
Earlier work this paper cites.
C. Brown and C. Morisset, “Simple and efficient identification of personally identifiable information on a public website,” in 2022 IEEE International Conference on Big Data (Big Data) . IEEE, 2022, pp. 4246–4255
2022
Earlier work this paper cites.
J. Huang, H. Shao, and K. C.-C. Chang, “Are large pre-trained language models leaking your personal information?” 2022
2022
Earlier work this paper cites.
C. Liu, H. Guo, M. Xu, S. Wang, D. Yu, J. Yu, and X. Cheng, “Extending on-chain trust to off-chain – trustworthy blockchain data collection using trusted execution environment (tee),” IEEE Transactions on Computers , vol. 71, no. 12, pp. 3268–3280, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
L. Luo, Y. Zhang, C. White, B. Keating, B. Pearson, X. Shao, Z. Ling, H. Yu, C. Zou, and X. Fu, “On security of trustzone-m-based iot systems,” IEEE Internet of Things Journal , vol. 9, no. 12, pp. 9683–9699, 2022
2022
Cited alongside, same era.
J. Weng, S. Zhijian, Y. Zhang, M. Li, W. Jiasi, Y. Wu, and L. Weiqi, “Peripheral-free secure pairing protocol by randomly switching power,” Mar. 1 2022, uS Patent 11,265,722
2022
Cited alongside, same era.
2023
Cited alongside, same era.
C. H. Song, J. Wu, C. Washington, B. M. Sadler, W.-L. Chao, and Y. Su, “Llm-planner: Few-shot grounded planning for embodied agents with large language models,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 2998–3009
2023
Cited alongside, same era.
S. Yu, J. P. Muñoz, and A. Jannesari, “Federated foundation models: Privacy-preserving and collaborative learning for large models,” 2023
2023
Later among the works it cites.
J. Sun, Z. Xu, H. Yin, D. Yang, D. Xu, Y. Chen, and H. R. Roth, “Fedbpt: Efficient federated black-box prompt tuning for large language models,” 2023
2023
Later among the works it cites.
T. Fan, Y. Kang, G. Ma, W. Chen, W. Wei, L. Fan, and Q. Yang, “Fate-llm: A industrial grade federated learning framework for large language models,” 2023
2023
Later among the works it cites.
M. Du, X. Yue, S. S. Chow, T. Wang, C. Huang, and H. Sun, “Dp-forward: Fine-tuning and inference on language models with differential privacy in forward pass,” in Proceedings of the 2023 ACM SIGSAC Conference on Computer and Communications Security , 2023, pp. 2665–2679
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
N. Subramani, S. Luccioni, J. Dodge, and M. Mitchell, “Detecting personal information in training corpora: an analysis,” in Proceedings of the 3rd Workshop on Trustworthy Natural Language Processing (TrustNLP 2023) , 2023, pp. 208–220
2023
Cited alongside, same era.
2023
Cited alongside, same era.
J. Zhao, “Privacy-preserving fine-tuning of artificial intelligence (ai) foundation models with federated learning, differential privacy, offsite tuning, and parameter-efficient fine-tuning (peft),” Authorea Preprints , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Later among the works it cites.
D. Zhang, P. Finckenberg-Broman, T. Hoang, S. Pan, Z. Xing, M. Staples, and X. Xu, “Right to be forgotten in the era of large language models: Implications, challenges, and solutions,” 2023
2023
Later among the works it cites.
J. Chen and D. Yang, “Unlearn what you want to forget: Efficient unlearning for llms,” 2023
2023
Later among the works it cites.
R. Eldan and M. Russinovich, “Who’s harry potter? approximate unlearning in llms,” 2023
2023
Later among the works it cites.
G. Xiao, J. Lin, and S. Han, “Offsite-tuning: Transfer learning without full model,” 2023
2023
Later among the works it cites.
W.-j. Lu, Z. Huang, Z. Gu, J. Li, J. Liu, K. Ren, C. Hong, T. Wei, and W. Chen, “Bumblebee: Secure two-party inference framework for large transformers,” Cryptology ePrint Archive , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Hou, J. Liu, J. Li, Y. Li, W.-j. Lu, C. Hong, and K. Ren, “Ciphergpt: Secure two-party gpt inference,” Cryptology ePrint Archive , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Akimoto, K. Fukuchi, Y. Akimoto, and J. Sakuma, “Privformer: Privacy-preserving transformer with mpc,” in 2023 IEEE 8th European Symposium on Security and Privacy (EuroS&P) . IEEE, 2023, pp. 392–410
2023
Later among the works it cites.
2023
Later among the works it cites.
K. Gupta, N. Jawalkar, A. Mukherjee, N. Chandran, D. Gupta, A. Panwar, and R. Sharma, “Sigma: secure gpt inference with function secret sharing,” Cryptology ePrint Archive , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
C. Dong, J. Weng, J. Liu, Y. Zhang, Y. Tong, A. Yang, Y. Cheng, and S. Hu, “Fusion: Efficient and secure inference resilient to malicious servers,” in 30th Annual Network and Distributed System Security Symposium, NDSS 2023, San Diego, California, USA, February 27 - March 3, 2023 . The Internet Society, 2023
2023
Later among the works it cites.
N. Lukas, A. Salem, R. Sim, S. Tople, L. Wutschitz, and S. Zanella-Béguelin, “Analyzing leakage of personally identifiable information in language models,” 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Kim, S. Yun, H. Lee, M. Gubri, S. Yoon, and S. J. Oh, “Propile: Probing privacy leakage in large language models,” 2023
2023
Later among the works it cites.
M. Phute, A. Helbling, M. Hull, S. Peng, S. Szyller, C. Cornelius, and D. H. Chau, “Llm self defense: By self examination, llms know they are being tricked,” 2023
2023
Later among the works it cites.
B. Chen, A. Paliwal, and Q. Yan, “Jailbreaker in jail: Moving target defense for large language models,” 2023
2023
Later among the works it cites.
N. Mireshghallah, H. Kim, X. Zhou, Y. Tsvetkov, M. Sap, R. Shokri, and Y. Choi, “Can llms keep a secret? testing privacy implications of language models via contextual integrity theory,” 2023
2023
Later among the works it cites.
D. Glukhov, I. Shumailov, Y. Gal, N. Papernot, and V. Papyan, “Llm censorship: A machine learning challenge or a computer security problem?” 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Chen, H. H. Chen, M. Sun, K. Li, Z. Chen, and X. Wang, “A verified confidential computing as a service framework for privacy preservation,” in 32nd USENIX Security Symposium (USENIX Security 23) , 2023, pp. 4733–4750
2023
Later among the works it cites.
G. Dhanuskodi, S. Guha, V. Krishnan, A. Manjunatha, M. O’Connor, R. Nertney, and P. Rogers, “Creating the first confidential gpus: The team at nvidia brings confidentiality and integrity to user code and data for accelerated computing.” Queue , vol. 21, no. 4, pp. 68–93, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
B. Meskó, “The impact of multimodal large language models on health care’s future,” Journal of Medical Internet Research , vol. 25, p. e52865, 2023
2023
Later among the works it cites.
B. Huang, S. Yu, J. Li, Y. Chen, S. Huang, S. Zeng, and S. Wang, “Firewallm: A portable data protection and recovery framework for llm services,” in International Conference on Data Mining and Big Data . Springer, 2023, pp. 16–30
2023
Later among the works it cites.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
S. Y. Gadre, G. Ilharco, A. Fang, J. Hayase, G. Smyrnis, T. Nguyen, R. Marten, M. Wortsman, D. Ghosh, J. Zhang et al. , “Datacomp: In search of the next generation of multimodal datasets,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
M. Xu, D. Cai, Y. Wu, X. Li, and S. Wang, “Fwdllm: Efficient fedllm using forward gradient,” 2024
2024
Closest in time.
J. Zhang, S. Vahidian, M. Kuo, C. Li, R. Zhang, T. Yu, Y. Zhou, G. Wang, and Y. Chen, “Towards building the federated gpt: Federated instruction tuning,” 2024
2024
Closest in time.
S. Liu, Y. Yao, J. Jia, S. Casper, N. Baracaldo, P. Hase, X. Xu, Y. Yao, H. Li, K. R. Varshney, M. Bansal, S. Koyejo, and Y. Liu, “Rethinking machine unlearning for large language models,” 2024
2024
Closest in time.
2024
Closest in time.