Fetching the paper…
Reading the bibliography…
Retrieval Augmented Generation (RAG) systems have shown great promise in natural language processing.
Membership inference attacks from first principles
Carlini, N., Chien, S., Nasr, M., Song, S., Terzis, A., and Tramer, F. (2022) · 1914
Earlier work this paper cites.
Membership inference attacks against machine learning models
Shokri, R., Stronati, M., Song, C., and Shmatikov, V. (2017) · 2017
Earlier work this paper cites.
Auditing data provenance in text-generation models
Song, C. and Shmatikov, V. (2019) · 2019
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Lewis, P., Perez, E., Piktus, A., Petroni, F., Karpukhin, V., Goyal, N., Küttler, H., Lewis, M., Yih, W.-t., Rocktäschel, T., et al. (2020) · 2020
Earlier work this paper cites.
On adaptive attacks to adversarial example defenses
Tramer, F., Carlini, N., Brendel, W., and Madry, A. (2020) · 2020
Earlier work this paper cites.
Dp-fp: Differentially private forward propagation for large models
Du, J. and Mi, H. (2021) · 2021
Earlier work this paper cites.
Membership inference on word embedding and beyond
Mahloujifar, S., Inan, H. A., Chase, M., Ghosh, E., and Hasegawa, M. (2021) · 2021
Earlier work this paper cites.
Prompt programming for large language models: Beyond the few-shot paradigm
Reynolds, L. and McDonell, K. (2021) · 2021
Earlier work this paper cites.
Membership inference attacks against nlp classification models
Shejwalkar, V., Inan, H. A., Houmansadr, A., and Sim, R. (2021) · 2021
Earlier work this paper cites.
Membership inference attacks against self-supervised speech models
Tseng, W.-C., Kao, W.-T., and Lee, H.-y. (2021) · 2021
Earlier work this paper cites.
Milvus: A purpose-built vector data management system
Wang, J., Yi, X., Guo, R., Jin, H., Xu, P., Li, S., Wang, X., Guo, X., Li, C., Xu, X., et al. (2021) · 2021
Earlier work this paper cites.
Langchain
Chase, H. (2022) · 2022
Earlier work this paper cites.
Manu: a cloud native vector database management system
Guo, R., Luan, X., Xiang, L., Yan, X., Yi, X., Luo, J., Cheng, Q., Xu, W., Luo, J., Liu, F., et al. (2022) · 2022
Earlier work this paper cites.
Membership inference attacks on machine learning: A survey
Hu, H., Salcic, Z., Sun, L., Dobbie, G., Yu, P. S., and Zhang, X. (2022) · 2022
Cited alongside, same era.
Differentially private decoding in large language models
Majmudar, J., Dupuy, C., Peris, C., Smaili, S., Gupta, R., and Zemel, R. (2022) · 2022
Cited alongside, same era.
Retrieval-augmented generation for large language models: A survey
Gao, Y., Xiong, Y., Gao, X., Jia, K., Pan, J., Bi, Y., Dai, Y., Sun, J., and Wang, H. (2023) · 2023
Cited alongside, same era.
Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection
Greshake, K., Abdelnabi, S., Mishra, S., Endres, C., Holz, T., and Fritz, M. (2023) · 2023
Cited alongside, same era.
Mistral 7B
Jiang, A. Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D. S., de las Casas, D., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., Lavaud, L. R., Lachaux, M.-A., Stock, P., Scao, T. L., Lavril, T., Wang, T., Lacroix, T., and Sayed, W. E. (2023) · 2023
Cited alongside, same era.
Do membership inference attacks work on large language models?
Duan, M., Suri, A., Mireshghallah, N., Min, S., Shi, W., Zettlemoyer, L., Tsvetkov, Y., Choi, Y., Evans, D., and Hajishirzi, H. (2024) · 2024
Closest in time.
Noisy neighbors: Efficient membership inference attacks against llms
Galli, F., Melis, L., and Cucinotta, T. (2024) · 2024
Closest in time.
Prompt perturbation in retrieval-augmented generation based large language models
Hu, Z., Wang, C., Shu, Y., Zhu, L., et al. (2024) · 2024
Closest in time.
Formalizing and benchmarking prompt injection attacks and defenses
Liu, Y., Jia, Y., Geng, R., Jia, J., and Gong, N. Z. (2024) · 2024
Closest in time.
Keeping llms aligned after fine-tuning: The crucial role of prompt templates
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
User inference attacks on llms
Kandpal, N., Pillutla, K., Oprea, A., Kairouz, P., Choquette-Choo, C., and Xu, Z. (2023) · 2023
Cited alongside, same era.
Improved membership inference attacks against language classification models
Shachor, S., Razinkov, N., and Goldsteen, A. (2023) · 2023
Cited alongside, same era.
UL2: Unifying Language Learning Paradigms
Tay, Y., Dehghani, M., Tran, V. Q., Garcia, X., Wei, J., Wang, X., Chung, H. W., Shakeri, S., Bahri, D., Schuster, T., Zheng, H. S., Zhou, D., Houlsby, N., and Metzler, D. (2023) · 2023
Cited alongside, same era.
Poisoning retrieval corpora by injecting adversarial passages
Zhong, Z., Huang, Z., Wettig, A., and Chen, D. (2023) · 2023
Cited alongside, same era.
Llama 3 model card
AI@Meta (2024) · 2024
Cited alongside, same era.
Private prediction for large-scale synthetic text generation
Amin, K., Bie, A., Kong, W., Kurakin, A., Ponomareva, N., Syed, U., Terzis, A., and Vassilvitskii, S. (2024) · 2024
Cited alongside, same era.
Sok: Reducing the vulnerability of fine-tuned language models to membership inference attacks
Amit, G., Goldsteen, A., and Farkash, A. (2024) · 2024
Cited alongside, same era.
Lyu, K., Zhao, H., Gu, X., Yu, D., Goyal, A., and Arora, S. (2024) · 2024
Closest in time.
What Is a Prompt Injection Attack? — IBM — ibm.com
Matthew Kosinski, A. F. (2024) · 2024
Closest in time.
Privacy auditing of large language models
Panda, A., Tang, X., Nasr, M., Choquette-Choo, C. A., and Mittal, P. (2024) · 2024
Closest in time.
Prompt injection attacks against GPT-3 — simonwillison.net
Willison, S. (2022) · 2024
Closest in time.
Differentially private synthetic data via foundation model apis 2: Text
Xie, C., Lin, Z., Backurs, A., Gopi, S., Yu, D., Inan, H. A., Nori, H., Jiang, H., Zhang, H., Lee, Y. T., et al. (2024) · 2024
Closest in time.
The Good and The Bad: Exploring Privacy Issues in Retrieval-Augmented Generation (RAG)
Zeng, S., Zhang, J., He, P., Xing, Y., Liu, Y., Xu, H., Ren, J., Wang, S., Yin, D., Chang, Y., et al. (2024) · 2024
Closest in time.
Min-k%++: Improved baseline for detecting pre-training data from large language models
Zhang, J., Sun, J., Yeats, E., Ouyang, Y., Kuo, M., Zhang, J., Yang, H. F., and Li, H. (2024) · 2024
Closest in time.
PoisonedRAG: Knowledge Poisoning Attacks to Retrieval-Augmented Generation of Large Language Models
Zou, W., Geng, R., Wang, B., and Jia, J. (2024) · 2024
Closest in time.