Fetching the paper…
Reading the bibliography…
Large Language Models~(LLMs) have gained immense popularity and are being increasingly applied in various domains.
W. M. Si, M. Backes, J. Blackburn, E. D. Cristofaro, G. Stringhini, S. Zannettou, and Y. Zhang, “Why So Toxic?: Measuring and Triggering Toxic Behavior in Open-Domain Chatbots,” in CCS , 2022, pp. 2659–2673
2022
Earlier work this paper cites.
“Chatgpt-4.0,” https://chat.openai.com/, 2023
2023
Earlier work this paper cites.
“GPT-3 powers the next generation of apps,” https://openai.com/blog/gpt-3-apps, 2023
2023
Earlier work this paper cites.
“Pathways Language Model (PaLM): Scaling to 540 Billion Parameters for Breakthrough Performance,” https://blog.research.google/2022/04/pathways-language-model-palm-scaling-to.html, 2023
2023
Earlier work this paper cites.
“PaLM 2,” https://ai.google/discover/palm2/, 2023
2023
Earlier work this paper cites.
“Introducing Llama 2,” https://ai.meta.com/llama/, 2023
2023
Earlier work this paper cites.
“Introducing LLaMA: A foundational, 65-billion-parameter large language model,” https://ai.meta.com/blog/large-language-model-llama-meta-ai/, 2023
2023
Earlier work this paper cites.
X. Liu, N. Xu, M. Chen, and C. Xiao, “Autodan: Generating stealthy jailbreak prompts on aligned large language models,” 2023
2023
Earlier work this paper cites.
Y. Liu, G. Deng, Z. Xu, Y. Li, Y. Zheng, Y. Zhang, L. Zhao, T. Zhang, and Y. Liu, “Jailbreaking chatgpt via prompt engineering: An empirical study,” 2023
2023
Earlier work this paper cites.
G. Deng, Y. Liu, Y. Li, K. Wang, Y. Zhang, Z. Li, H. Wang, T. Zhang, and Y. Liu, “Masterkey: Automated jailbreak across multiple large language model chatbots,” 2023
2023
Cited alongside, same era.
A. Zou, Z. Wang, N. Carlini, M. Nasr, J. Z. Kolter, and M. Fredrikson, “Universal and transferable adversarial attacks on aligned language models,” 2023
2023
Cited alongside, same era.
“What is retrieval-augmented generation?” https://research.ibm.com/blog/retrieval-augmented-generation-RAG, 2023
2023
Cited alongside, same era.
“Retrieval augmented generation (RAG) explained,” https://www.superannotate.com/blog/rag-explained, 2023
2023
Cited alongside, same era.
“Improve LLM responses in RAG use cases by interacting with the user,” https://aws.amazon.com/blogs/machine-learning/improve-llm-responses-in-rag-use-cases-by-interacting-with-the-user/, 2023
2023
Y. Wolf, N. Wies, Y. Levine, and A. Shashua, “Fundamental limitations of alignment in large language models,” arXiv preprint , 2023
2023
Later among the works it cites.
M. Shanahan, K. McDonell, and L. Reynolds, “Role-play with large language models,” arXiv preprint , 2023
2023
Later among the works it cites.
A. Rao, S. Vashistha, A. Naik, S. Aditya, and M. Choudhury, “Tricking LLMs into Disobedience: Understanding, Analyzing, and Preventing Jailbreaks,” arXiv preprint , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Liu, G. Deng, Y. Li, K. Wang, T. Zhang, Y. Liu, H. Wang, Y. Zheng, and Y. Liu, “Prompt injection attack against llm-integrated applications,” 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Retrieval Augmented Generation (RAG) in Azure AI Search,” https://learn.microsoft.com/en-us/azure/search/retrieval-augmented-generation-overview, 2023
2023
Cited alongside, same era.
“Introducing GPTs,” https://openai.com/blog/introducing-gpts, 2023
2023
Cited alongside, same era.
H. Li, D. Guo, W. Fan, M. Xu, J. Huang, F. Meng, and Y. Song, “Multi-step Jailbreaking Privacy Attacks on ChatGPT,” 2023
2023
Cited alongside, same era.
OpenAI, “GPT-4,” https://openai.com/research/gpt-4
Cited in the paper.
2023
Later among the works it cites.
R. Lapid, R. Langberg, and M. Sipper, “Open sesame! universal black box jailbreaking of large language models,” 2023
2023
Later among the works it cites.
A. Q. Jiang, A. Sablayrolles, A. Mensch, C. Bamford, D. S. Chaplot, D. de las Casas, F. Bressand, G. Lengyel, G. Lample, L. Saulnier, L. R. Lavaud, M.-A. Lachaux, P. Stock, T. L. Scao, T. Lavril, T. Wang, T. Lacroix, and W. E. Sayed, “Mistral 7b,” 2023
2023
Later among the works it cites.