Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have revolutionized Artificial Intelligence (AI) services due to their exceptional proficiency in understanding and generating human-like text.
D. Lawley, “A Generalization of Fisher’s Z Test,” Biometrika , vol. 30, no. 1/2, pp. 180–187, 1938
1938
Earlier work this paper cites.
I. Cohen, Y. Huang, J. Chen, J. Benesty, J. Benesty, J. Chen, Y. Huang, and I. Cohen, “Pearson correlation coefficient,” Noise reduction in speech processing , pp. 1–4, 2009
2009
Earlier work this paper cites.
F. L. Bauer, Cæsar Cipher . Boston, MA: Springer US, 2011, pp. 180–180. [Online]. Available: https://doi.org/10.1007/978-1-4419-5906-5_162
2011
Earlier work this paper cites.
I. Beltagy, K. Lo, and A. Cohan, “Scibert: A pretrained language model for scientific text,” in EMNLP , 2019
2019
Earlier work this paper cites.
S. Gehman, S. Gururangan, M. Sap, Y. Choi, and N. A. Smith, “RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models,” in EMNLP , 2020, pp. 3356–3369
2020
Earlier work this paper cites.
L. Reynolds and K. McDonell, “Prompt programming for large language models: Beyond the few-shot paradigm,” in CHI EA , 2021
2021
Earlier work this paper cites.
E. Bagdasaryan and V. Shmatikov, “Spinning Language Models: Risks of Propaganda-As-A-Service and Countermeasures,” in S&P . IEEE, 2022, pp. 769–786
2022
Earlier work this paper cites.
F. Perez and I. Ribeiro, “Ignore Previous Prompt: Attack Techniques For Language Models,” in NeurIPS ML Safety Workshop , 2022
2022
Earlier work this paper cites.
A. Salem, M. Backes, and Y. Zhang, “Get a Model! Model Hijacking Attack Against Machine Learning Models,” in NDSS , 2022
2022
Earlier work this paper cites.
W. M. Si, M. Backes, J. Blackburn, E. D. Cristofaro, G. Stringhini, S. Zannettou, and Y. Zhang, “Why So Toxic?: Measuring and Triggering Toxic Behavior in Open-Domain Chatbots,” in CCS , 2022, pp. 2659–2673
2022
Earlier work this paper cites.
W. Sun, Z. Shi, S. Gao, P. Ren, M. de Rijke, and Z. Ren, “Contrastive Learning Reduces Hallucination in Conversations,” arXiv preprint , 2022
2022
Earlier work this paper cites.
T. Van Ede, H. Aghakhani, N. Spahn, R. Bortolameotti, M. Cova, A. Continella, M. van Steen, A. Peter, C. Kruegel, and G. Vigna, “Deepcase: Semi-supervised Contextual Analysis of Security Events,” in IEEE S&P , 2022, pp. 522–539
2022
Earlier work this paper cites.
W. Xiang, C. Li, Y. Zhou, B. Wang, and L. Zhang, “Language Supervised Training for Skeleton-based Action Recognition,” 2022
2022
Earlier work this paper cites.
A. Yuan, A. Coenen, E. Reif, and D. Ippolito, “Wordcraft: Story writing with large language models,” in IUI , 2022, p. 841–852
2022
Cited alongside, same era.
Z. Zhang, L. Lyu, X. Ma, C. Wang, and X. Sun, “Fine-mixing: Mitigating Backdoors in Fine-tuned Language Models,” in EMNLP , 2022, pp. 355–372
2022
Cited alongside, same era.
2022
Cited alongside, same era.
“Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality — lmsys org,” https://lmsys.org/blog/2023-03-30-vicuna/
2023
Cited alongside, same era.
G. Apruzzese, H. S. Anderson, S. Dambra, D. Freeman, F. Pierazzi, and K. A. Roundy, “”Real Attackers Don’t Compute Gradients”: Bridging the Gap between Adversarial ML Research and Practice,” in SaTML , 2023
P. Manakul, A. Liusie, and M. J. Gales, “Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language models,” arXiv preprint , 2023
2023
Closest in time.
N. McKenna, T. Li, L. Cheng, M. J. Hosseini, M. Johnson, and M. Steedman, “Sources of Hallucination by Large Language Models on Inference Tasks,” arXiv preprint , 2023
2023
Closest in time.
K. Mei, Z. Li, Z. Wang, Y. Zhang, and S. Ma, “NOTABLE: Transferable Backdoor Attacks Against Prompt-based NLP Models,” in ACL , 2023
2023
Closest in time.
J. Oppenlaender, R. Linder, and J. Silvennoinen, “Prompting AI Art: An Investigation into the Creative Skill of Prompt Engineering,” arXiv preprint , 2023
2023
Closest in time.
R. Pryzant, D. Iter, J. Li, Y. T. Lee, C. Zhu, and M. Zeng, “Automatic Prompt Optimization with Gradient Descent and Beam Search,” arXiv preprint , 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
S. Diao, R. Pan, H. Dong, K. S. Shum, J. Zhang, W. Xiong, and T. Zhang, “Lmflow: An extensible toolkit for finetuning and inference of large foundation models,” 2023
2023
Cited alongside, same era.
H. Dong, W. Xiong, D. Goyal, Y. Zhang, W. Chow, R. Pan, S. Diao, J. Zhang, K. Shum, and T. Zhang, “Raft: Reward ranked finetuning for generative foundation model alignment,” 2023
2023
Cited alongside, same era.
K. Greshake, S. Abdelnabi, S. Mishra, C. Endres, T. Holz, and M. Fritz, “Not what you’ve signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection,” in arXiv preprint , 2023
2023
Cited alongside, same era.
Jay Peters, “The Bing AI bot has been secretly running GPT-4,” https://www.theverge.com/2023/3/14/23639928/microsoft-bing-chatbot-ai-gpt-4-llm
2023
Cited alongside, same era.
E. Kasneci, K. Sessler, S. Küchemann, M. Bannert, D. Dementieva, F. Fischer, U. Gasser, G. Groh, S. Günnemann, E. Hüllermeier, S. Krusche, G. Kutyniok, T. Michaeli, C. Nerdel, J. Pfeffer, O. Poquet, M. Sailer, A. Schmidt, T. Seidel, M. Stadler, J. Weller, J. Kuhn, and G. Kasneci, “Chatgpt for good? on opportunities and challenges of large language models for education,” Learning and Individual Differences , vol. 103, p. 102274, 2023
2023
Cited alongside, same era.
H. Li, D. Guo, W. Fan, M. Xu, J. Huang, F. Meng, and Y. Song, “Multi-step Jailbreaking Privacy Attacks on ChatGPT,” 2023
2023
Cited alongside, same era.
X. Liu and Z. Liu, “Llms can understand encrypted prompt: Towards privacy-computing friendly transformers,” 2023
2023
Cited alongside, same era.
2023
Closest in time.
A. Rao, S. Vashistha, A. Naik, S. Aditya, and M. Choudhury, “Tricking LLMs into Disobedience: Understanding, Analyzing, and Preventing Jailbreaks,” arXiv preprint , 2023
2023
Closest in time.
M. Shanahan, K. McDonell, and L. Reynolds, “Role-play with large language models,” arXiv preprint , 2023
2023
Closest in time.
W. M. Si, M. Backes, Y. Zhang, and A. Salem, “Two-in-One: A Model Hijacking Attack Against Text Generation Models,” arXiv preprint , 2023
2023
Closest in time.
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, and T. B. Hashimoto, “Stanford alpaca: An instruction-following llama model,” 2023
2023
Closest in time.
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, A. Rodriguez, A. Joulin, E. Grave, and G. Lample, “Llama: Open and efficient foundation language models,” 2023
2023
Closest in time.
J. White, Q. Fu, S. Hays, M. Sandborn, C. Olea, H. Gilbert, A. Elnashar, J. Spencer-Smith, and D. C. Schmidt, “A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT,” arXiv preprint , 2023
2023
Closest in time.
Y. Wolf, N. Wies, Y. Levine, and A. Shashua, “Fundamental limitations of alignment in large language models,” arXiv preprint , 2023
2023
Closest in time.
J. Zamfirescu-Pereira, R. Y. Wong, B. Hartmann, and Q. Yang, “Why Johnny Can’t Prompt: How Non-AI Experts Try (and fail) to Design LLM Prompts,” in CHI , 2023, pp. 1–21
2023
Closest in time.