Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have become increasingly integrated with various applications.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Reliability in content analysis: Some common misconceptions and recommendations
Klaus Krippendorff · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie · 2005
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul Francis Christiano, Jan Leike, Tom B. Brown, Miljan Martic, and Shane Legg et al · 2017
Earlier work this paper cites.
Towards an automatic Turing test: Learning to evaluate dialogue responses
Ryan Lowe, Michael Noseworthy, Iulian Vlad Serban, Nicolas Angelard-Gontier, Yoshua Bengio, Joelle Pineau, and Min-Yen Kan · 2017
Earlier work this paper cites.
https://bit.ly/3TRQD7P , 2018
doccano: Text annotation tool for human · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Klaus Kippendorff, Karthik Narasimhan, Tim Salimans, and Sutskever · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Fine-tuning language models from human preferences
Daniel M. Ziegler, Nisan Stiennon, Jeff Wu, Tom B. Brown, and Alec Radford et al · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, and Jared et al. Kaplan · 2020
Earlier work this paper cites.
BLEURT: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, Joyce Parikh, Ankur Chai, Natalie Schluter, and Joel Tetreault · 2020
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi · 2020
Earlier work this paper cites.
Safetynot: on the usage of the safetynet attestation api in android
Muhammad Ibrahim, Abdullah Imran, and Antonio Bianchi · 2021
Earlier work this paper cites.
Concealed data poisoning attacks on nlp models
Eric Wallace, Tony Zhao, Shi Feng, and Sameer Singh · 2021
Earlier work this paper cites.
Constitutional ai: Harmlessness from ai feedback
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, and Jackson et al. Kernion · 2022
Earlier work this paper cites.
Truthfulqa: Measuring how models mimic human falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans · 2022
Earlier work this paper cites.
Pinsql: Pinpoint root cause sqls to resolve performance issues in cloud databases
Xiaoze Liu, Zheng Yin, Chao Zhao, Congcong Ge, and Lu et al. Chen · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, and Carroll L. Wainwright et al · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, and Pamela Mishkin et al · 2022
Cited alongside, same era.
https://bit.ly/3ITv1BK , 2023
Generative ai prohibited use policy · 2023
Cited alongside, same era.
https://bit.ly/3TQWzOa , 2023
How to use chatgpt to summarize an article · 2023
Cited alongside, same era.
https://bit.ly/3TPxfs4 , 2023
An overview of bard: an early experiment with generative ai · 2023
Cited alongside, same era.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, and Amjad Almahairi et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin R. Stone, Peter Albert, and Amjad Almahairi et al · 2023
Later among the works it cites.
Poisoning language models during instruction tuning
Alexander Wan, Eric Wallace, Sheng Shen, and Dan Klein · 2023
Later among the works it cites.
Survey on factuality in large language models: Knowledge, retrieval and domain-specificity, 2023
Cunxiang Wang, Xiaoze Liu, Yuanhao Yue, Xiangru Tang, and Tianhang Zhang et al · 2023
Later among the works it cites.
Real-time workload pattern analysis for large-scale cloud databases, 2023
Jiaqi Wang, Tianyi Li, Anni Wang, Xiaoze Liu, and Lu Chen et al · 2023
Later among the works it cites.
Jailbroken: How does LLM safety training fail?
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
https://bit.ly/4abuBm8 , 2023
Universe 2023: Copilot transforms github into the ai-powered developer platform · 2023
Cited alongside, same era.
Large language model-based chatbot as a source of advice on first aid in heart attack
Alexei A Birkun and Adhish Gautam · 2023
Cited alongside, same era.
Jailbreaking black box large language models in twenty queries
Patrick Chao, Alexander Robey, Edgar Dobriban, Hamed Hassani, and George J et al. Pappas · 2023
Cited alongside, same era.
Gptscore: Evaluate as you desire, 2023
Jinlan Fu, See-Kiong Ng, Zhengbao Jiang, and Pengfei Liu · 2023
Cited alongside, same era.
Catastrophic jailbreak of open-source llms via exploiting generation
Yangsibo Huang, Samyak Gupta, Mengzhou Xia, Kai Li, and Danqi Chen · 2023
Cited alongside, same era.
Aot - attack on things: A security analysis of iot firmware updates
Muhammad Ibrahim, Andrea Continella, and Antonio Bianchi · 2023
Cited alongside, same era.
Alexander Wei, Nika Haghtalab, and Jacob Steinhardt · 2023
Later among the works it cites.
TrojanSQL: SQL injection against natural language interface to database
Jinchuan Zhang, Yan Zhou, Binyuan Hui, Yaxin Liu, Ziming Li, and Songlin Hu · 2023
Later among the works it cites.
Judging LLM-as-a-judge with MT-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, and Zhanghao Wu et al · 2023
Later among the works it cites.
Universal and transferable adversarial attacks on aligned language models
Andy Zou, Zifan Wang, J Zico Kolter, and Matt Fredrikson · 2023
Later among the works it cites.
https://blog.google/technology/developers/gemma-open-models/ , 2024
Gemma: Introducing new state-of-the-art open models · 2024
Closest in time.
https://ai.meta.com/blog/meta-llama-3/ , 2024
Introducing meta llama 3: The most capable openly available llm to date · 2024
Closest in time.
https://bit.ly/3vrCYLc , 2024
Usage policies — openai.com · 2024
Closest in time.
https://bit.ly/3Q0uS3t , 2024
Usenix submission replication · 2024
Closest in time.
Knowledge graphs meet multi-modal learning: A comprehensive survey, 2024
Zhuo Chen, Yichi Zhang, Yin Fang, Yuxia Geng, and Lingbing Guo et al · 2024
Closest in time.
Harmbench: A standardized evaluation framework for automated red teaming and robust refusal
Mantas Mazeika, Long Phan, Xuwang Yin, Andy Zou, and Zifan et al. Wang · 2024
Closest in time.
Attackeval: How to evaluate the effectiveness of jailbreak attacking on large language models, 2024
Dong shu, Mingyu Jin, Suiyuan Zhu, Beichen Wang, and Zihao Zhou et al · 2024
Closest in time.
Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Xiang Yue, Yuansheng Ni, Kai Zhang, Tianyu Zheng, and Ruoqi Liu et al · 2024
Closest in time.