Ultrafeedback: Boosting language models with high-quality feedback, 2023
Cui, G., Yuan, L., Ding, N., Yao, G., Zhu, W., Ni, Y., Xie, G., Liu, Z., and Sun, M · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Original
Gemini, T., Anil, R., Borgeaud, S., Wu, Y., Alayrac, J.-B., Yu, J., Soricut, R., Schalkwyk, J., Dai, A. M., Hauth, A., et al · 2023
Later among the works it cites.
Koala: A dialogue model for academic research
Geng, X., Gudibande, A., Liu, H., Wallace, E., Abbeel, P., Levine, S., and Song, D · 2023
Later among the works it cites.
Competition-level problems are effective llm evaluators
Original
Huang, Y., Lin, Z., Liu, X., Gong, Y., Lu, S., Lei, F., Liang, Y., Shen, Y., Lin, C., Duan, N., et al · 2023
Later among the works it cites.
Openassistant conversations–democratizing large language model alignment
Original
Köpf, A., Kilcher, Y., von Rütte, D., Anagnostidis, S., Tam, Z.-R., Stevens, K., Barhoum, A., Duc, N. M., Stanley, O., Nagyfi, R., et al · 2023
Later among the works it cites.
Alpacaeval: An automatic evaluator of instruction-following models
Li, X., Zhang, T., Dubois, Y., Taori, R., Gulrajani, I., Guestrin, C., Liang, P., and Hashimoto, T. B · 2023
Later among the works it cites.
ToxicChat: Unveiling hidden challenges of toxicity detection in real-world user-AI conversation
Lin, Z., Wang, Z., Tong, Y., Wang, Y., Guo, Y., Wang, Y., and Shang, J · 2023
Later among the works it cites.
Gpt-4 technical report
Original
OpenAI · 2023
Later among the works it cites.
Proving test set contamination in black box language models
Original
Oren, Y., Meister, N., Chatterji, N., Ladhak, F., and Hashimoto, T. B · 2023
Later among the works it cites.
Game-theoretic statistics and safe anytime-valid inference
Ramdas, A., Grünwald, P., Vovk, V., and Shafer, G · 2023
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Srivastava, A., Rastogi, A., Rao, A., Shoeb, A. A. M., Abid, A., Fisch, A., Brown, A. R., Santoro, A., Gupta, A., Garriga-Alonso, A., et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Original
Touvron, H., Martin, L., Stone, K., Albert, P., Almahairi, A., Babaei, Y., Bashlykov, N., Batra, S., Bhargava, P., Bhosale, S., et al · 2023
Later among the works it cites.
Self-instruct: Aligning language models with self-generated instructions
Wang, Y., Kordi, Y., Mishra, S., Liu, A., Smith, N. A., Khashabi, D., and Hajishirzi, H · 2023
Later among the works it cites.
Rethinking benchmark and contamination for language models with rephrased samples
Original
Yang, S., Chiang, W.-L., Zheng, L., Gonzalez, J. E., and Stoica, I · 2023
Later among the works it cites.
Agieval: A human-centric benchmark for evaluating foundation models
Original
Zhong, W., Cui, R., Guo, Y., Liang, Y., Lu, S., Wang, Y., Saied, A., Chen, W., and Duan, N · 2023
Later among the works it cites.
Starling-7b: Improving llm helpfulness & harmlessness with rlaif, November 2023
Zhu, B., Frick, E., Wu, T., Zhu, H., and Jiao, J · 2023
Later among the works it cites.