Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have shown remarkable capabilities across various natural language processing tasks but often struggle to excel uniformly in diverse or complex domains.
A density-based algorithm for discovering clusters in large spatial databases with noise
Martin Ester, Hans-Peter Kriegel, Jörg Sander, Xiaowei Xu, et al. 1996 · 1996
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2020 · 2009
Earlier work this paper cites.
Sentence transformers all-minilm-l6-v2
Hugging Face. 2023 · 2023
Earlier work this paper cites.
Routing to the expert: Efficient reward-guided ensemble of large language models
Keming Lu, Hongyi Yuan, Runji Lin, Junyang Lin, Zheng Yuan, Chang Zhou, and Jingren Zhou. 2023 · 2023
Earlier work this paper cites.
Boosted prompt ensembles for large language models
Silviu Pitis, Michael R Zhang, Andrew Wang, and Jimmy Ba. 2023 · 2023
Earlier work this paper cites.
Lora ensembles for large language model fine-tuning
Xi Wang, Laurence Aitchison, and Maja Rudolph. 2023 · 2023
Earlier work this paper cites.
Phi-3 technical report: A highly capable language model locally on your phone
Marah Abdin, Jyoti Aneja, Hany Awadalla, Ahmed Awadallah, Ammar Ahmad Awan, Nguyen Bach, Amit Bahree, Arash Bakhtiari, Jianmin Bao, Harkirat Behl, et al. 2024 · 2024
Earlier work this paper cites.
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, et al. 2024 · 2024
Earlier work this paper cites.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al. 2024 · 2024
Cited alongside, same era.
Chatglm: A family of large language models from glm-130b to glm-4 all tools
Team GLM, Aohan Zeng, Bin Xu, Bowen Wang, Chenhui Zhang, Da Yin, Dan Zhang, Diego Rojas, Guanyu Feng, Hanlin Zhao, et al. 2024 · 2024
Cited alongside, same era.
Ensemble learning for heterogeneous large language models with deep parallel collaboration
Yichong Huang, Xiaocheng Feng, Baohang Li, Yang Xiang, Hui Wang, Ting Liu, and Bing Qin. 2024 · 2024
Cited alongside, same era.
Jinliang Lu, Ziliang Pang, Min Xiao, Yaochen Zhu, Rui Xia, and Jiajun Zhang. 2024 · 2024
Cited alongside, same era.
Bridging the gap between different vocabularies for llm ensemble
Yangyifan Xu, Jinliang Lu, and Jiajun Zhang. 2024 · 2024
Later among the works it cites.
An Yang, Baosong Yang, Beichen Zhang, Binyuan Hui, Bo Zheng, Bowen Yu, Chengyuan Li, Dayiheng Liu, Fei Huang, Haoran Wei, et al. 2024 · 2024
Later among the works it cites.
Yi: Open foundation models by 01. ai
Alex Young, Bei Chen, Chao Li, Chengen Huang, Ge Zhang, Guanwei Zhang, Guoyin Wang, Heng Li, Jiangcheng Zhu, Jianqun Chen, et al. 2024 · 2024
Later among the works it cites.
Efficiently democratizing medical llms for 50 languages via a mixture of language family experts
Guorui Zheng, Xidong Wang, Juhao Liang, Nuo Chen, Yuping Zheng, and Benyou Wang. 2024 · 2024
Later among the works it cites.
Starling-7b: Improving helpfulness and harmlessness with rlaif
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaushal Kumar Maurya, KV Srivatsa, and Ekaterina Kochmar. 2024 · 2024
Cited alongside, same era.
Pack of llms: Model fusion at test-time via perplexity optimization
Costas Mavromatis, Petros Karypis, and George Karypis. 2024 · 2024
Cited alongside, same era.
Gemma: Open models based on gemini research and technology
Gemma Team, Thomas Mesnard, Cassidy Hardin, Robert Dadashi, Surya Bhupatiraju, Shreya Pathak, Laurent Sifre, Morgane Rivière, Mihir Sanjay Kale, Juliette Love, et al. 2024 · 2024
Cited alongside, same era.
Llm-topla: Efficient llm ensemble by maximising diversity
Selim Tekin, Fatih Ilhan, Tiansheng Huang, Sihao Hu, and Ling Liu. 2024 · 2024
Cited alongside, same era.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al. 2023a
Cited in the paper.
Llm-blender: Ensembling large language models with pairwise ranking and generative fusion
Dongfu Jiang, Xiang Ren, and Bill Yuchen Lin. 2023b
Cited in the paper.
Calibrating language models via augmented prompt ensembles
Mingjian Jiang, Yangjun Ruan, Sicong Huang, Saifei Liao, Silviu Pitis, Roger Baker Grosse, and Jimmy Ba. 2023c
Cited in the paper.
Banghua Zhu, Evan Frick, Tianhao Wu, Hanlin Zhu, Karthik Ganesan, Wei-Lin Chiang, Jian Zhang, and Jiantao Jiao. 2024 · 2024
Later among the works it cites.
A survey on large language models with some insights on their capabilities and limitations
Andrea Matarazzo and Riccardo Torlone. 2025 · 2025
Closest in time.
Hit the sweet spot! span-level ensemble for large language models
Yangyifan Xu, Jianghao Chen, Junhong Wu, and Jiajun Zhang. 2025 · 2025
Closest in time.