Fetching the paper…
Reading the bibliography…
Large language models (LLMs) exhibit complementary strengths in various tasks, motivating the research of LLM ensembling.
Minimum Bayes-risk decoding for statistical machine translation
S. Kumar and W. Byrne · 2004
Earlier work this paper cites.
Towards understanding ensemble, knowledge distillation and self-distillation in deep learning, 2020
Z. Allen-Zhu and Y. Li · 2012
Earlier work this paper cites.
Ensemble Machine Learning: Methods and Applications
C. Zhang and Y. Ma · 2012
Earlier work this paper cites.
Ensemble methods: foundations and algorithms
Z.-H. Zhou · 2012
Earlier work this paper cites.
Ensemble learning for multi-source neural machine translation
E. Garmash and C. Monz · 2016
Earlier work this paper cites.
TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension
M. Joshi, E. Choi, D. Weld, and L. Zettlemoyer · 2017
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge, 2018
P. Clark, I. Cowhey, O. Etzioni, T. Khot, A. Sabharwal, C. Schoenick, and O. Tafjord · 2018
Earlier work this paper cites.
Ensemble learning: A survey
O. Sagi and L. Rokach · 2018
Earlier work this paper cites.
Natural questions: A benchmark for question answering research
T. Kwiatkowski, J. Palomaki, O. Redfield, M. Collins, A. Parikh, C. Alberti, D. Epstein, I. Polosukhin, J. Devlin, K. Lee, K. Toutanova, L. Jones, M. Kelcey, M.-W. Chang, A. M. Dai, J. Uszkoreit, Q. Le, and S. Petrov · 2019
Earlier work this paper cites.
Relational knowledge distillation
W. Park, D. Kim, Y. Lu, and M. Cho · 2019
Earlier work this paper cites.
Piqa: Reasoning about physical commonsense in natural language
Y. Bisk, R. Zellers, R. Le bras, J. Gao, and Y. Choi · 2020
Earlier work this paper cites.
A survey on ensemble learning
X. Dong, Z. Yu, W. Cao, Y. Shi, and Q. Ma · 2020
Cited alongside, same era.
Training verifiers to solve math word problems, 2021
K. Cobbe, V. Kosaraju, M. Bavarian, M. Chen, H. Jun, L. Kaiser, M. Plappert, J. Tworek, J. Hilton, R. Nakano, C. Hesse, and J. Schulman · 2021
Cited alongside, same era.
Measuring massive multitask language understanding
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt · 2021
Cited alongside, same era.
Feature kernel distillation
B. He and M. Ozay · 2022
Cited alongside, same era.
No language left behind: Scaling human-centered machine translation, 2022
N. Team · 2022
Cited alongside, same era.
Tigerbot: An open multilingual multitask llm, 2023
Y. Chen, W. Cai, L. Wu, X. Li, Z. Xin, and C. Fu · 2023
Gpt-4 technical report, 2023
OpenAI · 2023
Later among the works it cites.
Internlm: A multilingual language model with progressively enhanced capabilities
I. Team · 2023
Later among the works it cites.
Skywork: A more open bilingual foundation model, 2023
T. Wei, L. Zhao, L. Zhang, B. Zhu, L. Wang, and at el · 2023
Later among the works it cites.
A survey of large language models, 2023
W. X. Zhao, K. Zhou, J. Li, T. Tang, X. Wang, Y. Hou, Y. Min, B. Zhang, J. Zhang, Z. Dong, Y. Du, C. Yang, Y. Chen, Z. Chen, J. Jiang, R. Ren, Y. Li, X. Tang, Z. Liu, P. Liu, J.-Y. Nie, and J.-R. Wen · 2023
Later among the works it cites.
Yi: Open foundation models by 01.ai, 2024
. AI · 2024
Closest in time.
Mixtral of experts, 2024
A. Q. Jiang, A. Sablayrolles, A. Roux, A. Mensch, B. Savary, and et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Epsilon sampling rocks: Investigating sampling strategies for minimum Bayes risk decoding for machine translation
M. Freitag, B. Ghorbani, and P. Fernandes · 2023
Cited alongside, same era.
Specializing smaller language models towards multi-step reasoning
Y. Fu, H. Peng, L. Ou, A. Sabharwal, and T. Khot · 2023
Cited alongside, same era.
LLM-blender: Ensembling large language models with pairwise ranking and generative fusion
D. Jiang, X. Ren, and B. Y. Lin · 2023
Cited alongside, same era.
Routing to the expert: Efficient reward-guided ensemble of large language models, 2023
K. Lu, H. Yuan, R. Lin, J. Lin, Z. Yuan, C. Zhou, and J. Zhou · 2023
Cited alongside, same era.
Relative representations enable zero-shot latent space communication
L. Moschella, V. Maiorca, M. Fumero, A. Norelli, F. Locatello, and E. Rodolà · 2023
Cited alongside, same era.
Mistral 7b, 2023a
A. Q. Jiang, A. Sablayrolles, A. Mensch, C. Bamford, D. S. Chaplot, D. de las Casas, F. Bressand, G. Lengyel, G. Lample, L. Saulnier, L. R. Lavaud, M.-A. Lachaux, P. Stock, T. L. Scao, T. Lavril, T. Wang, T. Lacroix, and W. E. Sayed
Cited in the paper.
Pack of llms: Model fusion at test-time via perplexity optimization, 2024
C. Mavromatis, P. Karypis, and G. Karypis · 2024
Closest in time.
Large language model routing with benchmark datasets, 2024
T. Shnitzer, A. Ou, M. Silva, K. Soule, Y. Sun, J. Solomon, N. Thompson, and M. Yurochkin · 2024
Closest in time.
Knowledge fusion of large language models
F. Wan, X. Huang, D. Cai, X. Quan, W. Bi, and S. Shi · 2024
Closest in time.
Fusing models with complementary expertise
H. Wang, F. M. Polo, Y. Sun, S. Kundu, E. Xing, and M. Yurochkin · 2024
Closest in time.
Bridging the gap between different vocabularies for llm ensemble, 2024
Y. Xu, J. Lu, and J. Zhang · 2024
Closest in time.