Fetching the paper…
Reading the bibliography…
Evolutionary model merging enables the creation of high-performing multi-task models but remains computationally prohibitive for consumer hardware.
Statistical theories of mental test scores
Lord, F., Novick, M., and Birnbaum, A · 1968
Earlier work this paper cites.
The feasibility of using item response theory as a psychometric model for the gre aptitude test
Kingston, N. M. and Dorans, N. J · 1982
Earlier work this paper cites.
Using item response theory to equate scholastic aptitude test scores
Petersen, N. S. et al · 1982
Earlier work this paper cites.
Animating rotation with quaternion curves
Shoemake, K · 1985
Earlier work this paper cites.
An overview of evolutionary algorithms for parameter optimization
Bäck, T. and Schwefel, H.-P · 1993
Earlier work this paper cites.
Evolutionary algorithms—an overview
Dasgupta, D. and Michalewicz, Z · 1997
Earlier work this paper cites.
A fast and elitist multiobjective genetic algorithm: Nsga-ii
Deb, K., Pratap, A., Agarwal, S., and Meyarivan, T · 2002
Earlier work this paper cites.
Self-adaptive simulated binary crossover for real-parameter optimization
Deb, K., Sindhya, K., and Okabe, T · 2007
Earlier work this paper cites.
Item response theory: What it is and how you can use the irt procedure to apply it
An, X. and Yung, Y.-F · 2014
Earlier work this paper cites.
Introduction to Evolutionary Computing
Eiben, A. and Smith, J · 2015
Earlier work this paper cites.
Item response theory
Cai, L., Choi, K., Hansen, M., and Harrell, L · 2016
Earlier work this paper cites.
Building an evaluation scale using item response theory
Lalor, J. P., Wu, H., and Yu, H · 2016
Earlier work this paper cites.
Bag of tricks for efficient text classification
Joulin, A., Grave, E., Bojanowski, P., and Mikolov, T · 2017
Earlier work this paper cites.
Evolutionary algorithms
Pétrowski, A. and Ben-Hamida, S · 2017
Earlier work this paper cites.
Think you have solved question answering? try arc, the AI2 reasoning challenge
Clark, P., Cowhey, I., Etzioni, O., Khot, T., Sabharwal, A., Schoenick, C., and Tafjord, O · 2018
Earlier work this paper cites.
Handbook of item response theory: Three volume set
Van der Linden, W. J · 2018
Earlier work this paper cites.
Regularized evolution for image classifier architecture search
Real, E., Aggarwal, A., Huang, Y., and Le, Q. V · 2019
Cited alongside, same era.
HellaSwag: Can a machine really finish your sentence?
Zellers, R., Holtzman, A., Bisk, Y., Farhadi, A., and Choi, Y · 2019
Cited alongside, same era.
Pymoo: Multi-objective optimization in python
Blank, J. and Deb, K · 2020
Cited alongside, same era.
Item response theory models in the measurement theory
Brzezińska, J · 2020
Cited alongside, same era.
Training verifiers to solve math word problems
Cobbe, K., Kosaraju, V., Bavarian, M., Chen, M., Jun, H., Kaiser, L., Plappert, M., Tworek, J., Hilton, J., Nakano, R., et al · 2021
Cited alongside, same era.
Winogrande: An adversarial winograd schema challenge at scale
Sakaguchi, K., Bras, R. L., Bhagavatula, C., and Choi, Y · 2021
Ties-merging: Resolving interference when merging models
Yadav, P., Tam, D., Choshen, L., Raffel, C. A., and Bansal, M · 2023
Later among the works it cites.
Efficiently measuring the cognitive ability of llms: An adaptive testing perspective
Zhuang, Y., Liu, Q., Ning, Y., Huang, W., Lv, R., Huang, Z., Zhao, G., Zhang, Z., Mao, Q., Wang, S., et al · 2023
Later among the works it cites.
Tower: An open multilingual large language model for translation-related tasks
Alves, D. M., Pombal, J., Guerreiro, N. M., Martins, P. H., Alves, J., Farajian, A., Peters, B., Rei, R., Fernandes, P., Agrawal, S., Colombo, P., de Souza, J. G. C., and Martins, A · 2024
Later among the works it cites.
A framework for few-shot language model evaluation, 07 2024
Gao, L., Tow, J., Abbasi, B., Biderman, S., Black, S., DiPofi, A., Foster, C., Golding, L., Hsu, J., Le Noac’h, A., Li, H., McDonell, K., Muennighoff, N., Ociepa, C., Phang, J., Reynolds, L., Schoelkopf, H., Skowron, A., Sutawika, L., Tang, E., Thite, A., Wang, B., Wang, K., and Zou, A · 2024
Later among the works it cites.
Arcee‘s MergeKit: A toolkit for merging large language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Comparing test sets with item response theory
Vania, C., Htut, P. M., Huang, W., Mungra, D., Pang, R. Y., Phang, J., Liu, H., Cho, K., and Bowman, S. R · 2021
Cited alongside, same era.
Git Re-Basin: Merging models modulo permutation symmetries
Ainsworth, S., Hayase, J., and Srinivasa, S · 2022
Cited alongside, same era.
Editing models with task arithmetic
Ilharco, G., Ribeiro, M. T., Wortsman, M., Gururangan, S., Schmidt, L., Hajishirzi, H., and Farhadi, A · 2022
Cited alongside, same era.
TruthfulQA: Measuring how models mimic human falsehoods
Lin, S., Hilton, J., and Evans, O · 2022
Cited alongside, same era.
Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time
Wortsman, M., Ilharco, G., Gadre, S. Y., Roelofs, R., Gontijo-Lopes, R., Morcos, A. S., Namkoong, H., Farhadi, A., Carmon, Y., Kornblith, S., and Schmidt, L · 2022
Cited alongside, same era.
Jiang, A., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D., de las Casas, D., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., Lavaud, L., Lachaux, M.-A., Stock, P., Le Scao, T., Lavril, T., Wang, T., Lacroix, T., and El Sayed, W · 2023
Cited alongside, same era.
Goddard, C., Siriwardhana, S., Ehghaghi, M., Meyers, L., Karpukhin, V., Benedict, B., McQuade, M., and Solawetz, J · 2024
Later among the works it cites.
Task arithmetic in the tangent space: Improved editing of pre-trained models
Ortiz-Jimenez, G., Favero, A., and Frossard, P · 2024
Later among the works it cites.
tinybenchmarks: evaluating llms with fewer examples
Polo, F. M., Weber, L., Choshen, L., Sun, Y., Xu, G., and Yurochkin, M · 2024
Later among the works it cites.
Towards cross-lingual llm evaluation for european languages, 2024
Thellmann, K., Stadler, B., Fromm, M., Buschhoff, J. S., Jude, A., Barth, F., Leveling, J., Flores-Herr, N., Köhler, J., Jäkel, R., and Ali, M · 2024
Later among the works it cites.
Localizing task information for improved model merging and compression
Wang, K., Dimitriadis, N., Ortiz-Jimenez, G., Fleuret, F., and Frossard, P · 2024
Later among the works it cites.
Language models are super mario: Absorbing abilities from homologous models as a free lunch
Yu, L., Yu, B., Yu, H., Huang, F., and Li, Y · 2024
Later among the works it cites.
Atm: Improving model merging by alternating tuning and merging
Zhou, L., Solombrino, D., Crisostomi, D., Bucarelli, M. S., Silvestri, F., and Rodolà, E · 2024
Later among the works it cites.
Evolutionary optimization of model merging recipes
Akiba, T., Shing, M., Tang, Y., Sun, Q., and Ha, D · 2025
Closest in time.
c 2 m 3 c^{2}m^{3} : Cycle-consistent multi-model merging
Crisostomi, D., Fumero, M., Baieri, D., Bernard, F., and Rodolà, E · 2025
Closest in time.
Model breadcrumbs: Scaling multi-task model merging with sparse masks
Davari, M. and Belilovsky, E · 2025
Closest in time.
Task singular vectors: Reducing task interference in model merging
Gargiulo, A. A., Crisostomi, D., Bucarelli, M. S., Scardapane, S., Silvestri, F., and Rodolà, E · 2025
Closest in time.