Fetching the paper…
Reading the bibliography…
As Large Language Models (LLMs) are frequently updated, LoRA weights trained on earlier versions quickly become obsolete.
J. Munkres, “Algorithms for the assignment and transportation problems,” Journal of the society for industrial and applied mathematics , vol. 5, no. 1, pp. 32–38, 1957
1957
Earlier work this paper cites.
L. Song, A. Smola, A. Gretton, J. Bedo, and K. Borgwardt, “Feature selection via dependence maximization,” The Journal of Machine Learning Research , vol. 13, pp. 1393–1434, 2012
2012
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 1026–1034
2015
Earlier work this paper cites.
R. Koncel-Kedziorski, S. Roy, A. Amini, N. Kushman, and H. Hajishirzi, “Mawps: A math word problem repository,” in Proceedings of the 2016 conference of the north american chapter of the association for computational linguistics: human language technologies , 2016, pp. 1152–1157
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Yamada, Y. Umezu, K. Fukumizu, and I. Takeuchi, “Post selection inference with kernels,” in International conference on artificial intelligence and statistics . PMLR, 2018, pp. 152–160
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Morcos, M. Raghu, and S. Bengio, “Insights on representational similarity in neural networks with canonical correlation,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for nlp,” in International conference on machine learning . PMLR, 2019, pp. 2790–2799
2019
Earlier work this paper cites.
S. Kornblith, M. Norouzi, H. Lee, and G. Hinton, “Similarity of neural network representations revisited,” in International conference on machine learning . PMLR, 2019, pp. 3519–3529
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
E. Strubell, A. Ganesh, and A. McCallum, “Energy and policy considerations for modern deep learning research,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 09, 2020, pp. 13 693–13 696
2020
Cited alongside, same era.
Y. Bisk, R. Zellers, J. Gao, Y. Choi et al. , “Piqa: Reasoning about physical commonsense in natural language,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 05, 2020, pp. 7432–7439
2020
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
A. Conmy, A. Mavor-Parker, A. Lynch, S. Heimersheim, and A. Garriga-Alonso, “Towards automated circuit discovery for mechanistic interpretability,” Advances in Neural Information Processing Systems , vol. 36, pp. 16 318–16 352, 2023
2023
Later among the works it cites.
Q. Zhang, M. Chen, A. Bukharin, P. He, Y. Cheng, W. Chen, and T. Zhao, “Adaptive budget allocation for parameter-efficient fine-tuning,” in The Eleventh International Conference on Learning Representations , 2023. [Online]. Available: https://openreview.net/forum?id=lq62uWRJjiY
2023
Later among the works it cites.
Y. Li, Y. Yu, Q. Zhang, C. Liang, P. He, W. Chen, and T. Zhao, “Losparse: Structured compression of large language models based on low-rank and sparse approximation,” in International Conference on Machine Learning . PMLR, 2023, pp. 20 336–20 350
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y.-L. Sung, V. Nair, and C. A. Raffel, “Training neural networks with fixed sparse masks,” Advances in Neural Information Processing Systems , vol. 34, pp. 24 193–24 205, 2021
2021
Cited alongside, same era.
T. Nguyen, M. Raghu, and S. Kornblith, “Do wide and deep networks learn the same things? uncovering how neural network representations vary with width and depth,” in International Conference on Learning Representations , 2021. [Online]. Available: https://openreview.net/forum?id=KJNcAkY8tY4
2021
Cited alongside, same era.
K. Sakaguchi, R. L. Bras, C. Bhagavatula, and Y. Choi, “Winogrande: An adversarial winograd schema challenge at scale,” Communications of the ACM , vol. 64, no. 9, pp. 99–106, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “Lora: Low-rank adaptation of large language models,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=nZeVKeeFYf9
2022
Cited alongside, same era.
K. Meng, D. Bau, A. Andonian, and Y. Belinkov, “Locating and editing factual associations in gpt,” Advances in Neural Information Processing Systems , vol. 35, pp. 17 359–17 372, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Later among the works it cites.
H. Wen, Y. Li, G. Liu, S. Zhao, T. Yu, T. J.-J. Li, S. Jiang, Y. Liu, Y. Zhang, and Y. Liu, “Autodroid: Llm-powered task automation in android,” in Proceedings of the 30th Annual International Conference on Mobile Computing and Networking , 2024, pp. 543–557
2024
Later among the works it cites.
A. Yang, B. Yang, B. Hui et al. , “Qwen2 technical report,” arXiv preprint arXiv:2407.10671 , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
M. Xu, D. Cai, Y. Wu, X. Li, and S. Wang, “Fwdllm: Efficient federated finetuning of large language models with perturbed inferences,” in 2024 USENIX Annual Technical Conference (USENIX ATC 24) , 2024, pp. 579–596
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Google, “Gemini nano with the google ai edge sdk,” 2025. [Online]. Available: https://developer.android.com/ai/gemini-nano
2025
Closest in time.
Meta, “The llama 4 herd: The beginning of a new era of natively multimodal ai innovation,” 2025. [Online]. Available: https://ai.meta.com/blog/llama-4-multimodal-intelligence/
2025
Closest in time.
Q. Sun, M. Pickett, A. K. Nain, and L. Jones, “Transformer layers as painters,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 39, no. 24, 2025, pp. 25 219–25 227
2025
Closest in time.