Fetching the paper…
Reading the bibliography…
In the rapidly evolving field of artificial intelligence, multimodal models, e.g., integrating vision and language into visual-language models (VLMs), have become pivotal for many applications, ranging from image captioning to multimodal search engines.
L. Fei-Fei, R. Fergus, and P. Perona, “Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,” in Computer Vision and Pattern Recognition (CVPR) Workshop , 2004
2004
Earlier work this paper cites.
M.-E. Nilsback and A. Zisserman, “Automated flower classification over a large number of classes,” in Indian Conference on Computer Vision, Graphics and Image Processing , 2008
2008
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” Master’s thesis, Department of Computer Science, University of Toronto, 2009
2009
Earlier work this paper cites.
J. Xiao, J. Hays, K. A. Ehinger, A. Oliva, and A. Torralba, “Sun database: Large-scale scene recognition from abbey to zoo,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2010
2010
Earlier work this paper cites.
O. M. Parkhi, A. Vedaldi, A. Zisserman, and C. V. Jawahar, “Cats and dogs,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
J. Krause, M. Stark, J. Deng, and L. Fei-Fei, “3d object representations for fine-grained categorization,” in International IEEE Workshop on 3D Representation and Recognition (3dRR) , 2013
2013
Earlier work this paper cites.
Dumitru, I. Goodfellow, W. Cukierski, and Y. Bengio, “Challenges in representation learning: Facial expression recognition challenge,” 2013
2013
Earlier work this paper cites.
M. Cimpoi, S. Maji, I. Kokkinos, S. Mohamed, , and A. Vedaldi, “Describing textures in the wild,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2014
2014
Earlier work this paper cites.
L. Bossard, M. Guillaumin, and L. Van Gool, “Food-101 – mining discriminative components with random forests,” in European Conference on Computer Vision (ECCV) , 2014
2014
Earlier work this paper cites.
J. Xiao, K. A. Ehinger, J. Hays, A. Torralba, and A. Oliva, “Sun database: Exploring a large collection of scene categories,” in International Journal of Computer Vision (IJCV) , 2014
2014
Earlier work this paper cites.
O. Vinyals, M. Fortunato, and N. Jaitly, “Pointer networks,” in Advances in Neural Information Processing Systems (NeurIPS) , 2015
2015
Earlier work this paper cites.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhutdinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in International Conference on Machine Learning (ICML) , 2015
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei, “ImageNet Large Scale Visual Recognition Challenge,” in International Journal of Computer Vision (IJCV) , 2015
2015
Earlier work this paper cites.
H. B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas, “Communication-efficient learning of deep networks from decentralized data,” in International Conference on Artificial Intelligence and Statistics (AISTATS) , 2017
2017
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
——, “Introducing eurosat: A novel dataset and deep learning benchmark for land use and land cover classification,” in IGARSS IEEE International Geoscience and Remote Sensing Symposium , 2018
2018
Cited alongside, same era.
Q. Li, B. He, and D. Song, “Model-contrastive federated learning,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “LoRA: Low-rank adaptation of large language models,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
M. Wortsman, G. Ilharco, J. W. Kim, M. Li, S. Kornblith, R. Roelofs, R. Gontijo-Lopes, H. Hajishirzi, A. Farhadi, H. Namkoong, and L. Schmidt, “Robust fine-tuning of zero-shot models,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Association for Computational Linguistics (ACL) , 2019
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Cited alongside, same era.
J. Lu, D. Batra, D. Parikh, and S. Lee, “Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,” in Advances in Neural Information Processing Systems (NeurIPS) , 2019
2019
Cited alongside, same era.
H. Tan and M. Bansal, “LXMERT: Learning cross-modality encoder representations from transformers,” in Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2019
2019
Cited alongside, same era.
P. Helber, B. Bischke, A. Dengel, and D. Borth, “Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing , 2019
2019
Cited alongside, same era.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” in Advances in Neural Information Processing Systems (NeurIPS) , 2020
2020
Cited alongside, same era.
X. Li, K. Huang, W. Yang, S. Wang, and Z. Zhang, “On the convergence of FedAvg on Non-IID data,” in International Conference on Learning Representations (ICML) , 2020
2020
Cited alongside, same era.
T. Lin, L. Kong, S. U. Stich, and M. Jaggi, “Ensemble distillation for robust model fusion in federated learning,” in Advances in Neural Information Processing Systems (NeurIPS) , 2020
2020
Cited alongside, same era.
S. Goyal, A. Kumar, S. Garg, Z. Kolter, and A. Raghunathan, “Finetune like you pretrain: Improved finetuning of zero-shot vision models,” in Computer Vision and Pattern Recognition Conference (CVPR) , 2022
2022
Later among the works it cites.
Q. Li, Y. Diao, Q. Chen, and B. He, “Federated learning on non-iid data silos: An experimental study,” in International Conference on Data Engineering (ICDE) , 2022
2022
Later among the works it cites.
T. Guo, S. Guo, J. Wang, X. Tang, and W. Xu, “Promptfl: Let federated participants cooperatively learn prompts instead of models - federated learning in age of foundation model,” in IEEE Transactions on Mobile Computing (TMC) , 2023
2023
Later among the works it cites.
T. Guo, S. Guo, and J. Wang, “pfedprompt: Learning personalized prompt for vision-language models in federated learning,” in ACM Web Conference (WWW) , 2023
2023
Later among the works it cites.
W. Lu, X. Hu, J. Wang, and X. Xie, “Fedclip: Fast generalization and personalization for CLIP in federated learning,” IEEE Data Eng. Bull. , vol. 46, no. 1, pp. 52–66, 2023. [Online]. Available: http://sites.computer.org/debull/A23mar/p52.pdf
2023
Later among the works it cites.
C. Qiu, X. Li, C. K. Mummadi, M. R. Ganesh, Z. Li, L. Peng, and W.-Y. Lin, “Text-driven prompt generation for vision-language models in federated learning,” in International Conference on Learning Representations (ICLR) , 2024
2024
Closest in time.
X. Wu, X. Liu, J. Niu, H. Wang, S. Tang, and G. Zhu, “FedloRA: When personalized federated learning meets low-rank adaptation,” 2024. [Online]. Available: https://openreview.net/forum?id=bZh06ptG9r
2024
Closest in time.
2024
Closest in time.
J. P. Muñoz, J. Yuan, Y. Zheng, and N. Jain, “Lonas: Elastic low-rank adapters for efficient large language models,” in The 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation , 2024
2024
Closest in time.
J. P. Muñoz, J. Yuan, and N. Jain, “Shears: Unstructured sparsity with neural low-rank adapter search,” in 2024 Annual Conference of the North American Chapter of the Association for Computational Linguistics - Industry Track , 2024
2024
Closest in time.