Fetching the paper…
Reading the bibliography…
Vision-Language models (VLMs) that use contrastive language-image pre-training have shown promising zero-shot classification performance.
Vapnik V (1991) Principles of risk minimization for learning theory. Advances in neural information processing systems 4
1991
Earlier work this paper cites.
Platt J, Cristianini N, Shawe-Taylor J (1999) Large margin dags for multiclass classification. NIPS 12
1999
Earlier work this paper cites.
Yang CY, Yang JS, Wang JJ (2009) Margin calibration in svm class-imbalanced learning. Neurocomputing 73(1-3):397–411
2009
Earlier work this paper cites.
He K, Zhang X, Ren S, et al (2016) Deep residual learning for image recognition. In: CVPR, pp 770–778
2016
Earlier work this paper cites.
Khan SH, Hayat M, Bennamoun M, et al (2017) Cost-sensitive learning of deep feature representations from imbalanced data. IEEE TNNLS 29(8):3573–3587
2017
Earlier work this paper cites.
Lin TY, Goyal P, Girshick R, et al (2017) Focal loss for dense object detection. In: ICCV, pp 2980–2988
2017
Earlier work this paper cites.
Wang YX, Ramanan D, Hebert M (2017) Learning to model the tail. In: NeurIPS, pp 7032–7042
2017
Earlier work this paper cites.
Zhou B, Lapedriza A, Khosla A, et al (2017) Places: A 10 million image database for scene recognition. IEEE TPAMI 40(6):1452–1464
2017
Earlier work this paper cites.
Van Horn G, Mac Aodha O, Song Y, et al (2018) The inaturalist species classification and detection dataset. In: CVPR, pp 8769–8778
2018
Earlier work this paper cites.
Byrd J, Lipton Z (2019) What is the effect of importance weighting in deep learning? In: ICML, PMLR, pp 872–881
2019
Earlier work this paper cites.
Kang B, Xie S, Rohrbach M, et al (2019) Decoupling representation and classifier for long-tailed recognition. In: ICML
2019
Earlier work this paper cites.
Liu Z, Miao Z, Zhan X, et al (2019) Large-scale long-tailed recognition in an open world. In: CVPR, pp 2537–2546
2019
Earlier work this paper cites.
Yin X, Yu X, Sohn K, et al (2019) Feature transfer learning for face recognition with under-represented data. In: CVPR
2019
Earlier work this paper cites.
Dosovitskiy A, Beyer L, Kolesnikov A, et al (2020) An image is worth 16x16 words: Transformers for image recognition at scale. In: International Conference on Learning Representations
2020
Cited alongside, same era.
Jamal MA, Brown M, Yang MH, et al (2020) Rethinking class-balanced methods for long-tailed visual recognition from a domain adaptation perspective. In: CVPR, pp 7610–7619
2020
Cited alongside, same era.
Menon AK, Jayasumana S, Rawat AS, et al (2020) Long-tail learning via logit adjustment. In: ICLR
2020
Cited alongside, same era.
Ren J, Yu C, Sheng S, et al (2020) Balanced meta-softmax for long-tailed visual recognition. arXiv preprint arXiv:200710740
2020
Cited alongside, same era.
Tan J, Wang C, Li B, et al (2020) Equalization loss for long-tailed object recognition. In: CVPR, pp 11662–11671
2020
Cited alongside, same era.
He K, Chen X, Xie S, et al (2022) Masked autoencoders are scalable vision learners. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 16000–16009
2022
Later among the works it cites.
Li J, Li D, Xiong C, et al (2022) Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation. In: International Conference on Machine Learning, PMLR, pp 12888–12900
2022
Later among the works it cites.
Liu Z, Hu H, Lin Y, et al (2022) Swin transformer v2: Scaling up capacity and resolution. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp 12009–12019
2022
Later among the works it cites.
Lüddecke T, Ecker A (2022) Image segmentation using text and image prompts. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 7086–7096
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tang K, Huang J, Zhang H (2020) Long-tailed classification by keeping the good and removing the bad momentum causal effect. NeurIPS 33
2020
Cited alongside, same era.
Yang Y, Xu Z (2020) Rethinking the value of labels for improving class-imbalanced learning. In: NeurIPS
2020
Cited alongside, same era.
Zhou B, Cui Q, Wei XS, et al (2020) Bbn: Bilateral-branch network with cumulative learning for long-tailed visual recognition. In: CVPR
2020
Cited alongside, same era.
Hong Y, Han S, Choi K, et al (2021) Disentangling label distribution for long-tailed visual recognition. In: CVPR, pp 6626–6636
2021
Cited alongside, same era.
Ma T, Geng S, Wang M, et al (2021) A simple long-tailed recognition baseline via vision-language model. arXiv preprint arXiv:211114745
2021
Cited alongside, same era.
Radford A, Kim JW, Hallacy C, et al (2021) Learning transferable visual models from natural language supervision. In: International conference on machine learning, PMLR, pp 8748–8763
2021
Cited alongside, same era.
Zhang S, Li Z, Yan S, et al (2021) Distribution alignment: A unified framework for long-tail visual recognition. In: CVPR
2021
Cited alongside, same era.
Schuhmann C, Beaumont R, Vencu R, et al (2022) Laion-5b: An open large-scale dataset for training next generation image-text models. In: Thirty-sixth Conference on Neural Information Processing Systems Datasets and Benchmarks Track
2022
Later among the works it cites.
Tian C, Wang W, Zhu X, et al (2022) Vl-ltr: Learning class-wise visual-linguistic representation for long-tailed visual recognition. In: Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXV, Springer, pp 73–91
2022
Later among the works it cites.
Wang Y, Zhang B, Hou W, et al (2022) Margin calibration for long-tailed visual recognition. In: Asian Conference on Machine Learning (ACML)
2022
Later among the works it cites.
Wei H, Tao L, Xie R, et al (2022) Open-sampling: Exploring out-of-distribution data for re-balancing long-tailed datasets. In: International Conference on Machine Learning, PMLR, pp 23615–23630
2022
Later among the works it cites.
Yang L, Jiang H, Song Q, et al (2022) A survey on long-tailed visual recognition. IJCV pp 1–36
2022
Later among the works it cites.
Yu J, Wang Z, Vasudevan V, et al (2022) Coca: Contrastive captioners are image-text foundation models. arXiv preprint arXiv:220501917
2022
Later among the works it cites.
Dehghani M, Djolonga J, Mustafa B, et al (2023) Scaling vision transformers to 22 billion parameters. arXiv preprint arXiv:230205442
2023
Closest in time.
Xu Z, Yang S, Wang X, et al (2023) Rethink long-tailed recognition with vision transforms. In: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE, pp 1–5
2023
Closest in time.