Fetching the paper…
Reading the bibliography…
Although vision models such as Contrastive Language-Image Pre-Training (CLIP) show impressive generalization performance, their zero-shot robustness is still limited under Out-of-Distribution (OOD) scenarios without fine-tuning.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A., Hinton, G., et al · 2009
Earlier work this paper cites.
Learning with noisy labels
Natarajan, N., Dhillon, I. S., Ravikumar, P. K., and Tewari, A · 2013
Earlier work this paper cites.
Classification with noisy labels by importance reweighting
Liu, T. and Tao, D · 2015
Earlier work this paper cites.
Deep learning face attributes in the wild
Liu, Z., Luo, P., Wang, X., and Tang, X · 2015
Earlier work this paper cites.
A baseline for detecting misclassified and out-of-distribution examples in neural networks
Hendrycks, D. and Gimpel, K · 2016
Earlier work this paper cites.
Improved regularization of convolutional neural networks with cutout
DeVries, T. and Taylor, G. W · 2017
Earlier work this paper cites.
mixup: Beyond empirical risk minimization
Zhang, H., Cisse, M., Dauphin, Y. N., and Lopez-Paz, D · 2017
Earlier work this paper cites.
Co-teaching: Robust training of deep neural networks with extremely noisy labels
Han, B., Yao, Q., Yu, X., Niu, G., Xu, M., Hu, W., Tsang, I., and Sugiyama, M · 2018
Earlier work this paper cites.
Deep anomaly detection with outlier exposure
Hendrycks, D., Mazeika, M., and Dietterich, T · 2018
Earlier work this paper cites.
Benchmarking neural network robustness to common corruptions and perturbations
Hendrycks, D. and Dietterich, T · 2019
Earlier work this paper cites.
Do imagenet classifiers generalize to imagenet?
Recht, B., Roelofs, R., Schmidt, L., and Shankar, V · 2019
Earlier work this paper cites.
Learning robust global representations by penalizing local predictive power
Wang, H., Ge, S., Lipton, Z., and Xing, E. P · 2019
Earlier work this paper cites.
Are anchor points really indispensable in label-noise learning?
Xia, X., Liu, T., Wang, N., Han, B., Gong, C., Niu, G., and Sugiyama, M · 2019
Earlier work this paper cites.
Cutmix: Regularization strategy to train strong classifiers with localizable features
Yun, S., Han, D., Oh, S. J., Chun, S., Choe, J., and Yoo, Y · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Earlier work this paper cites.
Invariant rationalization
Chang, S., Zhang, Y., Yu, M., and Jaakkola, T · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., et al · 2020
Earlier work this paper cites.
Gpt-3: Its nature, scope, limits, and consequences
Floridi, L. and Chiriatti, M · 2020
Earlier work this paper cites.
In search of lost domain generalization
Gulrajani, I. and Lopez-Paz, D · 2020
Earlier work this paper cites.
Energy-based out-of-distribution detection
Liu, W., Wang, X., Owens, J., and Li, Y · 2020
Earlier work this paper cites.
Dual t: Reducing estimation error for transition matrix in label-noise learning
Yao, Y., Liu, T., Han, B., Gong, M., Deng, J., Niu, G., and Sugiyama, M · 2020
Earlier work this paper cites.
In search of lost domain generalization
Gulrajani, I. and Lopez-Paz, D · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Hu, E. J., Shen, Y., Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W · 2021
Cited alongside, same era.
Supervision exists everywhere: A data efficient contrastive language-image pre-training paradigm
Li, Y., Liang, F., Zhao, L., Cui, Y., Ouyang, W., Shao, J., Yu, F., and Yan, J · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B · 2021
Cited alongside, same era.
Domain generalization using causal matching
Mahajan, D., Tople, S., and Sharma, A · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Cited alongside, same era.
Benchmarking robustness of 3d object detection to common corruptions
Dong, Y., Kang, C., Zhang, J., Zhu, Z., Wang, Y., Yang, X., Su, H., Wei, X., and Zhu, J · 2023
Closest in time.
Imagebind: One embedding space to bind them all
Girdhar, R., El-Nouby, A., Liu, Z., Singh, M., Alwala, K. V., Joulin, A., and Misra, I · 2023
Closest in time.
Multimodal-gpt: A vision and language model for dialogue with humans
Gong, T., Lyu, C., Zhang, S., Wang, Y., Zheng, M., Zhao, Q., Liu, K., Zhang, W., Luo, P., and Chen, K · 2023
Closest in time.
Finetune like you pretrain: Improved finetuning of zero-shot vision models
Goyal, S., Kumar, A., Garg, S., Kolter, Z., and Raghunathan, A · 2023
Closest in time.
Robo3d: Towards robust and reliable 3d perception against corruptions
Kong, L., Liu, Y., Li, X., Chen, R., Zhang, W., Ren, J., Pan, L., Chen, K., and Liu, Z · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pyramid vision transformer: A versatile backbone for dense prediction without convolutions
Wang, W., Xie, E., Li, X., Fan, D.-P., Song, K., Liang, D., Lu, T., Luo, P., and Shao, L · 2021
Cited alongside, same era.
An explanation of in-context learning as implicit bayesian inference
Xie, S. M., Raghunathan, A., Liang, P., and Ma, T · 2021
Cited alongside, same era.
Instance-dependent label-noise learning under a structural causal model
Yao, Y., Liu, T., Gong, M., Han, B., Niu, G., and Zhang, K · 2021
Cited alongside, same era.
Towards defending against adversarial examples via attack-invariant features
Zhou, D., Liu, T., Han, B., Wang, N., Peng, C., and Gao, X · 2021
Cited alongside, same era.
Flamingo: a visual language model for few-shot learning
Alayrac, J.-B., Donahue, J., Luc, P., Miech, A., Barr, I., Hasson, Y., Lenc, K., Mensch, A., Millican, K., Reynolds, M., et al · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models
Chung, H. W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al · 2022
Cited alongside, same era.
Viewfool: Evaluating the robustness of visual recognition to adversarial viewpoints
Dong, Y., Ruan, S., Su, H., Kang, C., Wei, X., and Zhu, J · 2022
Cited alongside, same era.
Liu, H., Li, C., Wu, Q., and Lee, Y. J · 2023
Closest in time.
Spawrious: A benchmark for fine control of spurious correlation biases
Lynch, A., Dovonon, G. J., Kaddour, J., and Silva, R · 2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.
Sam-guided unsupervised domain adaptation for 3d segmentation
Peng, X., Chen, R., Qiao, F., Kong, L., Liu, Y., Wang, T., Zhu, X., and Ma, Y · 2023
Closest in time.
Clipood: Generalizing clip to out-of-distributions
Shu, Y., Guo, X., Wu, J., Wang, X., Wang, J., and Long, M · 2023
Closest in time.
Pandagpt: One model to instruction-follow them all
Su, Y., Lan, T., Li, H., Xu, J., Wang, Y., and Cai, D · 2023
Closest in time.
Making binary classification from multiple unlabeled datasets almost free of supervision
Wu, Y., Xia, X., Yu, J., Han, B., Niu, G., Sugiyama, M., and Liu, T · 2023
Closest in time.
Which is better for learning with noisy labels: the semi-supervised method or modeling label noise?
Yao, Y., Gong, M., Du, Y., Yu, J., Han, B., Zhang, K., and Liu, T · 2023
Closest in time.
Retrieval-augmented multimodal language modeling
Yasunaga, M., Aghajanyan, A., Shi, W., James, R., Leskovec, J., Liang, P., Lewis, M., Zettlemoyer, L., and Yih, W.-t · 2023
Closest in time.
mplug-owl: Modularization empowers large language models with multimodality
Ye, Q., Xu, H., Xu, G., Ye, J., Yan, M., Zhou, Y., Wang, J., Hu, A., Shi, P., Shi, Y., et al · 2023
Closest in time.
Investigating the catastrophic forgetting in multimodal large language models
Zhai, Y., Tong, S., Li, X., Cai, M., Qu, Q., Lee, Y. J., and Ma, Y · 2023
Closest in time.
Mmicl: Empowering vision-language model with multi-modal in-context learning
Zhao, H., Cai, Z., Si, S., Ma, X., An, K., Chen, L., Liu, Z., Wang, S., Han, W., and Chang, B · 2023
Closest in time.
Towards label-free scene understanding by vision foundation models
Chen, R., Liu, Y., Kong, L., Chen, N., Zhu, X., Ma, Y., Liu, T., and Wang, W · 2024
Closest in time.
Improving non-transferable representation learning by harnessing content and style
Hong, Z., Wang, Z., Shen, L., Yao, Y., Huang, Z., Chen, S., Yang, C., Gong, M., and Liu, T · 2024
Closest in time.
Visual-augmented dynamic semantic prototype for generative zero-shot learning
Hou, W., Chen, S., Chen, S., Hong, Z., Wang, Y., Feng, X., Khan, S., Khan, F. S., and You, X · 2024
Closest in time.
Eliminating catastrophic overfitting via abnormal adversarial examples regularization
Lin, R., Yu, C., and Liu, T · 2024
Closest in time.
Mitigating label noise on graph via topological sample selection
Wu, Y., Yao, J., Xia, X., Yu, J., Wang, R., Han, B., and Liu, T · 2024
Closest in time.
Early stopping against label noise without validation data
Yuan, S., Feng, L., and Liu, T · 2024
Closest in time.