Fetching the paper…
Reading the bibliography…
Real-world recognition system often encounters the challenge of unseen labels.
Multi-label classification: An overview
Tsoumakas, G.; and Katakis, I. 2007 · 2007
Earlier work this paper cites.
NUS-WIDE: A Real-World Web Image Database from National University of Singapore
Chua, T.-S.; Tang, J.; Hong, R.; Li, H.; Luo, Z.; and Zheng, Y.-T. 2009 · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
Lampert, C. H.; Nickisch, H.; and Harmeling, S. 2009 · 2009
Earlier work this paper cites.
Classifier chains for multi-label classification
Read, J.; Pfahringer, B.; Holmes, G.; and Frank, E. 2011 · 2011
Earlier work this paper cites.
Distributed Representations of Words and Phrases and their Compositionality
Mikolov, T.; Sutskever, I.; Chen, K.; Corrado, G. S.; and Dean, J. 2013 · 2013
Earlier work this paper cites.
Deep Convolutional Ranking for Multilabel Image Annotation
Gong, Y.; Jia, Y.; Leung, T.; Toshev, A.; and Ioffe, S. 2014 · 2014
Earlier work this paper cites.
Attribute-Based Classification for Zero-Shot Visual Object Categorization
Lampert, C. H.; Nickisch, H.; and Harmeling, S. 2014 · 2014
Earlier work this paper cites.
Glove: Global Vectors for Word Representation
Pennington, J.; Socher, R.; and Manning, C. 2014 · 2014
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Simonyan, K.; and Zisserman, A. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G.; Vinyals, O.; Dean, J.; et al. 2015 · 2015
Earlier work this paper cites.
Zero-Shot Learning via Semantic Similarity Embedding
Zhang, Z.; and Saligrama, V. 2015 · 2015
Earlier work this paper cites.
Conditional Graphical Lasso for Multi-Label Image Classification
Li, Q.; Qiao, M.; Bian, W.; and Tao, D. 2016 · 2016
Earlier work this paper cites.
We Don’t Need No Bounding-Boxes: Training Object Class Detectors Using Only Human Verification
Papadopoulos, D. P.; Uijlings, J. R. R.; Keller, F.; and Ferrari, V. 2016 · 2016
Earlier work this paper cites.
CNN-RNN: A Unified Framework for Multi-Label Image Classification
Wang, J.; Yang, Y.; Mao, J.; Huang, Z.; Huang, C.; and Xu, W. 2016 · 2016
Earlier work this paper cites.
Fast Zero-Shot Image Tagging
Zhang, Y.; Gong, B.; and Shah, M. 2016 · 2016
Earlier work this paper cites.
Learning From Noisy Large-Scale Datasets With Minimal Supervision
Veit, A.; Alldrin, N.; Chechik, G.; Krasin, I.; Gupta, A.; and Belongie, S. 2017 · 2017
Earlier work this paper cites.
Multi-Label Image Recognition by Recurrently Discovering Attentional Regions
Wang, Z.; Chen, T.; Li, G.; Xu, R.; and Lin, L. 2017 · 2017
Earlier work this paper cites.
Zero-Shot Learning - the Good, the Bad and the Ugly
Xian, Y.; Schiele, B.; and Akata, Z. 2017 · 2017
Earlier work this paper cites.
Learning Spatial Regularization With Image-Level Supervisions for Multi-Label Image Classification
Zhu, F.; Li, H.; Ouyang, W.; Yu, N.; and Wang, X. 2017 · 2017
Cited alongside, same era.
Multi-Label Zero-Shot Learning With Structured Knowledge Graphs
Lee, C.-W.; Fang, W.; Yeh, C.-K.; and Wang, Y.-C. F. 2018 · 2018
Cited alongside, same era.
Multi-Label Image Recognition With Graph Convolutional Networks
Chen, Z.-M.; Wei, X.-S.; Wang, P.; and Guo, Y. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. N. 2018 · 2019
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, J.; Batra, D.; Parikh, D.; and Lee, S. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Generative Multi-Label Zero-Shot Learning
Gupta, A.; Narayan, S.; Khan, S.; Khan, F. S.; Shao, L.; and van de Weijer, J. 2021 · 2021
Later among the works it cites.
Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-Labeling
Huynh, D.; Kuen, J.; Lin, Z.; Gu, J.; and Elhamifar, E. 2021 · 2021
Later among the works it cites.
Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision
Jia, C.; Yang, Y.; Xia, Y.; Chen, Y.-T.; Parekh, Z.; Pham, H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T. 2021 · 2021
Later among the works it cites.
ViLT: Vision-and-Language Transformer Without Convolution or Region Supervision
Kim, W.; Son, B.; and Kim, I. 2021 · 2021
Later among the works it cites.
General Multi-label Image Classification with Transformers
Lanchantin, J.; Wang, T.; Ordonez, V.; and Qi, Y. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Radford, A.; Wu, J.; Child, R.; Luan, D.; Amodei, D.; Sutskever, I.; et al. 2019 · 2019
Cited alongside, same era.
UNITER: Learning UNiversal Image-TExt Representations
Chen, Y.-C.; Li, L.; Yu, L.; Kholy, A. E.; Ahmed, F.; Gan, Z.; Cheng, Y.; and Liu, J. J. 2020 · 2020
Cited alongside, same era.
A Shared Multi-Attention Framework for Multi-Label Zero-Shot Learning
Huynh, D.; and Elhamifar, E. 2020 · 2020
Cited alongside, same era.
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-Modal Pre-Training
Li, G.; Duan, N.; Fang, Y.; Gong, M.; and Jiang, D. 2020 · 2020
Cited alongside, same era.
Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks
Li, X.; Yin, X.; Li, C.; Zhang, P.; Hu, X.; Zhang, L.; Wang, L.; Hu, H.; Dong, L.; Wei, F.; Choi, Y.; and Gao, J. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2020 · 2020
Cited alongside, same era.
Semantic Diversity Learning for Zero-Shot Multi-Label Classification
Ben-Cohen, A.; Zamir, N.; Ben-Baruch, E.; Friedman, I.; and Zelnik-Manor, L. 2021 · 2021
Cited alongside, same era.
Li, X. L.; and Liang, P. 2021 · 2021
Later among the works it cites.
Query2Label: A Simple Transformer Way to Multi-Label Classification
Liu, S.; Zhang, L.; Yang, X.; Su, H.; and Zhu, J. 2021 · 2021
Later among the works it cites.
Discriminative Region-Based Multi-Label Zero-Shot Learning
Narayan, S.; Gupta, A.; Khan, S.; Khan, F. S.; Shao, L.; and Shah, M. 2021 · 2021
Later among the works it cites.
Learning Transferable Visual Models From Natural Language Supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I. 2021 · 2021
Later among the works it cites.
Open-Vocabulary Object Detection Using Captions
Zareian, A.; Rosa, K. D.; Hu, D. H.; and Chang, S.-F. 2021 · 2021
Later among the works it cites.
Learning to Prompt for Vision-Language Models
Zhou, K.; Yang, J.; Loy, C. C.; and Liu, Z. 2021 · 2021
Later among the works it cites.
Learning to Prompt for Open-Vocabulary Object Detection with Vision-Language Model
Du, Y.; Wei, F.; Zhang, Z.; Shi, M.; Gao, Y.; and Li, G. 2022 · 2022
Closest in time.
Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning
Liang, W.; Zhang, Y.; Kwon, Y.; Yeung, S.; and Zou, J. 2022 · 2022
Closest in time.
Open-Vocabulary One-Stage Detection with Hierarchical Visual-Language Knowledge Distillation
Ma, Z.; Luo, G.; Gao, J.; Li, L.; Chen, Y.; Wang, S.; Zhang, C.; and Hu, W. 2022 · 2022
Closest in time.
Dualcoop: Fast adaptation to multi-label recognition with limited annotations
Sun, X.; Hu, P.; and Saenko, K. 2022 · 2022
Closest in time.
A Dual Modality Approach For (Zero-Shot) Multi-Label Classification
Xu, S.; Li, Y.; Hsiao, J.; Ho, C.; and Qi, Z. 2022 · 2022
Closest in time.
Open-Vocabulary DETR with Conditional Matching
Zang, Y.; Li, W.; Zhou, K.; Huang, C.; and Loy, C. C. 2022 · 2022
Closest in time.