Fetching the paper…
Reading the bibliography…
Fine-grained visual classification (FGVC) which aims at recognizing objects from subcategories is a very challenging task due to the inherently subtle inter-class differences.
Transformer-xl: Attentive language models beyond a fixed-length context
Dai, Z.; Yang, Z.; Yang, Y.; Carbonell, J.; Le, Q. V.; and Salakhutdinov, R. 2019 · 1901
Earlier work this paper cites.
Serrano, S.; and Smith, N. A. 2019 · 1906
Earlier work this paper cites.
Learning deep bilinear transformation for fine-grained image representation
Zheng, H.; Fu, J.; Zha, Z.-J.; and Luo, J. 2019 · 1911
Earlier work this paper cites.
Quantifying attention flow in transformers
Abnar, S.; and Zuidema, W. 2020 · 2005
Earlier work this paper cites.
Ranking measures and loss functions in learning to rank
Chen, W.; Liu, T.-Y.; Lan, Y.; Ma, Z.-M.; and Li, H. 2009 · 2009
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; et al. 2020 · 2010
Earlier work this paper cites.
Deformable detr: Deformable transformers for end-to-end object detection
Zhu, X.; Su, W.; Lu, L.; Li, B.; Wang, X.; and Dai, J. 2020 · 2010
Earlier work this paper cites.
Novel Dataset for Fine-Grained Image Categorization
Khosla, A.; Jayadevaprakash, N.; Yao, B.; and Fei-Fei, L. 2011 · 2011
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
Wah, C.; Branson, S.; Welinder, P.; Perona, P.; and Belongie, S. 2011 · 2011
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks
Krizhevsky, A.; Sutskever, I.; and Hinton, G. E. 2012 · 2012
Earlier work this paper cites.
Transtrack: Multiple-object tracking with transformer
Sun, P.; Jiang, Y.; Zhang, R.; Xie, E.; Cao, J.; Hu, X.; Kong, T.; Yuan, Z.; Wang, C.; and Luo, P. 2020 · 2012
Earlier work this paper cites.
3D Object Representations for Fine-Grained Categorization
Krause, J.; Stark, M.; Deng, J.; and Fei-Fei, L. 2013 · 2013
Earlier work this paper cites.
Fine-grained visual classification of aircraft
Maji, S.; Rahtu, E.; Kannala, J.; Blaschko, M.; and Vedaldi, A. 2013 · 2013
Cited alongside, same era.
Bird species categorization using pose normalized deep convolutional nets
Branson, S.; Van Horn, G.; Belongie, S.; and Perona, P. 2014 · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K.; and Zisserman, A. 2014 · 2014
Cited alongside, same era.
Building a bird recognition app and large scale dataset with citizen scientists: The fine print in fine-grained dataset collection
Van Horn, G.; Branson, S.; Farrell, R.; Haber, S.; Barry, J.; Ipeirotis, P.; Perona, P.; and Belongie, S. 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
End-to-end object detection with transformers
Carion, N.; Massa, F.; Synnaeve, G.; Usunier, N.; Kirillov, A.; and Zagoruyko, S. 2020 · 2020
Later among the works it cites.
Fine-grained visual classification via progressive multi-granularity training of jigsaw patches
Du, R.; Chang, D.; Bhunia, A. K.; Xie, J.; Ma, Z.; Song, Y.-Z.; and Guo, J. 2020 · 2020
Later among the works it cites.
Channel Interaction Networks for Fine-Grained Image Categorization
Gao, Y.; Han, X.; Wang, X.; Huang, W.; and Scott, M. 2020 · 2020
Later among the works it cites.
Filtration and distillation: Enhancing region attention for fine-grained visual categorization
Liu, C.; Xie, H.; Zha, Z.-J.; Ma, L.; Yu, L.; and Zhang, Y. 2020 · 2020
Later among the works it cites.
Learning attentive pairwise interaction for fine-grained classification
Zhuang, P.; Wang, Y.; and Qiao, Y. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mask-cnn: Localizing parts and selecting descriptors for fine-grained image recognition
Wei, X.-S.; Xie, C.-W.; and Wu, J. 2016 · 2016
Cited alongside, same era.
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, L.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2018 · 2018
Cited alongside, same era.
The inaturalist species classification and detection dataset
Van Horn, G.; Mac Aodha, O.; Song, Y.; Cui, Y.; Sun, C.; Shepard, A.; Adam, H.; Perona, P.; and Belongie, S. 2018 · 2018
Cited alongside, same era.
Hierarchical bilinear pooling for fine-grained visual recognition
Yu, C.; Zhao, X.; Zheng, Q.; Zhang, P.; and You, X. 2018 · 2018
Cited alongside, same era.
Selective sparse sampling for fine-grained image recognition
Ding, Y.; Zhou, Y.; Zhu, Y.; Ye, Q.; and Jiao, J. 2019 · 2019
Cited alongside, same era.
Video action transformer network
Girdhar, R.; Carreira, J.; Doersch, C.; and Zisserman, A. 2019 · 2019
Cited alongside, same era.
Chen, J.; Lu, Y.; Yu, Q.; Luo, X.; Adeli, E.; Wang, Y.; Lu, L.; Yuille, A. L.; and Zhou, Y. 2021 · 2021
Closest in time.
Transreid: Transformer-based object re-identification
He, S.; Luo, H.; Wang, P.; Wang, F.; Li, H.; and Jiang, W. 2021 · 2021
Closest in time.
Max-deeplab: End-to-end panoptic segmentation with mask transformers
Wang, H.; Zhu, Y.; Adam, H.; Yuille, A.; and Chen, L.-C. 2021 · 2021
Closest in time.
Trans2Seg: Transparent Object Segmentation with Transformer
Xie, E.; Wang, W.; Wang, W.; Sun, P.; Xu, H.; Liang, D.; and Luo, P. 2021 · 2021
Closest in time.
Re-rank Coarse Classification with Local Region Enhanced Features for Fine-Grained Image Recognition
Yang, S.; Liu, S.; Yang, C.; and Wang, C. 2021 · 2021
Closest in time.
Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers
Zheng, S.; Lu, J.; Zhao, H.; Zhu, X.; Luo, Z.; Wang, Y.; Fu, Y.; Feng, J.; Xiang, T.; Torr, P. H.; et al. 2021 · 2021
Closest in time.