Fetching the paper…
Reading the bibliography…
Recently, the vision transformer (ViT) has made breakthroughs in image recognition.
Imagenet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, et al · 2015
Earlier work this paper cites.
Learning with high-level attributes
Thao Nguyen · 2019
Earlier work this paper cites.
Image-to-image translation via group-wise deep whitening-and-coloring transformation
Wonwoong Cho, Sungha Choi, David Keetae Park, Inkyu Shin, and Jaegul Choo · 2019
Earlier work this paper cites.
Selective sparse sampling for fine-grained image recognition
Yao Ding, Yanzhao Zhou, Yi Zhu, Qixiang Ye, and Jianbin Jiao · 2019
Earlier work this paper cites.
Attention convolutional binary neural tree for fine-grained visual categorization
Ruyi Ji, Longyin Wen, Libo Zhang, Dawei Du, Yanjun Wu, Chen Zhao, Xianglong Liu, and Feiyue Huang · 2020
Earlier work this paper cites.
Sportscap: Monocular 3d human motion capture and fine-grained understanding in challenging sports videos
Xin Chen, Anqi Pang, Wei Yang, Yuexin Ma, Lan Xu, and Jingyi Yu · 2021
Earlier work this paper cites.
Multi-branch channel-wise enhancement network for fine-grained visual recognition
Guangjun Li, Yongxiong Wang, and Fengting Zhu · 2021
Earlier work this paper cites.
Dynamic perception framework for fine-grained recognition
Yao Ding, Zhenjun Han, Yanzhao Zhou, Yi Zhu, Jie Chen, Qixiang Ye, and Jianbin Jiao · 2021
Earlier work this paper cites.
Complemental attention multi-feature fusion network for fine-grained classification
Zhuang Miao, Xun Zhao, Jiabao Wang, Yang Li, and Hang Li · 2021
Cited alongside, same era.
Where to focus: Investigating hierarchical attention relationship for fine-grained visual classification
Yang Liu, Lei Zhou, Pengcheng Zhang, Xiao Bai, Lin Gu, Xiaohan Yu, Jun Zhou, and Edwin R Hancock · 2022
Cited alongside, same era.
Dual cross-attention learning for fine-grained visual categorization and object re-identification
Haowei Zhu, Wenjing Ke, Dong Li, Ji Liu, Lu Tian, and Yi Shan · 2022
Cited alongside, same era.
Progressive erasing network with consistency loss for fine-grained visual classification
Jin Peng, Yongxiong Wang, and Zeping Zhou · 2022
Cited alongside, same era.
Multi-view active fine-grained recognition
Ruoyi Du, Wenqing Yu, Heqing Wang, Dongliang Chang, Ting-En Lin, Yongbin Li, and Zhanyu Ma · 2022
Cited alongside, same era.
Swin unetr: Swin transformers for semantic segmentation of brain tumors in mri images
Ali Hatamizadeh, Vishwesh Nath, Yucheng Tang, Dong Yang, Holger R Roth, and Daguang Xu · 2022
Closest in time.
Actionformer: Localizing moments of actions with transformers
Chen-Lin Zhang, Jianxin Wu, and Yin Li · 2022
Closest in time.
Fat-net: Feature adaptive transformers for automated skin lesion segmentation
Huisi Wu, Shihuai Chen, Guilian Chen, Wei Wang, Baiying Lei, and Zhenkun Wen · 2022
Closest in time.
Cross-enhancement transformer for action segmentation
Jiahui Wang, Zhenyou Wang, Shanna Zhuang, and Hui Wang · 2022
Closest in time.
Multiscale progressive complementary fusion network for fine-grained visual classification
Jingsheng Lei, Xinqi Yang, and Shengying Yang · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improving image classification through joint guided learning
Peipei Zhao, Hang Yao, Xiangzeng Liu, Ruyi Liu, and Qiguang Miao · 2022
Cited alongside, same era.
Vision transformers are robust learners
Sayak Paul and Pin-Yu Chen · 2022
Cited alongside, same era.
Pyramid convolution and multi-frequency spatial attention for fine-grained visual categorization
Qin Xu, Yun Li, Mengquan Zhang, Zhifu Tao, and Bin Luo
Cited in the paper.
Davt: Data augmentation vision transformer for fine-grained visual categorization
Xiaobin Hu, Tao Jiang, Shining Zhu, Jia Guo, and Xiaotong Zhu
Cited in the paper.
Havt: Hierarchical attention vision transformer for fine-grained visual classification
Xiaobin Hu, Shining Zhu, and Taile Peng
Cited in the paper.
Yu Wang, Shuo Ye, Shujian Yu, and Xinge You · 2022
Closest in time.
Vit-net: Interpretable vision transformers with neural tree decoder
Sangwon Kim, Jaeyeal Nam, and Byoung Chul Ko · 2022
Closest in time.