Fetching the paper…
Reading the bibliography…
Existing pedestrian attribute recognition (PAR) algorithms adopt pre-trained CNN (e.g., ResNet) as their backbone network for visual feature learning, which might obtain sub-optimal results due to the insufficient employment of the relations between pedestrian images and attribute labels.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell
1901
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in
2009
Earlier work this paper cites.
I. Sutskever, J. Martens, G. Dahl, and G. Hinton, “On the importance of initialization and momentum in deep learning,” in
2013
Earlier work this paper cites.
N. Zhang, M. Paluri, M. Ranzato, T. Darrell, and L. Bourdev, “Panda: Pose aligned networks for deep attribute modeling,” in
2014
Earlier work this paper cites.
Y. Deng, P. Luo, C. C. Loy, and X. Tang, “Pedestrian attribute recognition at far distance,” in
2014
Earlier work this paper cites.
Y. Tian, P. Luo, X. Wang, and X. Tang, “Pedestrian detection aided by deep learning semantic tasks,” in
2015
Earlier work this paper cites.
A. H. Abdulnabi, G. Wang, J. Lu, and K. Jia, “Multi-task cnn model for attribute prediction,”
2015
Earlier work this paper cites.
D. Li, X. Chen, and K. Huang, “Multi-attribute learning for pedestrian attribute recognition in surveillance scenarios,” in
2015
Earlier work this paper cites.
Z. Liu, P. Luo, X. Wang, and X. Tang, “Deep learning face attributes in the wild,” in
2015
Earlier work this paper cites.
G. Gkioxari, R. Girshick, and J. Malik, “Contextual action recognition with r*cnn,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Y. Li, C. Huang, C. C. Loy, and X. Tang, “Human attribute recognition by deep hierarchical contexts,” in
2016
Earlier work this paper cites.
J. Wang, X. Zhu, S. Gong, and W. Li, “Attribute recognition by joint recurrent learning of context and correlation,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,”
2017
Earlier work this paper cites.
X. Liu, H. Zhao, M. Tian, L. Sheng, J. Shao, S. Yi, J. Yan, and X. Wang, “Hydraplus-net: Attentive deep features for pedestrian analysis,” in
2017
Earlier work this paper cites.
Q. Dong, S. Gong, and X. Zhu, “Multi-task curriculum transfer deep learning of clothing attributes,” in
2017
Earlier work this paper cites.
F. Zhu, H. Li, W. Ouyang, N. Yu, and X. Wang, “Learning spatial regularization with image-level supervisions for multi-label image classification,” in
2017
Earlier work this paper cites.
M. M. Kalayeh, B. Gong, and M. Shah, “Improving facial attribute prediction using semantic segmentation,” in
2017
Earlier work this paper cites.
E. Hand and R. Chellappa, “Attributes for improved attributes: A multi-task network utilizing implicit and explicit relationships for facial attribute classification,” in
2017
Earlier work this paper cites.
A. Tealab, “Time series forecasting using artificial neural networks methodologies: A systematic review,”
2018
Earlier work this paper cites.
X. Zhao, L. Sang, G. Ding, Y. Guo, and X. Jin, “Grouping attribute recognition for pedestrian with joint recurrent learning.” in
2018
Earlier work this paper cites.
N. Sarafianos, X. Xu, and I. A. Kakadiaris, “Deep imbalanced attribute classification using visual attention aggregation,” in
2018
Earlier work this paper cites.
N. Sarafianos, X. Xu, and I. A. Kakadiaris, “Deep imbalanced attribute classification using visual attention aggregation,” in
2018
Earlier work this paper cites.
K. He, Y. Fu, W. Zhang, C. Wang, Y.-G. Jiang, F. Huang, and X. Xue, “Harnessing synthesized abstraction images to improve facial attribute recognition.” in
2018
Earlier work this paper cites.
J. Li, F. Zhao, J. Feng, S. Roy, S. Yan, and T. Sim, “Landmark free face attribute prediction,”
2018
Earlier work this paper cites.
J. D. M.-W. C. Kenton and L. K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in
2019
Cited alongside, same era.
D. Li, Z. Zhang, X. Chen, and K. Huang, “A richly annotated pedestrian dataset for person retrieval in real surveillance scenarios,”
2019
Cited alongside, same era.
2019
Cited alongside, same era.
H. Guo, K. Zheng, X. Fan, H. Yu, and S. Wang, “Visual attention consistency under image transforms for multi-label image classification,” in
2019
Cited alongside, same era.
C. Tang, L. Sheng, Z.-X. Zhang, and X. Hu, “Improving pedestrian attribute recognition with weakly-supervised multi-scale attribute-specific localization,” in
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,”
2022
Later among the works it cites.
Zhou, Kaiyang and Yang, Jingkang and Loy, Chen Change and Liu, Ziwei, “Conditional prompt learning for vision-language models,” in
2022
Later among the works it cites.
R. Zhang, W. Zhang, R. Fang, P. Gao, K. Li, J. Dai, Y. Qiao, and H. Li, “Tip-adapter: Training-free adaption of clip for few-shot classification,” in
2022
Later among the works it cites.
R. Zhang, Z. Guo, W. Zhang, K. Li, X. Miao, B. Cui, Y. Qiao, P. Gao, and H. Li, “Pointclip: Point cloud understanding by clip,” in
2022
Later among the works it cites.
J. Wu, Y. Huang, Z. Gao, Y. Hong, J. Zhao, and X. Du, “Inter-attribute awareness for pedestrian attribute recognition,”
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in
2020
Cited alongside, same era.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly
2020
Cited alongside, same era.
X. Li, X. Yin, C. Li, P. Zhang, X. Hu, L. Zhang, L. Wang, H. Hu, L. Dong, F. Wei
2020
Cited alongside, same era.
Z. Tan, Y. Yang, J. Wan, G. Guo, and S. Z. Li, “Relation-aware pedestrian attribute recognition with graph convolutional networks,”
2020
Cited alongside, same era.
J. Wu, H. Liu, J. Jiang, M. Qi, B. Ren, X. Li, and Y. Wang, “Person attribute recognition by sequence contextual relation learning,”
2020
Cited alongside, same era.
M. Wu, D. Huang, Y. Guo, and Y. Wang, “Distraction-aware feature learning for human attribute recognition via coarse-to-fine attention mechanism,” in
2020
Cited alongside, same era.
L. Mao, Y. Yan, J.-H. Xue, and H. Wang, “Deep multi-task multi-label cnn for effective facial attribute classification,”
2020
Cited alongside, same era.
L. C. andJingkuan Song andXuerui Zhang andMingsheng Shang, “Mcfl: multi-label contrastive focal loss for deep imbalanced pedestrian attribute recognition,”
2022
Later among the works it cites.
Z. Tang and J. Huang, “Drformer: Learning dual relations using transformer for pedestrian attribute recognition,”
2022
Later among the works it cites.
H. Fan, H.-M. Hu, S. Liu, W. Lu, and S. Pu, “Correlation graph convolutional network for pedestrian attribute recognition,”
2022
Later among the works it cites.
M. Wang, J. Xing, J. Mei, Y. Liu, and Y. Jiang, “Actionclip: Adapting language-image pretrained models for video action recognition,”
2023
Closest in time.
P. Gao, S. Geng, R. Zhang, T. Ma, R. Fang, Y. Zhang, H. Li, and Y. Qiao, “Clip-adapter: Better vision-language models with feature adapters,”
2023
Closest in time.
Z. Guo, R. Zhang, L. Qiu, X. Ma, X. Miao, X. He, and B. Cui, “Calip: Zero-shot enhancement of clip with parameter-free attention,” in
2023
Closest in time.
X. Wang, G. Chen, G. Qian, P. Gao, X.-Y. Wei, Y. Wang, Y. Tian, and W. Gao, “Large-scale multi-modal pre-trained models: A comprehensive survey,”
2023
Closest in time.
X. Zhu, R. Zhang, B. He, Z. Guo, Z. Zeng, Z. Qin, S. Zhang, and P. Gao, “Pointclip v2: Prompting clip and gpt for powerful 3d open-world learning,” in
2023
Closest in time.
L. Xue, M. Gao, C. Xing, R. Martín-Martín, J. Wu, C. Xiong, R. Xu, J. C. Niebles, and S. Savarese, “Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding,” in
2023
Closest in time.
Z. Guo, R. Zhang, L. Qiu, X. Li, and P.-A. Heng, “Joint-mae: 2d-3d joint masked autoencoders for 3d point cloud pre-training,” in
2023
Closest in time.
R. Girdhar, A. El-Nouby, Z. Liu, M. Singh, K. V. Alwala, A. Joulin, and I. Misra, “Imagebind: One embedding space to bind them all,” in
2023
Closest in time.
Z. Guo, R. Zhang, X. Zhu, Y. Tang, X. Ma, J. Han, K. Chen, P. Gao, X. Li, H. Li
2023
Closest in time.
R. Abdelfattah, Q. Guo, X. Li, X. Wang, and S. Wang, “Cdul: Clip-driven unsupervised learning for multi-label image classification,” in
2023
Closest in time.
X. Fan, Y. Zhang, Y. Lu, and H. Wang, “Parformer: Transformer-based multi-task network for pedestrian attribute recognition,”
2023
Closest in time.
S. Chen, X. Zhu, Y. Yan, S. Zhu, S.-Z. Li, and D.-H. Wang, “Identity-aware contrastive knowledge distillation for facial attribute recognition,”
2023
Closest in time.
A. Zheng, H. Wang, J. Wang, H. Huang, R. He, and A. Hussain, “Diverse features discovery transformer for pedestrian attribute recognition,”
2023
Closest in time.
L. Xue, N. Yu, S. Zhang, A. Panagopoulou, J. Li, R. Martín-Martín, J. Wu, C. Xiong, R. Xu, J. C. Niebles
2024
Closest in time.
B. Zhu, B. Lin, M. Ning, Y. Yan, J. Cui, W. HongFa, Y. Pang, W. Jiang, J. Zhang, Z. Li
2024
Closest in time.
R. Zhang, J. Han, C. Liu, A. Zhou, P. Lu, Y. Qiao, H. Li, and P. Gao, “LLaMA-adapter: Efficient fine-tuning of large language models with zero-initialized attention,” in
2024
Closest in time.
S. Huang, X. Li, Z.-Q. Cheng, Z. Zhang, and A. Hauptmann, “Gnas: A greedy neural architecture search method for multi-attribute learning,” in
2057
Closest in time.