Fetching the paper…
Reading the bibliography…
Language-image pre-training is an effective technique for learning powerful representations in general domains.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Object detection with discriminatively trained part-based models
Pedro F Felzenszwalb, Ross B Girshick, David McAllester, and Deva Ramanan · 2009
Earlier work this paper cites.
Detect what you can: Detecting and representing objects using holistic models and body parts
Xianjie Chen, Roozbeh Mottaghi, Xiaobai Liu, Sanja Fidler, Raquel Urtasun, and Alan Yuille · 2014
Earlier work this paper cites.
Pedestrian attribute recognition at far distance
Yubin Deng, Ping Luo, Chen Change Loy, and Xiaoou Tang · 2014
Earlier work this paper cites.
Multi-attribute learning for pedestrian attribute recognition in surveillance scenarios
Dangwei Li, Xiaotang Chen, and Kaiqi Huang · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Earlier work this paper cites.
Scalable person re-identification: A benchmark
Liang Zheng, Liyue Shen, Lu Tian, Shengjin Wang, Jingdong Wang, and Qi Tian · 2015
Earlier work this paper cites.
Attention to scale: Scale-aware semantic image segmentation
Liang-Chieh Chen, Yi Yang, Jiang Wang, Wei Xu, and Alan L Yuille · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Beyond triplet loss: a deep quadruplet network for person re-identification
Weihua Chen, Xiaotang Chen, Jianguo Zhang, and Kaiqi Huang · 2017
Earlier work this paper cites.
Vse++: Improving visual-semantic embeddings with hard negatives
Fartash Faghri, David J Fleet, Jamie Ryan Kiros, and Sanja Fidler · 2017
Earlier work this paper cites.
Look into person: Self-supervised structure-sensitive learning and a new benchmark for human parsing
Ke Gong, Xiaodan Liang, Dongyu Zhang, Xiaohui Shen, and Liang Lin · 2017
Earlier work this paper cites.
In defense of the triplet loss for person re-identification
Alexander Hermans, Lucas Beyer, and Bastian Leibe · 2017
Earlier work this paper cites.
Multiple-human parsing in the wild
Jianshu Li, Jian Zhao, Yunchao Wei, Congyan Lang, Yidong Li, Terence Sim, Shuicheng Yan, and Jiashi Feng · 2017
Earlier work this paper cites.
Person search with natural language description
Shuang Li, Tong Xiao, Hongsheng Li, Bolei Zhou, Dayu Yue, and Xiaogang Wang · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie · 2017
Earlier work this paper cites.
Neural person search machines
Hao Liu, Jiashi Feng, Zequn Jie, Karlekar Jayashree, Bo Zhao, Meibin Qi, Jianguo Jiang, and Shuicheng Yan · 2017
Earlier work this paper cites.
Hydraplus-net: Attentive deep features for pedestrian analysis
Xihui Liu, Haiyu Zhao, Maoqing Tian, Lu Sheng, Jing Shao, Junjie Yan, and Xiaogang Wang · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Yolo9000: better, faster, stronger
Joseph Redmon and Ali Farhadi · 2017
Earlier work this paper cites.
Attribute recognition by joint recurrent learning of context and correlation
Jingya Wang, Xiatian Zhu, Shaogang Gong, and Wei Li · 2017
Earlier work this paper cites.
Joint multi-person pose estimation and semantic part segmentation
Fangting Xia, Peng Wang, Xianjie Chen, and Alan L Yuille · 2017
Earlier work this paper cites.
Joint detection and identification feature learning for person search
Tong Xiao, Shuang Li, Bochao Wang, Liang Lin, and Xiaogang Wang · 2017
Earlier work this paper cites.
Joint detection and identification feature learning for person search
Tong Xiao, Shuang Li, Bochao Wang, Liang Lin, and Xiaogang Wang · 2017
Earlier work this paper cites.
Person re-identification in the wild
Liang Zheng, Hengheng Zhang, Shaoyan Sun, Manmohan Chandraker, Yi Yang, and Qi Tian · 2017
Earlier work this paper cites.
A discriminatively learned cnn embedding for person reidentification
Zhedong Zheng, Liang Zheng, and Yi Yang · 2017
Earlier work this paper cites.
Unlabeled samples generated by gan improve the person re-identification baseline in vitro
Zhedong Zheng, Liang Zheng, and Yi Yang · 2017
Earlier work this paper cites.
Re-ranking person re-identification with k-reciprocal encoding
Zhun Zhong, Liang Zheng, Donglin Cao, and Shaozi Li · 2017
Earlier work this paper cites.
Improving text-based person search by spatial matching and adaptive threshold
Tianlang Chen, Chenliang Xu, and Jiebo Luo · 2018
Earlier work this paper cites.
Stacked cross attention for image-text matching
Kuang-Huei Lee, Xi Chen, Gang Hua, Houdong Hu, and Xiaodong He · 2018
Earlier work this paper cites.
A richly annotated pedestrian dataset for person retrieval in real surveillance scenarios
Dangwei Li, Zhang Zhang, Xiaotang Chen, and Kaiqi Huang · 2018
Earlier work this paper cites.
Look into person: Joint body parsing & pose estimation network and a new benchmark
Xiaodan Liang, Ke Gong, Xiaohui Shen, and Liang Lin · 2018
Earlier work this paper cites.
Localization guided learning for pedestrian attribute recognition
Pengze Liu, Xihui Liu, Junjie Yan, and Jing Shao · 2018
Earlier work this paper cites.
Macro-micro adversarial network for human parsing
Yawei Luo, Zhedong Zheng, Liang Zheng, Tao Guan, Junqing Yu, and Yi Yang · 2018
Earlier work this paper cites.
Person re-identification with deep similarity-guided graph neural network
Yantao Shen, Hongsheng Li, Shuai Yi, Dapeng Chen, and Xiaogang Wang · 2018
Earlier work this paper cites.
Region-based quality estimation network for large-scale person re-identification
Guanglu Song, Biao Leng, Yu Liu, Congrui Hetang, and Shaofan Cai · 2018
Earlier work this paper cites.
Part-aligned bilinear representations for person re-identification
Yumin Suh, Jingdong Wang, Siyu Tang, Tao Mei, and Kyoung Mu Lee · 2018
Earlier work this paper cites.
Beyond part models: Person retrieval with refined part pooling (and a strong convolutional baseline)
Yifan Sun, Liang Zheng, Yi Yang, Qi Tian, and Shengjin Wang · 2018
Earlier work this paper cites.
Learning discriminative features with multiple granularities for person re-identification
Guanshuo Wang, Yufeng Yuan, Xiong Chen, Jiwei Li, and Xi Zhou · 2018
Earlier work this paper cites.
Person transfer gan to bridge domain gap for person re-identification
Longhui Wei, Shiliang Zhang, Wen Gao, and Qi Tian · 2018
Earlier work this paper cites.
Deep cross-modal projection learning for image-text matching
Ying Zhang and Huchuan Lu · 2018
Cited alongside, same era.
Abd-net: Attentive but diverse person re-identification
Tianlong Chen, Shaojin Ding, Jingyi Xie, Ye Yuan, Wuyang Chen, Yang Yang, Zhou Ren, and Zhangyang Wang · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Graphonomy: Universal human parsing via graph transfer learning
Ke Gong, Yiming Gao, Xiaodan Liang, Xiaohui Shen, Meng Wang, and Liang Lin · 2019
Cited alongside, same era.
Visual-semantic graph reasoning for pedestrian attribute recognition
Qiaozhe Li, Xin Zhao, Ran He, and Kaiqi Huang · 2019
Cited alongside, same era.
Unified vision-language pre-training for image captioning and vqa
Weperson: Learning a generalized re-identification model from all-weather virtual data
He Li, Mang Ye, and Bo Du · 2021
Later among the works it cites.
Sequential end-to-end network for efficient person search
Zhengjia Li and Duoqian Miao · 2021
Later among the works it cites.
TransMatcher: Deep Image Matching Through Transformers for Generalizable Person Re-identification
Shengcai Liao and Ling Shao · 2021
Later among the works it cites.
Clipcap: Clip prefix for image captioning
Ron Mokady, Amir Hertz, and Amit H Bermano · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhou Luowei, Palangi Hamid, Zhang Lei, Hu Houdong, Jason J. Corso, and Gao Jianfeng · 2019
Cited alongside, same era.
Query-guided end-to-end person search
Bharti Munjal, Sikandar Amin, Federico Tombari, and Fabio Galasso · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Devil in the details: Towards accurate single and multiple human parsing
Tao Ruan, Ting Liu, Zilong Huang, Yunchao Wei, Shikui Wei, and Yao Zhao · 2019
Cited alongside, same era.
Adversarial representation learning for text-to-image matching
Nikolaos Sarafianos, Xiang Xu, and Ioannis A Kakadiaris · 2019
Cited alongside, same era.
Learning compositional neural information fusion for human parsing
Wenguan Wang, Zhijie Zhang, Siyuan Qi, Jianbing Shen, Yanwei Pang, and Ling Shao · 2019
Cited alongside, same era.
Learning context graph for person search
Yichao Yan, Qiang Zhang, Bingbing Ni, Wendong Zhang, Minghao Xu, and Xiaokang Yang · 2019
Cited alongside, same era.
Lightningdot: Pre-training visual-semantic embeddings for real-time image-text retrieval
Siqi Sun, Yen-Chun Chen, Linjie Li, Shuohang Wang, Yuwei Fang, and Jingjing Liu · 2021
Later among the works it cites.
Lapscore: language-guided person search via color reasoning
Yushuang Wu, Zizheng Yan, Xiaoguang Han, Guanbin Li, Changqing Zou, and Shuguang Cui · 2021
Later among the works it cites.
Vtbr: Semantic-based pretraining for person re-identification
Suncheng Xiang, Zirui Zhang, Mengyuan Guan, Hao Chen, Binjie Yan, Ting Liu, and Yuzhuo Fu · 2021
Later among the works it cites.
Unrealperson: An adaptive pipeline towards costless person re-identification
Tianyu Zhang, Lingxi Xie, Longhui Wei, Zijie Zhuang, Yongfei Zhang, Bo Li, and Qi Tian · 2021
Later among the works it cites.
Fairmot: On the fairness of detection and re-identification in multiple object tracking
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenjun Zeng, and Wenyu Liu · 2021
Later among the works it cites.
Viewpoint transform matching model for person re-identification
Ruochen Zheng, Changxin Gao, and Nong Sang · 2021
Later among the works it cites.
Dssl: deep surroundings-person separation learning for text-based person retrieval
Aichun Zhu, Zijie Wang, Yifeng Li, Xili Wan, Jing Jin, Tian Wang, Fangqiang Hu, and Gang Hua · 2021
Later among the works it cites.
Context autoencoder for self-supervised representation learning
Xiaokang Chen, Mingyu Ding, Xiaodi Wang, Ying Xin, Shentong Mo, Yunhao Wang, Shumin Han, Ping Luo, Gang Zeng, and Jingdong Wang · 2022
Later among the works it cites.
Tipcb: A simple but effective part-based convolutional baseline for text-based person search
Yuhao Chen, Guoqing Zhang, Yujiang Lu, Zhenxing Wang, and Yuhui Zheng · 2022
Later among the works it cites.
A simple visual-textual baseline for pedestrian attribute recognition
Xinhua Cheng, Mengxi Jia, Qian Wang, and Jian Zhang · 2022
Later among the works it cites.
Part-based pseudo label refinement for unsupervised person re-identification
Yoonki Cho, Woo Jae Kim, Seunghoon Hong, and Sung-Eui Yoon · 2022
Later among the works it cites.
Large-scale pre-training for person re-identification with noisy labels
Dengpan Fu, Dongdong Chen, Hao Yang, Jianmin Bao, Lu Yuan, Lei Zhang, Houqiang Li, Fang Wen, and Dong Chen · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Later among the works it cites.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi · 2022
Later among the works it cites.
Self-correction for human parsing
Peike Li, Yunqiu Xu, Yunchao Wei, and Yi Yang · 2022
Later among the works it cites.
Label2label: A language modeling framework for multi-attribute learning
Wanhua Li, Zhexuan Cao, Jianjiang Feng, Jie Zhou, and Jiwen Lu · 2022
Later among the works it cites.
Cdgnet: Class distribution guided network for human parsing
Kunliang Liu, Ouk Choi, Jianming Wang, and Wonjun Hwang · 2022
Later among the works it cites.
Learning granularity-unified representations for text-to-image person re-identification
Zhiyin Shao, Xinyu Zhang, Meng Fang, Zhifeng Lin, Jian Wang, and Changxing Ding · 2022
Later among the works it cites.
See finer, see more: Implicit modality alignment for text-based person retrieval
Xiujun Shu, Wei Wen, Haoqian Wu, Keyu Chen, Yiran Song, Ruizhi Qiao, Bo Ren, and Xiao Wang · 2022
Later among the works it cites.
Adan: Adaptive nesterov momentum algorithm for faster optimizing deep models
Xingyu Xie, Pan Zhou, Huan Li, Zhouchen Lin, and Shuicheng Yan · 2022
Later among the works it cites.
Simmim: A simple framework for masked image modeling
Zhenda Xie, Zheng Zhang, Yue Cao, Yutong Lin, Jianmin Bao, Zhuliang Yao, Qi Dai, and Han Hu · 2022
Later among the works it cites.
Cloning Outfits from Real-World Images to 3D Characters for Generalizable Person Re-Identification
Xuezhi Liang Yanan Wang and Shengcai Liao · 2022
Later among the works it cites.
Unleashing potential of unsupervised pre-training with intra-identity regularization for person re-identification
Zizheng Yang, Xin Jin, Kecheng Zheng, and Feng Zhao · 2022
Later among the works it cites.
Implicit sample extension for unsupervised person re-identification
Xinyu Zhang, Dongdong Li, Zhigang Wang, Jian Wang, Errui Ding, Javen Qinfeng Shi, Zhaoxiang Zhang, and Jingdong Wang · 2022
Later among the works it cites.
Pass: Part-aware self-supervised pre-training for person re-identification
Kuan Zhu, Haiyun Guo, Tianyi Yan, Yousong Zhu, Jinqiao Wang, and Ming Tang · 2022
Later among the works it cites.
Rasa: Relation and sensitivity aware representation learning for text-based person search
Yang Bai, Min Cao, Daming Gao, Ziqiang Cao, Chen Chen, Zhenfeng Fan, Liqiang Nie, and Min Zhang · 2023
Closest in time.
Beyond appearance: a semantic controllable self-supervised learning framework for human-centric visual tasks
Weihua Chen, Xianzhe Xu, Jian Jia, Hao Luo, Yaohua Wang, Fan Wang, Rong Jin, and Xiuyu Sun · 2023
Closest in time.
Cross-modal implicit relation reasoning and aligning for text-to-image person retrieval
Ding Jiang and Mang Ye · 2023
Closest in time.
BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi · 2023
Closest in time.
Movienet-ps: A large-scale person search dataset in the wild
Jie Qin, Peng Zheng, Yichao Yan, Quan Rong, Xiaogang Cheng, and Bingbing Ni · 2023
Closest in time.
Self-supervised learning from untrimmed videos via hierarchical consistency
Zhiwu Qing, Shiwei Zhang, Ziyuan Huang, Yi Xu, Xiang Wang, Changxin Gao, Rong Jin, and Nong Sang · 2023
Closest in time.
Dual pseudo-labels interactive self-training for semi-supervised visible-infrared person re-identification
Jiangming Shi, Yachao Zhang, Xiangbo Yin, Yuan Xie, Zhizhong Zhang, Jianping Fan, Zhongchao Shi, and Yanyun Qu · 2023
Closest in time.
Clip-driven fine-grained text-image person re-identification, 2023
Shuanglin Yan, Neng Dong, Liyan Zhang, and Jinhui Tang · 2023
Closest in time.
Towards unified text-based person retrieval: A large-scale multi-attribute and language search benchmark
Shuyu Yang, Yinan Zhou, Yaxiong Wang, Yujiao Wu, Li Zhu, and Zhedong Zheng · 2023
Closest in time.
Understanding self-supervised pretraining with part-aware representation learning
Jie Zhu, Jiyang Qi, Mingyu Ding, Xiaokang Chen, Ping Luo, Xinggang Wang, Wenyu Liu, Leye Wang, and Jingdong Wang · 2023
Closest in time.
Ufinebench: Towards text-based person retrieval with ultra-fine granularity
Jialong Zuo, Hanyu Zhou, Ying Nie, Feng Zhang, Tianyu Guo, Nong Sang, Yunhe Wang, and Changxin Gao · 2023
Closest in time.
Multi-memory matching for unsupervised visible-infrared person re-identification
Jiangming Shi, Xiangbo Yin, Yeyun Chen, Yachao Zhang, Zhizhong Zhang, Yuan Xie, and Yanyun Qu · 2024
Closest in time.