Fetching the paper…
Reading the bibliography…
Existing text-based person retrieval datasets often have relatively coarse-grained text annotations.
Look before you leap: Improving text-based person retrieval by learning a consistent cross-modal common manifold
Zijie Wang, Aichun Zhu, Jingyi Xue, Xili Wan, Chao Liu, Tian Wang, and Yifeng Li · 1992
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Scalable person re-identification: A benchmark
Liang Zheng, Liyue Shen, Lu Tian, Shengjin Wang, Jingdong Wang, and Qi Tian · 2015
Earlier work this paper cites.
Person search with natural language description
Shuang Li, Tong Xiao, Hongsheng Li, Bolei Zhou, Dayu Yue, and Xiaogang Wang · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Unlabeled samples generated by gan improve the person re-identification baseline in vitro
Zhedong Zheng, Liang Zheng, and Yi Yang · 2017
Earlier work this paper cites.
Attribute-based person retrieval and search in video sequences
Arne Schumann, Andreas Specker, and Jürgen Beyerer · 2018
Earlier work this paper cites.
Person transfer gan to bridge domain gap for person re-identification
Longhui Wei, Shiliang Zhang, Wen Gao, and Qi Tian · 2018
Earlier work this paper cites.
Deep cross-modal projection learning for image-text matching
Ying Zhang and Huchuan Lu · 2018
Earlier work this paper cites.
Adversarial representation learning for text-to-image matching
Nikolaos Sarafianos, Xiang Xu, and Ioannis A Kakadiaris · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Earlier work this paper cites.
Vitaa: Visual-textual attributes alignment in person search by natural language
Zhe Wang, Zhiyuan Fang, Jun Wang, and Yezhou Yang · 2020
Earlier work this paper cites.
Dual-path convolutional image-text embeddings with instance loss
Zhedong Zheng, Liang Zheng, Michael Garrett, Yi Yang, Mingliang Xu, and Yi-Dong Shen · 2020
Earlier work this paper cites.
Semantically self-aligned network for text-to-image part-aware person re-identification
Zefeng Ding, Changxing Ding, Zhiyin Shao, and Dacheng Tao · 2021
Earlier work this paper cites.
Dsa-pr: discrete soft biometric attribute-based person retrieval in surveillance videos
Hiren Galiyawala, Mehul S Raval, and Dhyey Savaliya · 2021
Cited alongside, same era.
Contextual non-local alignment over full-scale representation for text-based person search
Chenyang Gao, Guanyu Cai, Xinyang Jiang, Feng Zheng, Jun Zhang, Yifei Gong, Pai Peng, Xiaowei Guo, and Xing Sun · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Improving attribute-based person retrieval by using a calibrated, weighted, and distribution-based distance metric
Andreas Specker and Jürgen Beyerer · 2021
Cited alongside, same era.
Lapscore: language-guided person search via color reasoning
Yushuang Wu, Zizheng Yan, Xiaoguang Han, Guanbin Li, Changqing Zou, and Shuguang Cui · 2021
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Closest in time.
Upar challenge: Pedestrian attribute recognition and attribute-based person retrieval–dataset, design, and results
Mickael Cormier, Andreas Specker, Julio Junior, CS Jacques, Lucas Florin, Jürgen Metzler, Thomas B Moeslund, Kamal Nasrollahi, Sergio Escalera, and Jürgen Beyerer · 2023
Closest in time.
Cross-modal implicit relation reasoning and aligning for text-to-image person retrieval
Ding Jiang and Mang Ye · 2023
Closest in time.
BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi · 2023
Closest in time.
Lightclip: Learning multi-level interaction for lightweight vision-language models
Ying Nie, Wei He, Kai Han, Yehui Tang, Tianyu Guo, Fanyi Du, and Yunhe Wang · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Filip: Fine-grained interactive language-image pre-training
Lewei Yao, Runhui Huang, Lu Hou, Guansong Lu, Minzhe Niu, Hang Xu, Xiaodan Liang, Zhenguo Li, Xin Jiang, and Chunjing Xu · 2021
Cited alongside, same era.
Fairmot: On the fairness of detection and re-identification in multiple object tracking
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenjun Zeng, and Wenyu Liu · 2021
Cited alongside, same era.
Dssl: Deep surroundings-person separation learning for text-based person retrieval
Aichun Zhu, Zijie Wang, Yifeng Li, Xili Wan, Jing Jin, Tian Wang, Fangqiang Hu, and Gang Hua · 2021
Cited alongside, same era.
Tipcb: A simple but effective part-based convolutional baseline for text-based person search
Yuhao Chen, Guoqing Zhang, Yujiang Lu, Zhenxing Wang, and Yuhui Zheng · 2022
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Cited alongside, same era.
Large-scale pre-training for person re-identification with noisy labels
Dengpan Fu, Dongdong Chen, Hao Yang, Jianmin Bao, Lu Yuan, Lei Zhang, Houqiang Li, Fang Wen, and Dong Chen · 2022
Cited alongside, same era.
Person retrieval in surveillance videos using attribute recognition
Hiren Galiyawala, Mehul S Raval, and Meet Patel · 2022
Cited alongside, same era.
Closest in time.
Fine-grained image-text matching by cross-modal hard aligning network
Zhengxin Pan, Fangyu Wu, and Bailing Zhang · 2023
Closest in time.
Attribute based spatio-temporal person retrieval in video surveillance
Rasha Shoitan, Mona M Moussa, and Heba A El Nemr · 2023
Closest in time.
Upar: Unified pedestrian attribute recognition and person retrieval
Andreas Specker, Mickael Cormier, and Jürgen Beyerer · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto · 2023
Closest in time.
Attribute-wise reasoning reinforcement learning for pedestrian attribute retrieval
Yaodong Wang, Zhenfei Hu, and Zhong Ji · 2023
Closest in time.
Clip-driven fine-grained text-image person re-identification
Shuanglin Yan, Neng Dong, Liyan Zhang, and Jinhui Tang · 2023
Closest in time.
Plip: Language-image pre-training for person representation learning
Jialong Zuo, Changqian Yu, Nong Sang, and Changxin Gao · 2023
Closest in time.
Species196: A one-million semi-supervised dataset for fine-grained species recognition
Wei He, Kai Han, Ying Nie, Chengcheng Wang, and Yunhe Wang · 2024
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2024
Closest in time.