Fetching the paper…
Reading the bibliography…
Tracking multiple objects based on textual queries is a challenging task that requires linking language understanding with object association across frames.
Pattern matching: The gestalt approach
John W Ratcliff, David E Metzener, et al · 1988
Earlier work this paper cites.
An introduction to the kalman filter
Greg Welch, Gary Bishop, et al · 1995
Earlier work this paper cites.
Finding and tracking people from the bottom up
D. Ramanan and D.A. Forsyth · 2003
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Transtrack: Multiple object tracking with transformer
Peize Sun, Jinkun Cao, Yi Jiang, Rufeng Zhang, Enze Xie, Zehuan Yuan, Changhu Wang, and Ping Luo · 2012
Earlier work this paper cites.
Efficient agglomerative hierarchical clustering
Athman Bouguettaya, Qi Yu, Xumin Liu, Xiangmin Zhou, and Andy Song · 2015
Earlier work this paper cites.
Simple online and realtime tracking with a deep association metric
Nicolai Wojke, Alex Bewley, and Dietrich Paulus · 2017
Earlier work this paper cites.
Second: Sparsely embedded convolutional detection
Yan Yan, Yuxing Mao, and Bo Li · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding, 2019
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 2019
Earlier work this paper cites.
Pointrcnn: 3d object proposal generation and detection from point cloud
Shaoshuai Shi, Xiaogang Wang, and Hongsheng Li · 2019
Earlier work this paper cites.
Pv-rcnn: Point-voxel feature set abstraction for 3d object detection
Shaoshuai Shi, Chaoxu Guo, Li Jiang, Zhe Wang, Jianping Shi, Xiaogang Wang, and Hongsheng Li · 2020
Earlier work this paper cites.
3d multi-object tracking: A baseline and new evaluation metrics
Xinshuo Weng, Jianren Wang, David Held, and Kris Kitani · 2020
Earlier work this paper cites.
Tracking objects as points
Xingyi Zhou, Vladlen Koltun, and Philipp Krähenbühl · 2020
Cited alongside, same era.
Hota: A higher order metric for evaluating multi-object tracking
Jonathon Luiten, Aljosa Osep, Patrick Dendorfer, Philip Torr, Andreas Geiger, Laura Leal-Taixé, and Bastian Leibe · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Cited alongside, same era.
Learning to track with object permanence
Pavel Tokmakov, Jie Li, Wolfram Burgard, and Adrien Gaidon · 2021
Cited alongside, same era.
Fairmot: On the fairness of detection and re-identification in multiple object tracking
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenjun Zeng, and Wenyu Liu · 2021
Cited alongside, same era.
Monocular quasi-dense 3d object tracking
Referring multi-object tracking
Dongming Wu, Wencheng Han, Tiancai Wang, Xingping Dong, Xiangyu Zhang, and Jianbing Shen · 2023
Later among the works it cites.
Phi-4 technical report, 2024
Marah Abdin, Jyoti Aneja, Harkirat Behl, Sébastien Bubeck, Ronen Eldan, Suriya Gunasekar, Michael Harrison, Russell J. Hewett, Mojan Javaheripi, Piero Kauffmann, James R. Lee, Yin Tat Lee, Yuanzhi Li, Weishung Liu, Caio C. T. Mendes, Anh Nguyen, Eric Price, Gustavo de Rosa, Olli Saarikivi, Adil Salim, Shital Shah, Xin Wang, Rachel Ward, Yue Wu, Dingli Yu, Cyril Zhang, and Yi Zhang · 2024
Later among the works it cites.
ikun: Speak to trackers without retraining
Yunhao Du, Cheng Lei, Zhicheng Zhao, and Fei Su · 2024
Later among the works it cites.
Visual-linguistic representation learning with deep cross-modality fusion for referring multi-object tracking
Wenyan He, Yajun Jian, Yang Lu, and Hanzi Wang · 2024
Later among the works it cites.
Romot: Referring-expression-comprehension open-set multi-object tracking
Wei Li, Bowen Li, Jingqi Wang, Weiliang Meng, Jiguang Zhang, and Xiaopeng Zhang · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hou-Ning Hu, Yung-Hsu Yang, Tobias Fischer, Trevor Darrell, Fisher Yu, and Min Sun · 2022
Cited alongside, same era.
Trackformer: Multi-object tracking with transformers
Tim Meinhardt, Alexander Kirillov, Laura Leal-Taixe, and Christoph Feichtenhofer · 2022
Cited alongside, same era.
Object permanence emerges in a random walk along memory
Pavel Tokmakov, Allan Jabri, Jie Li, and Adrien Gaidon · 2022
Cited alongside, same era.
Casa: A cascade attention network for 3-d object detection from lidar point clouds
Hai Wu, Jinhao Deng, Chenglu Wen, Xin Li, Cheng Wang, and Jonathan Li · 2022
Cited alongside, same era.
Bytetrack: Multi-object tracking by associating every detection box
Yifu Zhang, Peize Sun, Yi Jiang, Dongdong Yu, Fucheng Weng, Zehuan Yuan, Ping Luo, Wenyu Liu, and Xinggang Wang · 2022
Cited alongside, same era.
Qwen technical report, 2023
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, Binyuan Hui, Luo Ji, Mei Li, Junyang Lin, Runji Lin, Dayiheng Liu, Gao Liu, Chengqiang Lu, Keming Lu, Jianxin Ma, Rui Men, Xingzhang Ren, Xuancheng Ren, Chuanqi Tan, Sinan Tan, Jianhong Tu, Peng Wang, Shijie Wang, Wei Wang, Shengguang Wu, Benfeng Xu, Jin Xu, An Yang, Hao Yang, Jian Yang, Shusheng Yang, Yang Yao, Bowen Yu, Hongyi Yuan, Zheng Yuan, Jianwei Zhang, Xingxuan Zhang, Yichang Zhang, Zhenru Zhang, Chang Zhou, Jingren Zhou, Xiaohuan Zhou, and Tianhang Zhu · 2023
Cited alongside, same era.
Observation-centric sort: Rethinking sort for robust multi-object tracking
Jinkun Cao, Jiangmiao Pang, Xinshuo Weng, Rawal Khirodkar, and Kris Kitani · 2023
Cited alongside, same era.
Echotrack: Auditory referring multi-object tracking for autonomous driving
Jiacheng Lin, Jiajun Chen, Kunyu Peng, Xuan He, Zhiyong Li, Rainer Stiefelhagen, and Kailun Yang · 2024
Later among the works it cites.
Mls-track: Multilevel semantic interaction in rmot
Zeliang Ma, Song Yang, Zhe Cui, Zhicheng Zhao, Fei Su, Delong Liu, and Jingyu Wang · 2024
Later among the works it cites.
Ucmctrack: Multi-object tracking with uniform camera motion compensation
Kefu Yi, Kai Luo, Xiaolei Luo, Jiangui Huang, Hao Wu, Rongdong Hu, and Wei Hao · 2024
Later among the works it cites.
Bootstrapping referring multi-object tracking
Yani Zhang, Dongming Wu, Wencheng Han, and Xingping Dong · 2024
Later among the works it cites.
Multi-granularity localization transformer with collaborative understanding for referring multi-object tracking
Jiajun Chen, Jiacheng Lin, Guojin Zhong, You Yao, and Zhiyong Li · 2025
Closest in time.
Hybridtrack: A hybrid approach for robust multi-object tracking
Leandro Di Bella, Yangxintong Lyu, Bruno Cornelis, and Adrian Munteanu · 2025
Closest in time.
Mex: Memory-efficient approach to referring multi-object tracking
Huu-Thien Tran, Phuoc-Sang Pham, Thai-Son Tran, and Khoa Luu · 2025
Closest in time.