Fetching the paper…
Reading the bibliography…
Existing online multiple object tracking (MOT) algorithms often consist of two subtasks, detection and re-identification (ReID).
H. W. Kuhn, “The hungarian method for the assignment problem,” Naval Research Logistics Quarterly , vol. 2, no. 1-2, pp. 83–97, 1955
1955
Earlier work this paper cites.
J. Müller, “The hadamard multiplication theorem and applications in summability theory,” Complex Variables and Elliptic Equations , vol. 18, no. 3-4, pp. 155–166, 1992
1992
Earlier work this paper cites.
G. Welch, G. Bishop et al. , “An introduction to the kalman filter,” 1995
1995
Earlier work this paper cites.
A. Ess, B. Leibe, K. Schindler, and L. Van Gool, “A mobile vision system for robust multi-person tracking,” in 2008 IEEE Conference on Computer Vision and Pattern Recognition , 2008, pp. 1–8
2008
Earlier work this paper cites.
K. Bernardin and R. Stiefelhagen, “Evaluating multiple object tracking performance: the clear mot metrics,” EURASIP Journal on Image and Video Processing , vol. 2008, pp. 1–10, 2008
2008
Earlier work this paper cites.
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan, “Object detection with discriminatively trained part-based models,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 32, no. 9, pp. 1627–1645, 2009
2009
Earlier work this paper cites.
P. Dollár, C. Wojek, B. Schiele, and P. Perona, “Pedestrian detection: A benchmark,” in 2009 IEEE Conference on Computer Vision and Pattern Recognition , 2009, pp. 304–311
2009
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in European Conference on Computer Vision , 2014, pp. 740–755
2014
Earlier work this paper cites.
Y. Xiang, A. Alahi, and S. Savarese, “Learning to track: Online multi-object tracking by decision making,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 4705–4713
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
F. Yu, W. Li, Q. Li, Y. Liu, X. Shi, and J. Yan, “Poi: Multiple object tracking with high performance detection and appearance feature,” in European Conference on Computer Vision , 2016, pp. 36–42
2016
Earlier work this paper cites.
A. Bewley, Z. Ge, L. Ott, F. Ramos, and B. Upcroft, “Simple online and realtime tracking,” in 2016 IEEE International Conference on Image Processing , 2016, pp. 3464–3468
2016
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: towards real-time object detection with region proposal networks,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 6, pp. 1137–1149, 2016
2016
Earlier work this paper cites.
L. Bertinetto, J. Valmadre, J. F. Henriques, A. Vedaldi, and P. H. Torr, “Fully-convolutional siamese networks for object tracking,” in European Conference on Computer Vision , 2016, pp. 850–865
2016
Earlier work this paper cites.
R. Tao, E. Gavves, and A. W. Smeulders, “Siamese instance search for tracking,” in Proceedings of the IEEE Conference on Computer vision and Pattern Recognition , 2016, pp. 1420–1429
2016
Earlier work this paper cites.
F. Yang, W. Choi, and Y. Lin, “Exploit all the layers: Fast and accurate cnn object detector with scale dependent pooling and cascaded rejection classifiers,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 2129–2137
2016
Earlier work this paper cites.
K. Kang, H. Li, J. Yan, X. Zeng, B. Yang, T. Xiao, C. Zhang, Z. Wang, R. Wang, X. Wang et al. , “T-cnn: Tubelets with convolutional neural networks for object detection from videos,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 28, no. 10, pp. 2896–2907, 2017
2017
Earlier work this paper cites.
N. Takahashi, M. Gygli, and L. Van Gool, “Aenet: Learning deep audio features for video analysis,” IEEE Transactions on Multimedia , vol. 20, no. 3, pp. 513–524, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in Conference on Neural Information Processing Systems , 2017
2017
Earlier work this paper cites.
N. Wojke, A. Bewley, and D. Paulus, “Simple online and realtime tracking with a deep association metric,” in 2017 IEEE International Conference on Image Processing , 2017, pp. 3645–3649
2017
Earlier work this paper cites.
X. Zhu, Y. Wang, J. Dai, L. Yuan, and Y. Wei, “Flow-guided feature aggregation for video object detection,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 408–417
2017
Earlier work this paper cites.
C. Feichtenhofer, A. Pinz, and A. Zisserman, “Detect to track and track to detect,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 3038–3046
2017
Cited alongside, same era.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 2980–2988
2017
Cited alongside, same era.
S. Zhang, R. Benenson, and B. Schiele, “Citypersons: A diverse dataset for pedestrian detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 3213–3221
2017
Cited alongside, same era.
T. Xiao, S. Li, B. Wang, L. Lin, and X. Wang, “Joint detection and identification feature learning for person search,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 3415–3424
2017
Cited alongside, same era.
Z. Sun, J. Chen, C. Liang, W. Ruan, and M. Mukherjee, “A survey of multiple pedestrian tracking based on tracking-by-detection framework,” IEEE Transactions on Circuits and Systems for Video Technology , 2020
2020
Later among the works it cites.
S. Zhang, Q. Zhang, Y. Yang, X. Wei, P. Wang, B. Jiao, and Y. Zhang, “Person re-identification in aerial imagery,” IEEE Transactions on Multimedia , vol. 23, pp. 281–291, 2020
2020
Later among the works it cites.
G. Ciaparrone, F. L. Sánchez, S. Tabik, L. Troiano, R. Tagliaferri, and F. Herrera, “Deep learning in video multi-object tracking: A survey,” Neurocomputing , vol. 381, pp. 61–88, 2020
2020
Later among the works it cites.
X. Weng, Y. Wang, Y. Man, and K. M. Kitani, “Gnn3dmot: Graph neural network for 3d multi-object tracking with 2d-3d multi-feature learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 6499–6508
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Luo, B. Yang, and R. Urtasun, “Fast and furious: Real time end-to-end 3d detection, tracking and motion forecasting with a single convolutional net,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2018, pp. 3569–3577
2018
Cited alongside, same era.
Z. Cai and N. Vasconcelos, “Cascade r-cnn: Delving into high quality object detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 6154–6162
2018
Cited alongside, same era.
H. Law and J. Deng, “Cornernet: Detecting objects as paired keypoints,” in Proceedings of the European Conference on Computer Vision , 2018, pp. 734–750
2018
Cited alongside, same era.
L. Chen, H. Ai, Z. Zhuang, and C. Shang, “Real-time multiple people tracking with deeply learned candidate selection and person re-identification,” in 2018 IEEE International Conference on Multimedia and Expo , 2018, pp. 1–6
2018
Cited alongside, same era.
J. Zhu, H. Yang, N. Liu, M. Kim, W. Zhang, and M.-H. Yang, “Online multi-object tracking with dual matching attention networks,” in Proceedings of the European Conference on Computer Vision , 2018, pp. 366–382
2018
Cited alongside, same era.
2018
Cited alongside, same era.
B. Li, J. Yan, W. Wu, Z. Zhu, and X. Hu, “High performance visual tracking with siamese region proposal network,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 8971–8980
2018
Cited alongside, same era.
Z. Zhong, L. Zheng, Z. Zheng, S. Li, and Y. Yang, “Camstyle: A novel data augmentation method for person re-identification,” IEEE Transactions on Image Processing , vol. 28, no. 3, pp. 1176–1190, 2018
2018
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
Z. Zhang, C. Lan, W. Zeng, X. Jin, and Z. Chen, “Relation-aware global attention for person re-identification,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 3186–3195
2020
Later among the works it cites.
M. Yin, Z. Yao, Y. Cao, X. Li, Z. Zhang, S. Lin, and H. Hu, “Disentangled non-local neural networks,” in European Conference on Computer Vision , 2020, pp. 191–207
2020
Later among the works it cites.
2020
Later among the works it cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in European Conference on Computer Vision . Springer, 2020, pp. 213–229
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Peng, C. Wang, F. Wan, Y. Wu, Y. Wang, Y. Tai, C. Wang, J. Li, F. Huang, and Y. Fu, “Chained-tracker: Chaining paired attentive regression results for end-to-end joint multiple-object detection and tracking,” in European Conference on Computer Vision , 2020, pp. 145–161
2020
Later among the works it cites.
2020
Later among the works it cites.
X. Zhou, V. Koltun, and P. Krähenbühl, “Tracking objects as points,” in European Conference on Computer Vision , 2020, pp. 474–490
2020
Later among the works it cites.
2020
Later among the works it cites.
Z. Li, Y. Li, Y. Liu, P. Wang, R. Lu, and H. B. Gooi, “Deep learning based densely connected network for load forecasting,” IEEE Transactions on Power Systems , 2020
2020
Later among the works it cites.
B. Pang, Y. Li, Y. Zhang, M. Li, and C. Lu, “Tubetk: Adopting tubes to track multi-object in a one-step training model,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 6308–6318
2020
Later among the works it cites.
Y. Zhang, H. Sheng, Y. Wu, S. Wang, W. Ke, and Z. Xiong, “Multiplex labeling graph for near-online tracking in crowded scenes,” IEEE Internet of Things Journal , vol. 7, no. 9, pp. 7892–7902, 2020
2020
Later among the works it cites.
2021
Closest in time.
2021
Closest in time.
J. Luiten, A. Osep, P. Dendorfer, P. Torr, A. Geiger, L. Leal-Taixé, and B. Leibe, “Hota: A higher order metric for evaluating multi-object tracking,” International Journal of Computer Vision , vol. 129, no. 2, pp. 548–578, 2021
2021
Closest in time.