Fetching the paper…
Reading the bibliography…
Transformers have proven superior performance for a wide variety of tasks since they were introduced.
The hungarian method for the assignment problem
Harold W. Kuhn and Bryn Yaw · 1955
Earlier work this paper cites.
Evaluating multiple object tracking performance: the clear mot metrics
Keni Bernardin and Rainer Stiefelhagen · 2008
Earlier work this paper cites.
A mobile vision system for robust multi-person tracking
A. Ess, B. Leibe, K. Schindler, , and L. van Gool · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng Jia, Dong Wei, Sochera Richard, Li Li-Jia, Li Kai, and Fei-Fei Li · 2009
Earlier work this paper cites.
Learning to associate: Hybridboosted multi-target tracker for crowded scene
Yuan Li, Chang Huang, and Ram Nevatia · 2009
Earlier work this paper cites.
Pedestrian detection: A benchmark
Dollár Piotr, Wojek Christian, Schiele Bernt, and Perona Pietro · 2009
Earlier work this paper cites.
Pedestrian detection: An evaluation of the state of the art
Piotr Dollár, Christian Wojek, Bernt Schiele, and Pietro Perona · 2011
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Deepreid: Deep filter pairing neural network for person re-identification
Wei Li, Rui Zhao, Tong Xiao, and Xiaogang Wang · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Joint probabilistic data association revisited
Seyed Hamid Rezatofighi, Anton Milan, Zhen Zhang, Qinfeng Shi, Anthony Dick, and Ian Reid · 2015
Earlier work this paper cites.
Subgraph decomposition for multi-target tracking
Siyu Tang, Bjoern Andres, Miykhaylo Andriluka, and Bernt Schiele · 2015
Earlier work this paper cites.
Learning to track: Online multi-object tracking by decision making
Yu Xiang, Alexandre Alahi, and Silvio Savarese · 2015
Earlier work this paper cites.
Scalable person re-identification: A benchmark
Liang Zheng, Liyue Shen, Lu Tian, Shengjin Wang, Jingdong Wang, and Qi Tian · 2015
Earlier work this paper cites.
Tracking multiple persons based on a variational bayesian model
Yutong Ban, Sileye Ba, Xavier Alameda-Pineda, and Radu Horaud · 2016
Earlier work this paper cites.
Simple online and realtime tracking
Alex Bewley, Zongyuan Ge, Lionel Ott, Fabio Ramos, and Ben Upcroft · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
A multi-cut formulation for joint segmentation and tracking of multiple objects
Margret Keuper, Siyu Tang, Yu Zhongjie, Bjoern Andres, Thomas Brox, and Bernt Schiele · 2016
Earlier work this paper cites.
MOT16: A benchmark for multi-object tracking
Anton Milan, Laura Leal-Taixé, Ian D. Reid, Stefan Roth, and Konrad Schindler · 2016
Earlier work this paper cites.
You only look once: Unified, real-time object detection
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi · 2016
Earlier work this paper cites.
Performance measures and a data set for multi-target, multi-camera tracking
Ergys Ristani, Francesco Solera, Roger Zou, Rita Cucchiara, and Carlo Tomasi · 2016
Earlier work this paper cites.
Multi-person tracking by multicut and deep matching
Siyu Tang, Bjoern Andres, Mykhaylo Andriluka, and Bernt Schiele · 2016
Earlier work this paper cites.
End-to-end deep learning for person search
Tong Xiao, Shuang Li, Bochao Wang, Liang Lin, and Xiaogang Wang · 2016
Earlier work this paper cites.
Deformable convolutional networks
Jifeng Dai, Haozhi Qi, Yuwen Xiong, Yi Li, Guodong Zhang, Han Hu, and Yichen Wei · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Online multi-target tracking using recurrent neural networks
Anton Milan, S Hamid Rezatofighi, Anthony Dick, Ian Reid, and Konrad Schindler · 2017
Earlier work this paper cites.
Tracking the untrackable: Learning to track multiple cues with long-term dependencies
Amir Sadeghian, Alexandre Alahi, and Silvio Savarese · 2017
Earlier work this paper cites.
Pathtrack: Fast trajectory annotation with path supervision
Manen Santiago, Gygli Michael, Dai Dengxin, and Van Gool Luc · 2017
Earlier work this paper cites.
Fast online tracking with detection refinement
Jianbing Shen, Dajiang Yu, Leyao Deng, and Xingping Dong · 2017
Earlier work this paper cites.
Multiple people tracking by lifted multicut and person re-identification
Siyu Tang, Mykhaylo Andriluka, Bjoern Andres, and Bernt Schiele · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Citypersons: A diverse dataset for pedestrian detection
Shanshan Zhang, Rodrigo Benenson, and Bernt Schiele · 2017
Earlier work this paper cites.
Person re-identification in the wild
Liang Zheng, Hengheng Zhang, Shaoyan Sun, Manmohan Chandraker, Yi Yang, and Qi Tian · 2017
Cited alongside, same era.
Real-time multiple people tracking with deeply learned candidate selection and person re-identification
Long Chen, Haizhou Ai, Zijie Zhuang, and Chong Shang · 2018
Cited alongside, same era.
Liteflownet: A lightweight convolutional neural network for optical flow estimation
Tak-Wai Hui, Xiaoou Tang, and Chen Change Loy · 2018
Cited alongside, same era.
Cornernet: Detecting objects as paired keypoints
Hei Law and Jia Deng · 2018
Cited alongside, same era.
Crowdhuman: A benchmark for detecting human in a crowd
Shuai Shao, Zijian Zhao, Boxun Li, Tete Xiao, Gang Yu, Xiangyu Zhang, and Jian Sun · 2018
Cited alongside, same era.
Gnn3dmot: Graph neural network for 3d multi-object tracking with 2d-3d multi-feature learning
Xinshuo Weng, Yongxin Wang, Yunze Man, and Kris M Kitani · 2020
Later among the works it cites.
Joint 3d tracking and forecasting with graph neural network and diversity sampling
Xinshuo Weng, Ye Yuan, and Kris Kitani · 2020
Later among the works it cites.
How to train your deep multi-object tracker
Yihong Xu, Aljosa Osep, Yutong Ban, Radu Horaud, Laura Leal-Taixé, and Xavier Alameda-Pineda · 2020
Later among the works it cites.
Learning texture transformer network for image super-resolution
Fuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu, and Baining Guo · 2020
Later among the works it cites.
A unified object motion and affinity model for online multi-object tracking
Junbo Yin, Wenguan Wang, Qinghao Meng, Ruigang Yang, and Jianbing Shen · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nathanael L Baisa · 2019
Cited alongside, same era.
Variational bayesian inference for audio-visual tracking of multiple speakers
Yutong Ban, Xavier Alameda-Pineda, Laurent Girin, and Radu Horaud · 2019
Cited alongside, same era.
Tracking without bells and whistles
Philipp Bergmann, Tim Meinhardt, and Laura Leal-Taixe · 2019
Cited alongside, same era.
Self-supervised moving vehicle tracking with stereo sound
Chuang Gan, Hang Zhao, Peihao Chen, David Cox, and Antonio Torralba · 2019
Cited alongside, same era.
Gnn3dmot: Graph neural network for 3d multi-object tracking with 2d-3d multi-feature learning
Hasith Karunasekera, Han Wang, and Handuo Zhang · 2019
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Ramprasaath R. Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra · 2019
Cited alongside, same era.
Bag of freebies for training object detection neural networks
Zhi Zhang, Tong He, Hang Zhang, Zhongyue Zhang, Junyuan Xie, and Mu Li · 2019
Cited alongside, same era.
Multiplex labeling graph for near-online tracking in crowded scenes
Yang Zhang, Hao Sheng, Yubin Wu, Shuai Wang, Wei Ke, and Zhang Xiong · 2020
Later among the works it cites.
Fairmot: On the fairness of detection and re-identification in multiple object tracking
Yifu Zhang, Chunyu Wang, Xinggang Wang, Wenjun Zeng, and Wenyu Liu · 2020
Later among the works it cites.
Tracking objects as points
Xingyi Zhou, Vladlen Koltun, and Philipp Krähenbühl · 2020
Later among the works it cites.
Deformable detr: Deformable transformers for end-to-end object detection
Xizhou Zhu, Weijie Su, Lewei Lu, Bin Li, Xiaogang Wang, and Jifeng Dai · 2020
Later among the works it cites.
Yolox: Exceeding yolo series in 2021
Zheng Ge, Songtao Liu, Feng Wang, Zeming Li, and Jian Sun · 2021
Closest in time.
Online multiple object tracking with cross-task synergy
Song Guo, Jingya Wang, Xinchao Wang, and Dacheng Tao · 2021
Closest in time.
Learnable graph matching: Incorporating graph partitioning with deep feature learning for multiple object tracking
Jiawei He, Zehao Huang, Naiyan Wang, and Zhaoxiang Zhang · 2021
Closest in time.
Transreid: Transformer-based object re-identification
Shuting He, Hao Luo, Pichao Wang, Fan Wang, Hao Li, and Wei Jiang · 2021
Closest in time.
Transgan: Two transformers can make one strong gan
Yifan Jiang, Shiyu Chang, and Zhangyang Wang · 2021
Closest in time.
Discriminative appearance modeling with multi-track pooling for real-time multi-object tracking
Chanho Kim, Fuxin Li, Mazen Alotaibi, and James M Rehg · 2021
Closest in time.
Trackformer: Multi-object tracking with transformers
Tim Meinhardt, Alexander Kirillov, Laura Leal-Taixe, and Christoph Feichtenhofer · 2021
Closest in time.
Quasi-dense similarity learning for multiple object tracking
Jiangmiao Pang, Linlu Qiu, Xia Li, Haofeng Chen, Qi Li, Trevor Darrell, and Fisher Yu · 2021
Closest in time.
Trackmpnn: A message passing graph neural architecture for multi-object tracking
Akshay Rangesh, Pranav Maheshwari, Mez Gebre, Siddhesh Mhatre, Vahid Ramezani, and Mohan M. Trivedi · 2021
Closest in time.
Probabilistic tracklet scoring and inpainting for multiple object tracking
Fatemeh Saleh, Sadegh Aliakbarian, Hamid Rezatofighi, Mathieu Salzmann, and Stephen Gould · 2021
Closest in time.
Siammot: Siamese multi-object tracking
Bing Shuai, Andrew Berneshawi, Xinyu Li, Davide Modolo, and Joseph Tighe · 2021
Closest in time.
Open synthetic dataset for improving cyclist detection
Phillip Thomas, Lars Pandikow, Alex Kim, Michael Stanley, and James Grieve · 2021
Closest in time.
Learning to track with object permanence
Pavel Tokmakov, Jie Li, Wolfram Burgard, and Adrien Gaidon · 2021
Closest in time.
Multiple object tracking with correlation learning
Qiang Wang, Yun Zheng, Pan Pan, and Yinghui Xu · 2021
Closest in time.
Joint object detection and multi-object tracking with graph neural networks
Yongxin Wang, Kris Kitani, and Xinshuo Weng · 2021
Closest in time.
Track to detect and segment: An online multi-object tracker
Jialian Wu, Jiale Cao, Liangchen Song, Yu Wang, Ming Yang, and Junsong Yuan · 2021
Closest in time.
Relationtrack: Relation-aware multiple object tracking with decoupled representation
En Yu, Zhuoling Li, Shoudong Han, and Hongwei Wang · 2021
Closest in time.
End-to-end multiple-object tracking with transformer
Fangao Zeng, Bin Dong, Tiancai Wang, Xiangyu Zhang, and Yichen Wei · 2021
Closest in time.
Bytetrack: Multi-object tracking by associating every detection box
Yifu Zhang, Peize Sun, Yi Jiang, Dongdong Yu, Zehuan Yuan, Ping Luo, Wenyu Liu, and Xinggang Wang · 2021
Closest in time.
Improving multiple object tracking with single object tracking
Linyu Zheng, Ming Tang, Yingying Chen, Guibo Zhu, Jinqiao Wang, and Hanqing Lu · 2021
Closest in time.
Looking beyond two frames: End-to-end multi-object tracking using spatial and temporal transformers
Tianyu Zhu, Markus Hiller, Mahsa Ehsanpour, Rongkai Ma, Tom Drummond, and Hamid Rezatofighi · 2021
Closest in time.
Unsupervised multiple-object tracking with a dynamical variational autoencoder
Xiaoyu Lin, Laurent Girin, and Xavier Alameda-Pineda · 2022
Closest in time.
Pvt v2: Improved baselines with pyramid vision transformer
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao · 2022
Closest in time.