Fetching the paper…
Reading the bibliography…
Transformer-based trackers have achieved strong accuracy on the standard benchmarks.
Compressing deep convolutional networks using vector quantization
Yunchao Gong, Liu Liu, Ming Yang, and Lubomir Bourdev · 2014
Earlier work this paper cites.
Microsoft COCO: common objects in context
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
Fitnets: Hints for thin deep nets
Romero Adriana, Ballas Nicolas, K Samira Ebrahimi, Chassang Antoine, Gatta Carlo, and B Yoshua · 2015
Earlier work this paper cites.
High-speed tracking with kernelized correlation filters
João F. Henriques, Rui Caseiro, Pedro Martins, and Jorge Batista · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network (2015)
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Earlier work this paper cites.
Fully-convolutional siamese networks for object tracking
Luca Bertinetto, Jack Valmadre, Joao F Henriques, Andrea Vedaldi, and Philip HS Torr · 2016
Earlier work this paper cites.
A benchmark and simulator for UAV tracking
Matthias Mueller, Neil Smith, and Bernard Ghanem · 2016
Earlier work this paper cites.
Channel pruning for accelerating very deep neural networks
Yihui He, Xiangyu Zhang, and Jian Sun · 2017
Earlier work this paper cites.
Mimicking very efficient network for object detection
Quanquan Li, Shengying Jin, and Junjie Yan · 2017
Earlier work this paper cites.
High performance visual tracking with siamese region proposal network
Bo Li, Junjie Yan, Wei Wu, Zheng Zhu, and Xiaolin Hu · 2018
Earlier work this paper cites.
Trackingnet: A large-scale dataset and benchmark for object tracking in the wild
Matthias Müller, Adel Bibi, Silvio Giancola, Salman Al-Subaihi, and Bernard Ghanem · 2018
Earlier work this paper cites.
Learning discriminative model prediction for tracking
Goutam Bhat, Martin Danelljan, Luc Van Gool, and Radu Timofte · 2019
Earlier work this paper cites.
ATOM: accurate tracking by overlap maximization
Martin Danelljan, Goutam Bhat, Fahad Shahbaz Khan, and Michael Felsberg · 2019
Earlier work this paper cites.
Neural architecture search: A survey
Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter · 2019
Earlier work this paper cites.
Lasot: A high-quality benchmark for large-scale single object tracking
Heng Fan, Liting Lin, Fan Yang, Peng Chu, Ge Deng, Sijia Yu, Hexin Bai, Yong Xu, Chunyuan Liao, and Haibin Ling · 2019
Earlier work this paper cites.
Siamrpn++: Evolution of siamese visual tracking with very deep networks
Bo Li, Wei Wu, Qiang Wang, Fangyi Zhang, Junliang Xing, and Junjie Yan · 2019
Earlier work this paper cites.
Haq: Hardware-aware automated quantization with mixed precision
Kuan Wang, Zhijian Liu, Yujun Lin, Ji Lin, and Song Han · 2019
Earlier work this paper cites.
Siamese box adaptive network for visual tracking
Zedu Chen, Bineng Zhong, Guorong Li, Shengping Zhang, and Rongrong Ji · 2020
Earlier work this paper cites.
Probabilistic regression for visual tracking
Martin Danelljan, Luc Van Gool, and Radu Timofte · 2020
Earlier work this paper cites.
The eighth visual object tracking VOT2020 challenge results
Matej Kristan, Ales Leonardis, and et. al · 2020
Cited alongside, same era.
Metadistiller: Network self-boosting via meta-learned top-down distillation
Benlin Liu, Yongming Rao, Jiwen Lu, Jie Zhou, and Cho-Jui Hsieh · 2020
Cited alongside, same era.
D3S - A discriminative single shot segmentation tracker
Alan Lukezic, Jiri Matas, and Matej Kristan · 2020
Cited alongside, same era.
Siamfc++: Towards robust and accurate visual tracking with target estimation guidelines
Yinda Xu, Zeyu Wang, Zuoxin Li, Ye Yuan, and Gang Yu · 2020
Cited alongside, same era.
Efficient visual tracking with exemplar transformers
Philippe Blatter, Menelaos Kanakis, Martin Danelljan, and Luc Van Gool · 2021
Cited alongside, same era.
Lighttrack: Finding lightweight neural networks for object tracking via one-shot architecture search
Bin Yan, Houwen Peng, Kan Wu, Dong Wang, Jianlong Fu, and Huchuan Lu · 2021
Later among the works it cites.
Vision transformer slimming: Multi-dimension searching in continuous optimization space
Arnav Chavan, Zhiqiang Shen, Zhuang Liu, Zechun Liu, Kwang-Ting Cheng, and Eric P Xing · 2022
Later among the works it cites.
Backbone is all your need: A simplified architecture for visual object tracking
Boyu Chen, Peixia Li, Lei Bai, Lei Qiao, Qiuhong Shen, Bo Li, Weihao Gan, Wei Wu, and Wanli Ouyang · 2022
Later among the works it cites.
Efficient visual tracking via hierarchical cross-attention transformer
Xin Chen, Dong Wang, Dongdong Li, and Huchuan Lu · 2022
Later among the works it cites.
Fully convolutional online tracking
Yutao Cui, Cheng Jiang, Limin Wang, and Gangshan Wu · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vasyl Borsuk, Roman Vei, Orest Kupyn, Tetiana Martyniuk, Igor Krashenyi, and Jiři Matas · 2021
Cited alongside, same era.
Autoformer: Searching transformers for visual recognition
Minghao Chen, Houwen Peng, Jianlong Fu, and Haibin Ling · 2021
Cited alongside, same era.
Transformer tracking
Xin Chen, Bin Yan, Jiawen Zhu, Dong Wang, Xiaoyun Yang, and Huchuan Lu · 2021
Cited alongside, same era.
Target transformed regression for accurate tracking
Yutao Cui, Cheng Jiang, Limin Wang, and Gangshan Wu · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Cited alongside, same era.
Nasvit: Neural architecture search for efficient vision transformers with gradient conflict aware supernet training
Chengyue Gong, Dilin Wang, Meng Li, Xinlei Chen, Zhicheng Yan, Yuandong Tian, Vikas Chandra, et al · 2021
Cited alongside, same era.
Distilling object detectors via decoupled features
Jianyuan Guo, Kai Han, Yunhe Wang, Han Wu, Xinghao Chen, Chunjing Xu, and Chang Xu · 2021
Cited alongside, same era.
Mixformer: End-to-end tracking with iterative mixed attention
Yutao Cui, Cheng Jiang, Limin Wang, and Gangshan Wu · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Later among the works it cites.
Swintrack: A simple and strong baseline for transformer tracking
Liting Lin, Heng Fan, Yong Xu, and Haibin Ling · 2022
Later among the works it cites.
Transforming model prediction for tracking
Christoph Mayer, Martin Danelljan, Goutam Bhat, Matthieu Paul, Danda Pani Paudel, Fisher Yu, and Luc Van Gool · 2022
Later among the works it cites.
Transformer tracking with cyclic shifting window attention
Zikai Song, Junqing Yu, Yi-Ping Phoebe Chen, and Wei Yang · 2022
Later among the works it cites.
Correlation-aware deep tracking
Fei Xie, Chunyu Wang, Guangting Wang, Yue Cao, Wankou Yang, and Wenjun Zeng · 2022
Later among the works it cites.
Evo-vit: Slow-fast token evolution for dynamic vision transformer
Yifan Xu, Zhijie Zhang, Mengdan Zhang, Kekai Sheng, Ke Li, Weiming Dong, Liqing Zhang, Changsheng Xu, and Xing Sun · 2022
Later among the works it cites.
Vitkd: Practical guidelines for vit feature knowledge distillation
Zhendong Yang, Zhe Li, Ailing Zeng, Zexian Li, Chun Yuan, and Yu Li · 2022
Later among the works it cites.
Joint feature learning and relation modeling for tracking: A one-stream framework
Botao Ye, Hong Chang, Bingpeng Ma, and Shiguang Shan · 2022
Later among the works it cites.
Minivit: Compressing vision transformers with weight multiplexing
Jinnian Zhang, Houwen Peng, Kan Wu, Mengchen Liu, Bin Xiao, Jianlong Fu, and Lu Yuan · 2022
Later among the works it cites.
Localization distillation for object detection
Zhaohui Zheng, Rongguang Ye, Qibin Hou, Dongwei Ren, Ping Wang, Wangmeng Zuo, and Ming-Ming Cheng · 2022
Later among the works it cites.
The tenth visual object tracking vot2022 challenge results
Matej Kristan, Aleš Leonardis, Jiří Matas, Michael Felsberg, Roman Pflugfelder, Joni-Kristian Kämäräinen, Hyung Jin Chang, Martin Danelljan, Luka Čehovin Zajc, Alan Lukežič, et al · 2023
Closest in time.
Compact transformer tracker with correlative masked modeling
Zikai Song, Run Luo, Junqing Yu, Yi-Ping Phoebe Chen, and Wei Yang · 2023
Closest in time.
Mixformer: End-to-end tracking with iterative mixed attention
Yutao Cui, Cheng Jiang, Gangshan Wu, and Limin Wang · 2024
Closest in time.