Fetching the paper…
Reading the bibliography…
Transformer framework has been showing superior performances in visual object tracking for its great strength in information aggregation across the template and search image with the well-known attention mechanism.
Deep Attentive Tracking via Reciprocative Learning
Pu, S.; Song, Y.; Ma, C.; Zhang, H.; and Yang, M.-H. 2018 · 1941
Earlier work this paper cites.
Histograms of oriented gradients for human detection
Dalal, N.; and Triggs, B. 2005 · 2005
Earlier work this paper cites.
Visual object tracking using adaptive correlation filters
Bolme, D. S.; Beveridge, J. R.; Draper, B. A.; and Lui, Y. M. 2010 · 2010
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A.; Sutskever, I.; and Hinton, G. E. 2012 · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y.; Maire, M.; Belongie, S.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
High-Speed Tracking with Kernelized Correlation Filters
Henriques, J. F.; Caseiro, R.; Martins, P.; and Batista, J. 2015 · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Simonyan, K.; and Zisserman, A. 2015 · 2015
Earlier work this paper cites.
Fully-Convolutional Siamese Networks for Object Tracking
Bertinetto, L.; Valmadre, J.; Henriques, J. F.; Vedaldi, A.; and Torr, P. H. S. 2016 · 2016
Earlier work this paper cites.
Beyond Correlation Filters: Learning Continuous Convolution Operators for Visual Tracking
Danelljan, M.; Robinson, A.; Khan, F. S.; and Felsberg, M. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
A benchmark and simulator for uav tracking
Mueller, M.; Smith, N.; and Ghanem, B. 2016 · 2016
Earlier work this paper cites.
Learning Multi-Domain Convolutional Neural Networks for Visual Tracking
Nam, H.; and Han, B. 2016 · 2016
Earlier work this paper cites.
ECO: Efficient Convolution Operators for Tracking
Danelljan, M.; Bhat, G.; Shahbaz Khan, F.; and Felsberg, M. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Acquisition of localization confidence for accurate object detection
Jiang, B.; Luo, R.; Mao, J.; Xiao, T.; and Jiang, Y. 2018 · 2018
Earlier work this paper cites.
High Performance Visual Tracking With Siamese Region Proposal Network
Li, B.; Yan, J.; Wu, W.; Zhu, Z.; and Hu, X. 2018 · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Loshchilov, I.; and Hutter, F. 2018 · 2018
Earlier work this paper cites.
TrackingNet: A Large-Scale Dataset and Benchmark for Object Tracking in the Wild
Muller, M.; Bibi, A.; Giancola, S.; Alsubaihi, S.; and Ghanem, B. 2018 · 2018
Cited alongside, same era.
Learning Discriminative Model Prediction for Tracking
Bhat, G.; Danelljan, M.; Gool, L. V.; and Timofte, R. 2019 · 2019
Cited alongside, same era.
ATOM: Accurate Tracking by Overlap Maximization
Danelljan, M.; Bhat, G.; Khan, F. S.; and Felsberg, M. 2019 · 2019
Cited alongside, same era.
LaSOT: A High-Quality Benchmark for Large-Scale Single Object Tracking
Fan, H.; Lin, L.; Yang, F.; Chu, P.; Deng, G.; Yu, S.; Bai, H.; Xu, Y.; Liao, C.; and Ling, H. 2019 · 2019
Cited alongside, same era.
GOT-10k: A Large High-Diversity Benchmark for Generic Object Tracking in the Wild
Huang, L.; Zhao, X.; and Huang, K. 2019 · 2019
Cited alongside, same era.
SiamRPN++: Evolution of Siamese Visual Tracking With Very Deep Networks
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2021
Later among the works it cites.
Stmtrack: Template-free visual tracking with space-time memory networks
Fu, Z.; Liu, Q.; Fu, Z.; and Wang, Y. 2021 · 2021
Later among the works it cites.
Graph attention tracking
Guo, D.; Shao, Y.; Cui, Y.; Wang, Z.; Zhang, L.; and Shen, C. 2021 · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Liu, Z.; Lin, Y.; Cao, Y.; Hu, H.; Wei, Y.; Zhang, Z.; Lin, S.; and Guo, B. 2021 · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Ramesh, A.; Pavlov, M.; Goh, G.; Gray, S.; Voss, C.; Radford, A.; Chen, M.; and Sutskever, I. 2021 · 2021
Later among the works it cites.
Transformer Meets Tracker: Exploiting Temporal Context for Robust Visual Tracking
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Li, B.; Wu, W.; Wang, Q.; Zhang, F.; Xing, J.; and Yan, J. 2019 · 2019
Cited alongside, same era.
Generalized intersection over union: A metric and a loss for bounding box regression
Rezatofighi, H.; Tsoi, N.; Gwak, J.; Sadeghian, A.; Reid, I.; and Savarese, S. 2019 · 2019
Cited alongside, same era.
Know Your Surroundings: Exploiting Scene Information for Object Tracking
Bhat, G.; Danelljan, M.; Van Gool, L.; and Timofte, R. 2020 · 2020
Cited alongside, same era.
End-to-end object detection with transformers
Carion, N.; Massa, F.; Synnaeve, G.; Usunier, N.; Kirillov, A.; and Zagoruyko, S. 2020 · 2020
Cited alongside, same era.
Generative pretraining from pixels
Chen, M.; Radford, A.; Child, R.; Wu, J.; Jun, H.; Luan, D.; and Sutskever, I. 2020 · 2020
Cited alongside, same era.
The eighth visual object tracking VOT2020 challenge results
Kristan, M.; Leonardis, A.; Matas, J.; Felsberg, M.; Pflugfelder, R.; Kämäräinen, J.-K.; Danelljan, M.; Zajc, L. Č.; Lukežič, A.; Drbohlav, O.; et al. 2020 · 2020
Cited alongside, same era.
Siam r-cnn: Visual tracking by re-detection
Voigtlaender, P.; Luiten, J.; Torr, P. H.; and Leibe, B. 2020 · 2020
Cited alongside, same era.
Wang, N.; Zhou, W.; Wang, J.; and Li, H. 2021 · 2021
Later among the works it cites.
Cvt: Introducing convolutions to vision transformers
Wu, H.; Xiao, B.; Codella, N.; Liu, M.; Dai, X.; Yuan, L.; and Zhang, L. 2021 · 2021
Later among the works it cites.
Learning spatio-temporal transformer for visual tracking
Yan, B.; Peng, H.; Fu, J.; Wang, D.; and Lu, H. 2021 · 2021
Later among the works it cites.
Learn to match: Automatic matching network design for visual tracking
Zhang, Z.; Liu, Y.; Wang, X.; Li, B.; and Hu, W. 2021 · 2021
Later among the works it cites.
MixFormer: End-to-End Tracking With Iterative Mixed Attention
Cui, Y.; Jiang, C.; Wang, L.; and Wu, G. 2022 · 2022
Later among the works it cites.
Masked Autoencoders Are Scalable Vision Learners
He, K.; Chen, X.; Xie, S.; Li, Y.; Dollár, P.; and Girshick, R. 2022 · 2022
Later among the works it cites.
Exploring plain vision transformer backbones for object detection
Li, Y.; Mao, H.; Girshick, R.; and He, K. 2022 · 2022
Later among the works it cites.
Unsupervised Learning of Accurate Siamese Tracking
Shen, Q.; Qiao, L.; Guo, J.; Li, P.; Li, X.; Li, B.; Feng, W.; Gan, W.; Wu, W.; and Ouyang, W. 2022 · 2022
Later among the works it cites.
Transformer Tracking With Cyclic Shifting Window Attention
Song, Z.; Yu, J.; Chen, Y.-P. P.; and Yang, W. 2022 · 2022
Later among the works it cites.
Masked feature prediction for self-supervised visual pre-training
Wei, C.; Fan, H.; Xie, S.; Wu, C.-Y.; Yuille, A.; and Feichtenhofer, C. 2022 · 2022
Later among the works it cites.
SimMIM: A Simple Framework for Masked Image Modeling
Xie, Z.; Zhang, Z.; Cao, Y.; Lin, Y.; Bao, J.; Yao, Z.; Dai, Q.; and Hu, H. 2022 · 2022
Later among the works it cites.