Fetching the paper…
Reading the bibliography…
Over the past decade, significant progress has been made in visual object tracking, largely due to the availability of large-scale datasets.
A new approach to linear filtering and prediction problems
Kalman, R · 1960
Earlier work this paper cites.
Fully-convolutional siamese networks for object tracking
Bertinetto, L., Valmadre, J., Henriques, J. F. et al · 2016
Earlier work this paper cites.
It’s moving! a probabilistic model for causal motion segmentation in moving camera videos
Bideau, P. and Learned-Miller, E · 2016
Earlier work this paper cites.
Siamese cascaded region proposal networks for real-time visual tracking
Fan, H. and Ling, H · 2019
Earlier work this paper cites.
Lasot: A high-quality benchmark for large-scale single object tracking
Fan, H., Lin, L., Yang, F. et al · 2019
Earlier work this paper cites.
Distilling channels for efficient deep tracking
Ge, S., Luo, Z., Zhang, C. et al · 2019
Earlier work this paper cites.
Got-10k: A large high-diversity benchmark for generic object tracking in the wild
Huang, L., Zhao, X., and Huang, K · 2019
Earlier work this paper cites.
Siamrpn++: Evolution of siamese visual tracking with very deep networks
Li, B., Wu, W., Wang, Q. et al · 2019
Earlier work this paper cites.
Learning adaptive discriminative correlation filters via temporal consistency preserving spatial feature selection for robust visual object tracking
Xu, T., Feng, Z.-H., Wu, X.-J. et al · 2019
Earlier work this paper cites.
Robust deep tracking with two-step augmentation discriminative correlation filters
Zhang, C., Ge, S., Hua, Y. et al · 2019
Earlier work this paper cites.
Cascaded correlation refinement for robust deep tracking
Ge, S., Zhang, C., Li, S. et al · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A. et al · 2020
Earlier work this paper cites.
Urvos: Unified referring video object segmentation network with a large-scale benchmark
Seo, S., Lee, J.-Y., and Han, B · 2020
Earlier work this paper cites.
Accurate uav tracking with distance-injected overlap maximization
Zhang, C., Ge, S., Zhang, K. et al · 2020
Earlier work this paper cites.
Lasot: A high-quality large-scale single object tracking benchmark
Fan, H., Bai, H., Lin, L. et al · 2021
Cited alongside, same era.
Visual tracking of deepwater animals using machine learning-controlled robotic underwater vehicles
Katija, K., Roberts, P. L., Daniels, J. et al · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C. et al · 2021
Cited alongside, same era.
Transformer meets tracker: Exploiting temporal context for robust visual tracking
Wang, N., Zhou, W., Wang, J. et al · 2021
Cited alongside, same era.
Utb180: A high-quality benchmark for underwater tracking
Alawode, B., Guo, Y., Ummar, M. et al · 2022
Cited alongside, same era.
Implicit motion handling for video camouflaged object detection
Cheng, X., Xiong, H., Fan, D.-P. et al · 2022
Watb: wild animal tracking benchmark
Wang, F., Cao, P., Li, F. et al · 2023
Later among the works it cites.
Autoregressive visual tracking
Wei, X., Bai, Y., Zheng, Y. et al · 2023
Later among the works it cites.
Single-model and any-modality for video object tracking
Wu, Z., Zheng, J., Ren, X. et al · 2023
Later among the works it cites.
Track anything: Segment anything meets videos
Yang, J., Gao, M., Li, Z. et al · 2023
Later among the works it cites.
Camouflaged object tracking: A benchmark
Guo, X., Zhong, P., Zhang, H. et al · 2024
Closest in time.
Mamba-fetrack: Frame-event tracking via state space model
Huang, J., Wang, S., Wang, S. et al · 2024
Closest in time.
Mambavt: Spatio-temporal contextual modeling for robust rgb-t tracking
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Divert more attention to vision-language tracking
Guo, M., Zhang, Z., Fan, H. et al · 2022
Cited alongside, same era.
Joint feature learning and relation modeling for tracking: A one-stream framework
Ye, B., Chang, H., Ma, B. et al · 2022
Cited alongside, same era.
Improving underwater visual tracking with a large scale dataset and image enhancement
Alawode, B., Dharejo, F. A., Ummar, M. et al · 2023
Cited alongside, same era.
Mixformerv2: Efficient fully transformer tracking
Cui, Y., Song, T., Wu, G. et al · 2023
Cited alongside, same era.
Sam-da: Uav tracks anything at night with sam-powered domain adaptation
Fu, C., Yao, L., Zuo, H. et al · 2023
Cited alongside, same era.
Citetracker: Correlating image and text for visual tracking
Li, X., Huang, Y., He, Z. et al · 2023
Cited alongside, same era.
Lai, S., Liu, C., Zhu, J. et al · 2024
Closest in time.
Tracking meets lora: Faster training, larger model, stronger performance
Lin, L., Fan, H., Zhang, Z. et al · 2024
Closest in time.
Vasttrack: Vast category visual object tracking
Peng, L., Gao, J., Liu, X. et al · 2024
Closest in time.
Sam 2: Segment anything in images and videos
Ravi, N., Gabeur, V., Hu, Y.-T. et al · 2024
Closest in time.
Samurai: Adapting segment anything model for zero-shot visual tracking with motion-aware memory
Yang, C.-Y., Huang, H.-W., Chai, W. et al · 2024
Closest in time.
ISAT with Segment Anything: An Interactive Semi-Automatic Annotation Tool, 2024
Ji, S. and Zhang, H · 2025
Closest in time.
A distractor-aware memory for visual object tracking with sam2
Videnovic, J., Lukezic, A., and Kristan, M · 2025
Closest in time.