Fetching the paper…
Reading the bibliography…
The Segment Anything Model 2 (SAM 2) has emerged as a powerful foundation model for object segmentation in both images and videos, paving the way for various downstream video applications.
An algorithm for tracking multiple targets
Donald Reid · 1979
Earlier work this paper cites.
An efficient implementation of reid’s multiple hypothesis tracking algorithm and its evaluation for the purpose of visual tracking
Ingemar J. Cox and Sunita L. Hingorani · 1996
Earlier work this paper cites.
Object segmentation by long term analysis of point trajectories
Thomas Brox and Jitendra Malik · 2010
Earlier work this paper cites.
Multiple hypothesis tracking revisited
Chanho Kim, Fuxin Li, Arridhana Ciptadi, and James M Rehg · 2015
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Earlier work this paper cites.
Learning video object segmentation from static images
Federico Perazzi, Anna Khoreva, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Earlier work this paper cites.
The 2017 davis challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alex Sorkine-Hornung, and Luc Van Gool · 2017
Earlier work this paper cites.
Online adaptation of convolutional neural networks for video object segmentation
Paul Voigtlaender and Bastian Leibe · 2017
Earlier work this paper cites.
Cnn in mrf: Video object segmentation via inference in a cnn-based higher-order spatio-temporal mrf
Linchao Bao, Baoyuan Wu, and Wei Liu · 2018
Earlier work this paper cites.
Blazingly fast video object segmentation with pixel-wise metric learning
Yuhua Chen, Jordi Pont-Tuset, Alberto Montes, and Luc Van Gool · 2018
Earlier work this paper cites.
Video object segmentation with joint re-identification and attention-aware mask propagation
Xiaoxiao Li and Chen Change Loy · 2018
Earlier work this paper cites.
Video object segmentation without temporal information
K-K Maninis, Sergi Caelles, Yuhua Chen, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2018
Earlier work this paper cites.
Fast video object segmentation by reference-guided mask propagation
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Earlier work this paper cites.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas Huang · 2018
Earlier work this paper cites.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos K Katsaggelos · 2018
Earlier work this paper cites.
Got-10k: A large high-diversity benchmark for generic object tracking in the wild
Lianghua Huang, Xin Zhao, and Kaiqi Huang · 2019
Earlier work this paper cites.
A generative appearance model for end-to-end video object segmentation
Joakim Johnander, Martin Danelljan, Emil Brissman, Fahad Shahbaz Khan, and Michael Felsberg · 2019
Earlier work this paper cites.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Earlier work this paper cites.
Rvos: End-to-end recurrent network for video object segmentation
Carles Ventura, Miriam Bellver, Andreu Girbau, Amaia Salvador, Ferran Marques, and Xavier Giro-i Nieto · 2019
Earlier work this paper cites.
Feelvos: Fast end-to-end embedding learning for video object segmentation
Paul Voigtlaender, Yuning Chai, Florian Schroff, Hartwig Adam, Bastian Leibe, and Liang-Chieh Chen · 2019
Earlier work this paper cites.
Fast online object tracking and segmentation: A unifying approach
Qiang Wang, Li Zhang, Luca Bertinetto, Weiming Hu, and Philip HS Torr · 2019
Cited alongside, same era.
Fast video object segmentation via dynamic targeting network
Lu Zhang, Zhe Lin, Jianming Zhang, Huchuan Lu, and You He · 2019
Cited alongside, same era.
Learning what to learn for video object segmentation
Goutam Bhat, Felix Järemo Lawin, Martin Danelljan, Andreas Robinson, Michael Felsberg, Luc Van Gool, and Radu Timofte · 2020
Cited alongside, same era.
Fast video object segmentation using the global context module
Yu Li, Zhuoran Shen, and Ying Shan · 2020
Cited alongside, same era.
Video object segmentation with adaptive feature bank and uncertain-region refinement
Yongqing Liang, Xin Li, Navid Jafari, and Jim Chen · 2020
Cited alongside, same era.
Learning fast and robust target models for video object segmentation
Tarvis: A unified approach for target-based video segmentation
Ali Athar, Alexander Hermans, Jonathon Luiten, Deva Ramanan, and Bastian Leibe · 2023
Later among the works it cites.
Xmem++: Production-level video segmentation from few annotated frames
Maksym Bekuzarov, Ariana Bermudez, Joon-Young Lee, and Hao Li · 2023
Later among the works it cites.
Robust object modeling for visual tracking
Yidong Cai, Jie Liu, Jie Tang, and Gangshan Wu · 2023
Later among the works it cites.
Seqtrack: Sequence to sequence learning for visual object tracking
Xin Chen, Houwen Peng, Dong Wang, Huchuan Lu, and Han Hu · 2023
Later among the works it cites.
Tracking anything with decoupled video segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian Price, Alexander Schwing, and Joon-Young Lee · 2023
Later among the works it cites.
Generalized relation modeling for transformer tracking
Shenyuan Gao, Chunluan Zhou, and Jun Zhang · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Andreas Robinson, Felix Jaremo Lawin, Martin Danelljan, Fahad Shahbaz Khan, and Michael Felsberg · 2020
Cited alongside, same era.
Kernelized memory network for video object segmentation
Hongje Seong, Junhyuk Hyun, and Euntai Kim · 2020
Cited alongside, same era.
Collaborative video object segmentation by foreground-background integration
Zongxin Yang, Yunchao Wei, and Yi Yang · 2020
Cited alongside, same era.
Rethinking space-time networks with improved memory coverage for efficient video object segmentation
Ho Kei Cheng, Yu-Wing Tai, and Chi-Keung Tang · 2021
Cited alongside, same era.
Sstvos: Sparse spatiotemporal transformers for video object segmentation
Brendan Duke, Abdalla Ahmed, Christian Wolf, Parham Aarabi, and Graham W Taylor · 2021
Cited alongside, same era.
Lasot: A high-quality large-scale single object tracking benchmark
Heng Fan, Hexin Bai, Liting Lin, Fan Yang, Peng Chu, Ge Deng, Sijia Yu, Harshit, Mingzhen Huang, Juehuan Liu, et al · 2021
Cited alongside, same era.
Learning target candidate association to keep track of what not to track
Christoph Mayer, Martin Danelljan, Danda Pani Paudel, and Luc Van Gool · 2021
Cited alongside, same era.
Later among the works it cites.
Lvos: A benchmark for long-term video object segmentation
Lingyi Hong, Wenchao Chen, Zhongying Liu, Wei Zhang, Pinxue Guo, Zhaoyu Chen, and Wenqiang Zhang · 2023
Later among the works it cites.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Semantics meets temporal correspondence: Self-supervised object-centric learning in videos
Rui Qian, Shuangrui Ding, Xian Liu, and Dahua Lin · 2023
Later among the works it cites.
Breaking the “object” in video object segmentation
Pavel Tokmakov, Jie Li, and Adrien Gaidon · 2023
Later among the works it cites.
Look before you match: Instance understanding matters in video object segmentation
Junke Wang, Dongdong Chen, Zuxuan Wu, Chong Luo, Chuanxin Tang, Xiyang Dai, Yucheng Zhao, Yujia Xie, Lu Yuan, and Yu-Gang Jiang · 2023
Later among the works it cites.
Joint modeling of feature, correspondence, and a compressed memory for video object segmentation
Jiaming Zhang, Yutao Cui, Gangshan Wu, and Limin Wang · 2023
Later among the works it cites.
Putting the object back into video object segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian Price, Joon-Young Lee, and Alexander Schwing · 2024
Closest in time.
Lvos: A benchmark for large-scale long-term video object segmentation
Lingyi Hong, Zhongying Liu, Wenchao Chen, Chenzhi Tan, Yuang Feng, Xinyu Zhou, Pinxue Guo, Jinglun Li, Zhaoyu Chen, Shuyong Gao, et al · 2024
Closest in time.
Diffusiontrack: Diffusion model for multi-object tracking
Run Luo, Zikai Song, Lintao Ma, Jinlin Wei, Wei Yang, and Min Yang · 2024
Closest in time.
Rethinking image-to-video adaptation: An object-centric perspective
Rui Qian, Shuangrui Ding, and Dahua Lin · 2024
Closest in time.
Sam 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, et al · 2024
Closest in time.
Odtrack: Online dense temporal token learning for visual tracking
Yaozong Zheng, Bineng Zhong, Qihua Liang, Zhiyi Mo, Shengping Zhang, and Xianxian Li · 2024
Closest in time.
Rmem: Restricted memory banks improve video object segmentation
Junbao Zhou, Ziqi Pang, and Yu-Xiong Wang · 2024
Closest in time.
Keyframe-guided creative video inpainting
Yuwei Guo, Ceyuan Yang, Anyi Rao, Chenlin Meng, Omer Bar-Tal, Shuangrui Ding, Maneesh Agrawala, Dahua Lin, and Bo Dai · 2025
Closest in time.
Tracking meets lora: Faster training, larger model, stronger performance
Liting Lin, Heng Fan, Zhipeng Zhang, Yaowei Wang, Yong Xu, and Haibin Ling · 2025
Closest in time.