Fetching the paper…
Reading the bibliography…
Current prevailing Video Object Segmentation methods follow the pipeline of extraction-then-matching, which first extracts features on current and reference frames independently, and then performs dense matching between them.
Video segmentation and its applications
King Ngi Ngan and Hongliang Li · 2011
Earlier work this paper cites.
Video segmentation and its applications
King Ngi Ngan and Hongliang Li · 2011
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Training deep neural networks on noisy labels with bootstrapping
Scott E. Reed, Honglak Lee, Dragomir Anguelov, Christian Szegedy, Dumitru Erhan, and Andrew Rabinovich · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Training deep neural networks on noisy labels with bootstrapping
Scott E. Reed, Honglak Lee, Dragomir Anguelov, Christian Szegedy, Dumitru Erhan, and Andrew Rabinovich · 2015
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus H. Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
Hierarchical image saliency detection on extended CSSD
Jianping Shi, Qiong Yan, Li Xu, and Jiaya Jia · 2016
Earlier work this paper cites.
Video segmentation via object flow
Yi-Hsuan Tsai, Ming-Hsuan Yang, and Michael J. Black · 2016
Earlier work this paper cites.
Instance-level segmentation for autonomous driving with deep densely connected mrfs
Ziyu Zhang, Sanja Fidler, and Raquel Urtasun · 2016
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus H. Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
Hierarchical image saliency detection on extended CSSD
Jianping Shi, Qiong Yan, Li Xu, and Jiaya Jia · 2016
Earlier work this paper cites.
Video segmentation via object flow
Yi-Hsuan Tsai, Ming-Hsuan Yang, and Michael J. Black · 2016
Earlier work this paper cites.
Instance-level segmentation for autonomous driving with deep densely connected mrfs
Ziyu Zhang, Sanja Fidler, and Raquel Urtasun · 2016
Earlier work this paper cites.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Earlier work this paper cites.
Segflow: Joint learning for video object segmentation and optical flow
Jingchun Cheng, Yi-Hsuan Tsai, Shengjin Wang, and Ming-Hsuan Yang · 2017
Earlier work this paper cites.
Lucid data dreaming for object tracking
Anna Khoreva, Rodrigo Benenson, Eddy Ilg, Thomas Brox, and Bernt Schiele · 2017
Earlier work this paper cites.
Learning video object segmentation from static images
Federico Perazzi, Anna Khoreva, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Earlier work this paper cites.
The 2017 DAVIS challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbelaez, Alexander Sorkine-Hornung, and Luc Van Gool · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Online adaptation of convolutional neural networks for video object segmentation
Paul Voigtlaender and Bastian Leibe · 2017
Earlier work this paper cites.
Learning to detect salient objects with image-level supervision
Lijun Wang, Huchuan Lu, Yifan Wang, Mengyang Feng, Dong Wang, Baocai Yin, and Xiang Ruan · 2017
Earlier work this paper cites.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Earlier work this paper cites.
Segflow: Joint learning for video object segmentation and optical flow
Jingchun Cheng, Yi-Hsuan Tsai, Shengjin Wang, and Ming-Hsuan Yang · 2017
Earlier work this paper cites.
Lucid data dreaming for object tracking
Anna Khoreva, Rodrigo Benenson, Eddy Ilg, Thomas Brox, and Bernt Schiele · 2017
Earlier work this paper cites.
Learning video object segmentation from static images
Federico Perazzi, Anna Khoreva, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Earlier work this paper cites.
The 2017 DAVIS challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbelaez, Alexander Sorkine-Hornung, and Luc Van Gool · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Online adaptation of convolutional neural networks for video object segmentation
Paul Voigtlaender and Bastian Leibe · 2017
Earlier work this paper cites.
Learning to detect salient objects with image-level supervision
Lijun Wang, Huchuan Lu, Yifan Wang, Mengyang Feng, Dong Wang, Baocai Yin, and Xiang Ruan · 2017
Earlier work this paper cites.
Blazingly fast video object segmentation with pixel-wise metric learning
Yuhua Chen, Jordi Pont-Tuset, Alberto Montes, and Luc Van Gool · 2018
Earlier work this paper cites.
Videomatch: Matching based video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G. Schwing · 2018
Earlier work this paper cites.
Video object segmentation with joint re-identification and attention-aware mask propagation
Xiaoxiao Li and Chen Change Loy · 2018
Earlier work this paper cites.
Premvos: Proposal-generation, refinement and merging for video object segmentation
Jonathon Luiten, Paul Voigtlaender, and Bastian Leibe · 2018
Earlier work this paper cites.
Fast video object segmentation by reference-guided mask propagation
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Earlier work this paper cites.
Monet: Deep motion exploitation for video object segmentation
Huaxin Xiao, Jiashi Feng, Guosheng Lin, Yu Liu, and Maojun Zhang · 2018
Earlier work this paper cites.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas S. Huang · 2018
Earlier work this paper cites.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos K. Katsaggelos · 2018
Earlier work this paper cites.
Blazingly fast video object segmentation with pixel-wise metric learning
Yuhua Chen, Jordi Pont-Tuset, Alberto Montes, and Luc Van Gool · 2018
Earlier work this paper cites.
Videomatch: Matching based video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G. Schwing · 2018
Earlier work this paper cites.
Video object segmentation with joint re-identification and attention-aware mask propagation
Xiaoxiao Li and Chen Change Loy · 2018
Earlier work this paper cites.
Premvos: Proposal-generation, refinement and merging for video object segmentation
Jonathon Luiten, Paul Voigtlaender, and Bastian Leibe · 2018
Earlier work this paper cites.
Fast video object segmentation by reference-guided mask propagation
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Earlier work this paper cites.
Monet: Deep motion exploitation for video object segmentation
Huaxin Xiao, Jiashi Feng, Guosheng Lin, Yu Liu, and Maojun Zhang · 2018
Earlier work this paper cites.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas S. Huang · 2018
Earlier work this paper cites.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos K. Katsaggelos · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Earlier work this paper cites.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Earlier work this paper cites.
FEELVOS: fast end-to-end embedding learning for video object segmentation
Paul Voigtlaender, Yuning Chai, Florian Schroff, Hartwig Adam, Bastian Leibe, and Liang-Chieh Chen · 2019
Earlier work this paper cites.
Towards high-resolution salient object detection
Yi Zeng, Pingping Zhang, Zhe L. Lin, Jianming Zhang, and Huchuan Lu · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Cited alongside, same era.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Cited alongside, same era.
FEELVOS: fast end-to-end embedding learning for video object segmentation
Paul Voigtlaender, Yuning Chai, Florian Schroff, Hartwig Adam, Bastian Leibe, and Liang-Chieh Chen · 2019
Cited alongside, same era.
Towards high-resolution salient object detection
Yi Zeng, Pingping Zhang, Zhe L. Lin, Jianming Zhang, and Huchuan Lu · 2019
Cited alongside, same era.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Epic-kitchens visor benchmark: Video segmentations and object relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu, Jian Ma, Amlan Kar, Richard Higgins, Sanja Fidler, David Fouhey, and Dima Damen · 2022
Later among the works it cites.
Convmae: Masked convolution meets masked autoencoders
Peng Gao, Teli Ma, Hongsheng Li, Ziyi Lin, Jifeng Dai, and Yu Qiao · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross B. Girshick · 2022
Later among the works it cites.
Vita: Video instance segmentation via object token association
Miran Heo, Sukjun Hwang, Seoung Wug Oh, Joon-Young Lee, and Seon Joo Kim · 2022
Later among the works it cites.
LVOS: A benchmark for long-term video object segmentation
Lingyi Hong, Wenchao Chen, Zhongying Liu, Wei Zhang, Pinxue Guo, Zhaoyu Chen, and Wenqiang Zhang · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Cascadepsp: Toward class-agnostic and very high-resolution segmentation via global and local refinement
Ho Kei Cheng, Jihoon Chung, Yu-Wing Tai, and Chi-Keung Tang · 2020
Cited alongside, same era.
Dice loss for data-imbalanced NLP tasks
Xiaoya Li, Xiaofei Sun, Yuxian Meng, Junjun Liang, Fei Wu, and Jiwei Li · 2020
Cited alongside, same era.
FSS-1000: A 1000-class dataset for few-shot segmentation
Xiang Li, Tianhan Wei, Yau Pun Chen, Yu-Wing Tai, and Chi-Keung Tang · 2020
Cited alongside, same era.
Video object segmentation with adaptive feature bank and uncertain-region refinement
Yongqing Liang, Xin Li, Navid Jafari, and Jim Chen · 2020
Cited alongside, same era.
Kernelized memory network for video object segmentation
Hongje Seong, Junhyuk Hyun, and Euntai Kim · 2020
Cited alongside, same era.
Collaborative video object segmentation by foreground-background integration
Zongxin Yang, Yunchao Wei, and Yi Yang · 2020
Cited alongside, same era.
DN-DETR: accelerate DETR training by introducing query denoising
Feng Li, Hao Zhang, Shilong Liu, Jian Guo, Lionel M. Ni, and Lei Zhang · 2022
Later among the works it cites.
Recurrent dynamic embedding for video object segmentation
Mingxing Li, Li Hu, Zhiwei Xiong, Bang Zhang, Pan Pan, and Dong Liu · 2022
Later among the works it cites.
Exploring plain vision transformer backbones for object detection
Yanghao Li, Hanzi Mao, Ross B. Girshick, and Kaiming He · 2022
Later among the works it cites.
Learning quality-aware dynamic memory for video object segmentation
Yong Liu, Ran Yu, Fei Yin, Xinyuan Zhao, Wei Zhao, Weihao Xia, and Yujiu Yang · 2022
Later among the works it cites.
Deit III: revenge of the vit
Hugo Touvron, Matthieu Cord, and Hervé Jégou · 2022
Later among the works it cites.
Evo-vit: Slow-fast token evolution for dynamic vision transformer
Yifan Xu, Zhijie Zhang, Mengdan Zhang, Kekai Sheng, Ke Li, Weiming Dong, Liqing Zhang, Changsheng Xu, and Xing Sun · 2022
Later among the works it cites.
Collaborative video object segmentation by multi-scale foreground-background integration
Zongxin Yang, Yunchao Wei, and Yi Yang · 2022
Later among the works it cites.
Decoupling features in hierarchical propagation for video object segmentation
Zongxin Yang and Yi Yang · 2022
Later among the works it cites.
Joint feature learning and relation modeling for tracking: A one-stream framework
Botao Ye, Hong Chang, Bingpeng Ma, Shiguang Shan, and Xilin Chen · 2022
Later among the works it cites.
Segvit: Semantic segmentation with plain vision transformers
Bowen Zhang, Zhi Tian, Quan Tang, Xiangxiang Chu, Xiaolin Wei, Chunhua Shen, and Yifan Liu · 2022
Later among the works it cites.
HODOR: high-level object descriptors for object re-segmentation in video learned from static images
Ali Athar, Jonathon Luiten, Alexander Hermans, Deva Ramanan, and Bastian Leibe · 2022
Later among the works it cites.
Backbone is all your need: A simplified architecture for visual object tracking
Boyu Chen, Peixia Li, Lei Bai, Lei Qiao, Qiuhong Shen, Bo Li, Weihao Gan, Wei Wu, and Wanli Ouyang · 2022
Later among the works it cites.
Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model
Ho Kei Cheng and Alexander G. Schwing · 2022
Later among the works it cites.
Mixformer: End-to-end tracking with iterative mixed attention
Yutao Cui, Cheng Jiang, Limin Wang, and Gangshan Wu · 2022
Later among the works it cites.
Epic-kitchens visor benchmark: Video segmentations and object relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu, Jian Ma, Amlan Kar, Richard Higgins, Sanja Fidler, David Fouhey, and Dima Damen · 2022
Later among the works it cites.
Convmae: Masked convolution meets masked autoencoders
Peng Gao, Teli Ma, Hongsheng Li, Ziyi Lin, Jifeng Dai, and Yu Qiao · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross B. Girshick · 2022
Later among the works it cites.
Vita: Video instance segmentation via object token association
Miran Heo, Sukjun Hwang, Seoung Wug Oh, Joon-Young Lee, and Seon Joo Kim · 2022
Later among the works it cites.
LVOS: A benchmark for long-term video object segmentation
Lingyi Hong, Wenchao Chen, Zhongying Liu, Wei Zhang, Pinxue Guo, Zhaoyu Chen, and Wenqiang Zhang · 2022
Later among the works it cites.
DN-DETR: accelerate DETR training by introducing query denoising
Feng Li, Hao Zhang, Shilong Liu, Jian Guo, Lionel M. Ni, and Lei Zhang · 2022
Later among the works it cites.
Recurrent dynamic embedding for video object segmentation
Mingxing Li, Li Hu, Zhiwei Xiong, Bang Zhang, Pan Pan, and Dong Liu · 2022
Later among the works it cites.
Exploring plain vision transformer backbones for object detection
Yanghao Li, Hanzi Mao, Ross B. Girshick, and Kaiming He · 2022
Later among the works it cites.
Learning quality-aware dynamic memory for video object segmentation
Yong Liu, Ran Yu, Fei Yin, Xinyuan Zhao, Wei Zhao, Weihao Xia, and Yujiu Yang · 2022
Later among the works it cites.
Deit III: revenge of the vit
Hugo Touvron, Matthieu Cord, and Hervé Jégou · 2022
Later among the works it cites.
Evo-vit: Slow-fast token evolution for dynamic vision transformer
Yifan Xu, Zhijie Zhang, Mengdan Zhang, Kekai Sheng, Ke Li, Weiming Dong, Liqing Zhang, Changsheng Xu, and Xing Sun · 2022
Later among the works it cites.
Collaborative video object segmentation by multi-scale foreground-background integration
Zongxin Yang, Yunchao Wei, and Yi Yang · 2022
Later among the works it cites.
Decoupling features in hierarchical propagation for video object segmentation
Zongxin Yang and Yi Yang · 2022
Later among the works it cites.
Joint feature learning and relation modeling for tracking: A one-stream framework
Botao Ye, Hong Chang, Bingpeng Ma, Shiguang Shan, and Xilin Chen · 2022
Later among the works it cites.
Segvit: Semantic segmentation with plain vision transformers
Bowen Zhang, Zhi Tian, Quan Tang, Xiangxiang Chu, Xiaolin Wei, Chunhua Shen, and Yifan Liu · 2022
Later among the works it cites.
Tracking anything with decoupled video segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Alexander G. Schwing, and Joon-Young Lee · 2023
Closest in time.
Mixformerv2: Efficient fully transformer tracking
Yutao Cui, Tianhui Song, Gangshan Wu, and Limin Wang · 2023
Closest in time.
MOSE: A new dataset for video object segmentation in complex scenes
Henghui Ding, Chang Liu, Shuting He, Xudong Jiang, Philip H. S. Torr, and Song Bai · 2023
Closest in time.
Which tokens to use? investigating token reduction in vision transformers
Joakim Bruslund Haurum, Sergio Escalera, Graham W. Taylor, and Thomas B. Moeslund · 2023
Closest in time.
Tracking through containers and occluders in the wild
Basile Van Hoorick, Pavel Tokmakov, Simon Stent, Jie Li, and Carl Vondrick · 2023
Closest in time.
Breaking the ”object” in video object segmentation
Pavel Tokmakov, Jie Li, and Adrien Gaidon · 2023
Closest in time.
Look before you match: Instance understanding matters in video object segmentation
Junke Wang, Dongdong Chen, Zuxuan Wu, Chong Luo, Chuanxin Tang, Xiyang Dai, Yucheng Zhao, Yujia Xie, Lu Yuan, and Yu-Gang Jiang · 2023
Closest in time.
Dropmae: Masked autoencoders with spatial-attention dropout for tracking tasks
Qiangqiang Wu, Tianyu Yang, Ziquan Liu, Baoyuan Wu, Ying Shan, and Antoni B. Chan · 2023
Closest in time.
Scalable video object segmentation with simplified framework
Qiangqiang Wu, Tianyu Yang, Wei Wu, and Antoni B. Chan · 2023
Closest in time.
Tracking anything with decoupled video segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Alexander G. Schwing, and Joon-Young Lee · 2023
Closest in time.
Mixformerv2: Efficient fully transformer tracking
Yutao Cui, Tianhui Song, Gangshan Wu, and Limin Wang · 2023
Closest in time.
MOSE: A new dataset for video object segmentation in complex scenes
Henghui Ding, Chang Liu, Shuting He, Xudong Jiang, Philip H. S. Torr, and Song Bai · 2023
Closest in time.
Which tokens to use? investigating token reduction in vision transformers
Joakim Bruslund Haurum, Sergio Escalera, Graham W. Taylor, and Thomas B. Moeslund · 2023
Closest in time.
Tracking through containers and occluders in the wild
Basile Van Hoorick, Pavel Tokmakov, Simon Stent, Jie Li, and Carl Vondrick · 2023
Closest in time.
Breaking the ”object” in video object segmentation
Pavel Tokmakov, Jie Li, and Adrien Gaidon · 2023
Closest in time.
Look before you match: Instance understanding matters in video object segmentation
Junke Wang, Dongdong Chen, Zuxuan Wu, Chong Luo, Chuanxin Tang, Xiyang Dai, Yucheng Zhao, Yujia Xie, Lu Yuan, and Yu-Gang Jiang · 2023
Closest in time.
Dropmae: Masked autoencoders with spatial-attention dropout for tracking tasks
Qiangqiang Wu, Tianyu Yang, Ziquan Liu, Baoyuan Wu, Ying Shan, and Antoni B. Chan · 2023
Closest in time.
Scalable video object segmentation with simplified framework
Qiangqiang Wu, Tianyu Yang, Wei Wu, and Antoni B. Chan · 2023
Closest in time.
Putting the object back into video object segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Joon-Young Lee, and Alexander G. Schwing · 2024
Closest in time.
Putting the object back into video object segmentation
Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Joon-Young Lee, and Alexander G. Schwing · 2024
Closest in time.