Fetching the paper…
Reading the bibliography…
This paper delves into the challenges of achieving scalable and effective multi-object modeling for semi-supervised Video Object Segmentation (VOS).
B. T. Polyak and A. B. Juditsky, “Acceleration of stochastic approximation by averaging,”
1992
Earlier work this paper cites.
V. Badrinarayanan, F. Galasso, and R. Cipolla, “Label propagation in video sequences,” in
2010
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (voc) challenge,”
2010
Earlier work this paper cites.
B. Hariharan, P. Arbeláez, L. Bourdev, S. Maji, and J. Malik, “Semantic contours from inverse detectors,” in
2011
Earlier work this paper cites.
S. Vijayanarasimhan and K. Grauman, “Active frame selection for label propagation in videos,” in
2012
Earlier work this paper cites.
S. Avinash Ramakanth and R. Venkatesh Babu, “Seamseg: Video object segmentation using patch seams,” in
2014
Earlier work this paper cites.
S. Nowozin, “Optimal decisions from probabilistic models: the intersection-over-union case,” in
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in
2014
Earlier work this paper cites.
M.-M. Cheng, N. J. Mitra, X. Huang, P. H. Torr, and S.-M. Hu, “Global contrast based salient region detection,”
2014
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,”
2015
Earlier work this paper cites.
L. Chen, J. Shen, W. Wang, and B. Ni, “Video object segmentation via dense trajectories,”
2015
Earlier work this paper cites.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in
2015
Earlier work this paper cites.
J. Shi, Q. Yan, L. Xu, and J. Jia, “Hierarchical image saliency detection on extended cssd,”
2015
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in
2015
Earlier work this paper cites.
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. Van Gool, M. Gross, and A. Sorkine-Hornung, “A benchmark dataset and evaluation methodology for video object segmentation,” in
2016
Earlier work this paper cites.
S. Teerapittayanon, B. McDanel, and H.-T. Kung, “Branchynet: Fast inference via early exiting from deep neural networks,” in
2016
Earlier work this paper cites.
J. L. Ba, J. R. Kiros, and G. E. Hinton, “Layer normalization,” in
2016
Earlier work this paper cites.
D. Hendrycks and K. Gimpel, “Gaussian error linear units (gelus),”
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
G. Huang, Y. Sun, Z. Liu, D. Sedra, and K. Q. Weinberger, “Deep networks with stochastic depth,” in
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Caelles, K.-K. Maninis, J. Pont-Tuset, L. Leal-Taixé, D. Cremers, and L. Van Gool, “One-shot video object segmentation,” in
2017
Earlier work this paper cites.
P. Voigtlaender and B. Leibe, “Online adaptation of convolutional neural networks for video object segmentation,” in
2017
Earlier work this paper cites.
F. Perazzi, A. Khoreva, R. Benenson, B. Schiele, and A. Sorkine-Hornung, “Learning video object segmentation from static images,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Earlier work this paper cites.
T. Bolukbasi, J. Wang, O. Dekel, and V. Saligrama, “Adaptive neural networks for efficient inference,” in
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, and S. Belongie, “Feature pyramid networks for object detection,” in
2017
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” in
2018
Earlier work this paper cites.
H. Xiao, J. Feng, G. Lin, Y. Liu, and M. Zhang, “Monet: Deep motion exploitation for video object segmentation,” in
2018
Earlier work this paper cites.
J. Luiten, P. Voigtlaender, and B. Leibe, “Premvos: Proposal-generation, refinement and merging for video object segmentation,” in
2018
Earlier work this paper cites.
L. Yang, Y. Wang, X. Xiong, J. Yang, and A. K. Katsaggelos, “Efficient video object segmentation via network modulation,” in
2018
Earlier work this paper cites.
Y. Chen, J. Pont-Tuset, A. Montes, and L. Van Gool, “Blazingly fast video object segmentation with pixel-wise metric learning,” in
2018
Earlier work this paper cites.
Y.-T. Hu, J.-B. Huang, and A. G. Schwing, “Videomatch: Matching based video object segmentation,” in
2018
Earlier work this paper cites.
S. Wug Oh, J.-Y. Lee, K. Sunkavalli, and S. Joo Kim, “Fast video object segmentation by reference-guided mask propagation,” in
2018
Earlier work this paper cites.
W. Wang, J. Shen, F. Porikli, and R. Yang, “Semi-supervised video object segmentation with super-trajectories,”
2018
Earlier work this paper cites.
J. Yu, L. Yang, N. Xu, J. Yang, and T. Huang, “Slimmable neural networks,” in
2018
Cited alongside, same era.
N. Parmar, A. Vaswani, J. Uszkoreit, L. Kaiser, N. Shazeer, A. Ku, and D. Tran, “Image transformer,” in
2018
Cited alongside, same era.
Y. Wu and K. He, “Group normalization,” in
2018
Cited alongside, same era.
S. W. Oh, J.-Y. Lee, N. Xu, and S. J. Kim, “Video object segmentation using space-time memory networks,” in
2019
Cited alongside, same era.
P. Voigtlaender, Y. Chai, F. Schroff, H. Adam, B. Leibe, and L.-C. Chen, “Feelvos: Fast end-to-end embedding learning for video object segmentation,” in
2019
Cited alongside, same era.
Z. Yang, P. Li, Q. Feng, Y. Wei, and Y. Yang, “Going deeper into embedding learning for video object segmentation,” in
W. Wang, M. Feiszli, H. Wang, and D. Tran, “Unidentified video objects: A benchmark for dense, open-world segmentation,” in
2021
Later among the works it cites.
B. Duke, A. Ahmed, C. Wolf, P. Aarabi, and G. W. Taylor, “Sstvos: Sparse spatiotemporal transformers for video object segmentation,” in
2021
Later among the works it cites.
H. Seong, S. W. Oh, J.-Y. Lee, S. Lee, S. Lee, and E. Kim, “Hierarchical memory matching network for video object segmentation,” in
2021
Later among the works it cites.
T. Zhou, F. Porikli, D. J. Crandall, L. Van Gool, and W. Wang, “A survey on deep learning technique for video segmentation,”
2022
Closest in time.
X. Xu, J. Wang, X. Li, and Y. Lu, “Reliable propagation-correction modulation for video object segmentation,” in
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,”
2019
Cited alongside, same era.
H. Lin, X. Qi, and J. Jia, “Agss-vos: Attention guided single-shot video object segmentation,” in
2019
Cited alongside, same era.
I. Loshchilov and F. Hutter, “Decoupled weight decay regularization,” in
2019
Cited alongside, same era.
H. Seong, J. Hyun, and E. Kim, “Kernelized memory network for video object segmentation,” in
2020
Cited alongside, same era.
X. Lu, W. Wang, M. Danelljan, T. Zhou, J. Shen, and L. Van Gool, “Video object segmentation with episodic graph memory networks,” in
2020
Cited alongside, same era.
S. Cho, H. Lee, M. Lee, C. Park, S. Jang, M. Kim, and S. Lee, “Tackling background distraction in video object segmentation,” in
2022
Closest in time.
J. Miao, X. Wang, Y. Wu, W. Li, X. Zhang, Y. Wei, and Y. Yang, “Large-scale video panoptic segmentation in the wild: A benchmark,” in
2022
Closest in time.
F. Zhu, Z. Yang, X. Yu, Y. Yang, and Y. Wei, “Instance as identity: A generic online paradigm for video instance segmentation,” in
2022
Closest in time.
Z. Yang and Y. Yang, “Decoupling features in hierarchical propagation for video object segmentation,” in
2022
Closest in time.
Y. Yu, J. Yuan, G. Mittal, L. Fuxin, and M. Chen, “Batman: Bilateral attention transformer in motion-appearance neighboring space for video object segmentation,” in
2022
Closest in time.
H. Seong, J. Hyun, and E. Kim, “Video object segmentation using kernelized memory network with multiple kernels,”
2022
Closest in time.
H. K. Cheng and A. G. Schwing, “Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model,” in
2022
Closest in time.
M. Li, L. Hu, Z. Xiong, B. Zhang, P. Pan, and D. Liu, “Recurrent dynamic embedding for video object segmentation,” in
2022
Closest in time.
X. Wang, L. Zhu, Z. Zheng, M. Xu, and Y. Yang, “Align and tell: Boosting text-video retrieval with local alignment and fine-grained supervision,” 2022
2022
Closest in time.
Z. Liu, J. Ning, Y. Cao, Y. Wei, Z. Zhang, S. Lin, and H. Hu, “Video swin transformer,” in
2022
Closest in time.
B. Yan, Y. Jiang, P. Sun, D. Wang, Z. Yuan, P. Luo, and H. Lu, “Towards grand unification of object tracking,” in
2022
Closest in time.
T. Meinhardt, A. Kirillov, L. Leal-Taixe, and C. Feichtenhofer, “Trackformer: Multi-object tracking with transformers,” in
2022
Closest in time.
Y. Liu, R. Yu, J. Wang, X. Zhao, Y. Wang, Y. Tang, and Y. Yang, “Global spectral filter memory network for video object segmentation,” in
2022
Closest in time.
Y. Liu, R. Yu, F. Yin, X. Zhao, W. Zhao, W. Xia, and Y. Yang, “Learning quality-aware dynamic memory for video object segmentation,” in
2022
Closest in time.
K. Park, S. Woo, S. W. Oh, I. S. Kweon, and J.-Y. Lee, “Per-clip video object segmentation,” in
2022
Closest in time.
R. Miles, M. K. Yucel, B. Manganelli, and A. Saà-Garriga, “Mobilevos: Real-time video object segmentation contrastive learning meets knowledge distillation,” in
2023
Closest in time.
Y. Cheng, L. Li, Y. Xu, X. Li, Z. Yang, W. Wang, and Y. Yang, “Segment and track anything,”
2023
Closest in time.
Y. Xu, Z. Yang, and Y. Yang, “Integrating boxes and masks: A multi-object framework for unified visual tracking and segmentation,” in
2023
Closest in time.
K. Li, Z. Yang, L. Chen, Y. Yang, and J. Xiao, “Catr: Combinatorial-dependence audio-queried transformer for audio-visual video segmentation,” in
2023
Closest in time.
Y. Xu, Z. Yang, and Y. Yang , “Video object segmentation in panoptic wild scenes,” in
2023
Closest in time.
2023
Closest in time.
P. Tokmakov, J. Li, and A. Gaidon, “Breaking the” object” in video object segmentation,” in
2023
Closest in time.
C. Liang, W. Wang, T. Zhou, J. Miao, Y. Luo, and Y. Yang, “Local-global context aware transformer for language-guided video segmentation,”
2023
Closest in time.
Y. Zhang, L. Li, W. Wang, R. Xie, L. Song, and W. Zhang, “Boosting video object segmentation via space-time correspondence learning,” in
2023
Closest in time.
J. Wang, D. Chen, Z. Wu, C. Luo, C. Tang, X. Dai, Y. Zhao, Y. Xie, L. Yuan, and Y.-G. Jiang, “Look before you match: Instance understanding matters in video object segmentation,” in
2023
Closest in time.
Y. Lu, F. Ni, H. Wang, X. Guo, L. Zhu, Z. Yang, R. Song, L. Cheng, and Y. Yang, “Show me a video: A large-scale narrated video dataset for coherent story illustration,”
2023
Closest in time.
2023
Closest in time.
P. Chu, J. Wang, Q. You, H. Ling, and Z. Liu, “Transmot: Spatial-temporal graph transformer for multiple object tracking,” in
2023
Closest in time.
M. Lan, J. Zhang, L. Zhang, and D. Tao, “Learning to learn better for video object segmentation,” in
2023
Closest in time.
Z. Yang, G. Chen, X. Li, W. Wang, and Y. Yang, “Doraemongpt: Toward understanding dynamic scenes with large language models,” 2024
2024
Closest in time.
C. Mayer, M. Danelljan, M.-H. Yang, V. Ferrari, L. Van Gool, and A. Kuznetsova, “Beyond sot: It’s time to track multiple generic objects at once,” in
2024
Closest in time.