Fetching the paper…
Reading the bibliography…
In this paper we introduce a Transformer-based approach to video object segmentation (VOS).
Do convnets learn correspondence?
Jonathan L Long, Ning Zhang, and Trevor Darrell · 2014
Earlier work this paper cites.
Flownet: Learning optical flow with convolutional networks
Alexey Dosovitskiy, Philipp Fischer, Eddy Ilg, Philip Hausser, Caner Hazirbas, Vladimir Golkov, Patrick van der Smagt, Daniel Cremers, and Thomas Brox · 2015
Earlier work this paper cites.
Very deep convolutional networks for Large-Scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. Van Gool, M. Gross, and A. Sorkine-Hornung · 2016
Earlier work this paper cites.
Video segmentation via object flow
Yi-Hsuan Tsai, Ming-Hsuan Yang, and Michael J. Black · 2016
Earlier work this paper cites.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Earlier work this paper cites.
Segflow: Joint learning for video object segmentation and optical flow
Jingchun Cheng, Yi-Hsuan Tsai, Shengjin Wang, and Ming-Hsuan Yang · 2017
Earlier work this paper cites.
Fusionseg: Learning to combine motion and appearance for fully automatic segmentation of generic objects in videos
Suyog Dutt Jain, Bo Xiong, and Kristen Grauman · 2017
Earlier work this paper cites.
Maskrnn: Instance level video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander Schwing · 2017
Earlier work this paper cites.
Video propagation networks
Varun Jampani, Raghudeep Gadde, and Peter V. Gehler · 2017
Earlier work this paper cites.
Online video object segmentation via convolutional trident network
Won-Dong Jang and Chang-Su Kim · 2017
Earlier work this paper cites.
Learning video object segmentation from static images
Anna Khoreva, Federico Perazzi, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Earlier work this paper cites.
The 2017 davis challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alexander Sorkine-Hornung, and Luc Van Gool · 2017
Earlier work this paper cites.
Recurrent neural networks for semantic instance segmentation
Amaia Salvador, Miriam Bellver, Manel Baradad, Ferran Marques, Jordi Torres, and Xavier Giro-i Nieto · 2017
Earlier work this paper cites.
Learning video object segmentation with visual memory
Pavel Tokmakov, Karteek Alahari, and Cordelia Schmid · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Online adaptation of convolutional neural networks for video object segmentation
Paul Voigtlaender and Bastian Leibe · 2017
Earlier work this paper cites.
CNN in MRF: video object segmentation via inference in A cnn-based higher-order spatio-temporal MRF
Linchao Bao, Baoyuan Wu, and Wei Liu · 2018
Earlier work this paper cites.
Fast and accurate online video object segmentation via tracking parts
Jingchun Cheng, Yi-Hsuan Tsai, Wei-Chih Hung, Shengjin Wang, and Ming-Hsuan Yang · 2018
Earlier work this paper cites.
Motion-guided cascaded refinement network for video object segmentation
Ping Hu, Gang Wang, Xiangfei Kong, Jason Kuen, and Yap-Peng Tan · 2018
Cited alongside, same era.
Videomatch: Matching based video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G. Schwing · 2018
Cited alongside, same era.
Video object segmentation with joint re-identification and attention-aware mask propagation
Xiaoxiao Li and Chen Change Loy · 2018
Cited alongside, same era.
Premvos: Proposal-generation, refinement and merging for video object segmentation
Jonathon Luiten, Paul Voigtlaender, and Bastian Leibe · 2018
Cited alongside, same era.
Video object segmentation without temporal information
Kevis-Kokitsi Maninis, Sergi Caelles, Yuhua Chen, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2018
Cited alongside, same era.
Fast video object segmentation by reference-guided mask propagation
ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
J. Lu, D. Batra, D. Parikh, and S. Lee · 2019
Later among the works it cites.
RWTH ASR Systems for LibriSpeech: Hybrid vs Attention
Christoph Lüscher, Eugen Beck, Kazuki Irie, Markus Kitza, Wilfried Michel, Albert Zeyer, Ralf Schlüter, and Hermann Ney · 2019
Later among the works it cites.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Later among the works it cites.
Learning Video Representations using Contrastive Bidirectional Transformer
C. Sun, F. Baradel, K. Murphy, and C. Schmid · 2019
Later among the works it cites.
LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Hao Tan and Mohit Bansal · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Cited alongside, same era.
Non-local Neural Networks
X. Wang, R. Girshick, A. Gupta, and K. He · 2018
Cited alongside, same era.
Youtube-vos: Sequence-to-sequence video object segmentation
Ning Xu, Linjie Yang, Yuchen Fan, Jianchao Yang, Dingcheng Yue, Yuchen Liang, Brian Price, Scott Cohen, and Thomas Huang · 2018
Cited alongside, same era.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas S. Huang · 2018
Cited alongside, same era.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos K. Katsaggelos · 2018
Cited alongside, same era.
Icnet for real-time semantic segmentation on high-resolution images
Hengshuang Zhao, Xiaojuan Qi, Xiaoyong Shen, Jianping Shi, and Jiaya Jia · 2018
Cited alongside, same era.
Generating long sequences with sparse transformers
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever · 2019
Cited alongside, same era.
EfficientNet: Rethinking model scaling for convolutional neural networks
Mingxing Tan and Quoc Le · 2019
Later among the works it cites.
Rvos: End-to-end recurrent network for video object segmentation
Carles Ventura, Miriam Bellver, Andreu Girbau, Amaia Salvador, Ferran Marques, and Xavier Giro-i Nieto · 2019
Later among the works it cites.
Feelvos: Fast end-to-end embedding learning for video object segmentation
Paul Voigtlaender, Yuning Chai, Florian Schroff, Hartwig Adam, Bastian Leibe, and Liang-Chieh Chen · 2019
Later among the works it cites.
BoLTVOS: Box-Level Tracking for Video Object Segmentation
Paul Voigtlaender, Jonathon Luiten, and Bastian Leibe · 2019
Later among the works it cites.
Analyzing Multi-Head Self-Attention: Specialized Heads Do the Heavy Lifting, the Rest Can Be Pruned
Elena Voita, David Talbot, Fedor Moiseev, Rico Sennrich, and Ivan Titov · 2019
Later among the works it cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Anonymous · 2020
Later among the works it cites.
Video Action Transformer Network
R. Girdhar, J. Carreira, C. Doersch, and A. Zisserman · 2020
Later among the works it cites.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya · 2020
Later among the works it cites.
Kernelized memory network for video object segmentation
Hongje Seong, Junhyuk Hyun, and Euntai Kim · 2020
Later among the works it cites.
Collaborative video object segmentation by foreground-background integration
Zongxin Yang, Yunchao Wei, and Yi Yang · 2020
Later among the works it cites.
Exploring Self-attention for Image Recognition
H. Zhao, J. Jia, and V. Koltun · 2020
Later among the works it cites.
Sparse spatiotemporal transformer
Brendan Duke · 2021
Closest in time.
How Transferrable are Reasoning Patterns in VQA?
C. Kervadec, T. Jaunet, G. Antipov, M. Baccouche, R. Vuillemot, and C. Wolf · 2021
Closest in time.