Fetching the paper…
Reading the bibliography…
The task of object segmentation in videos is usually accomplished by processing appearance and motion information separately using standard 2D convolutional networks, followed by a learned fusion of the two sources of information.
A benchmark for the comparison of 3-d motion segmentation algorithms
Roberto Tron and René Vidal · 2007
Earlier work this paper cites.
Primary object segmentation in videos based on region augmentation and reduction
Chang-Su Kim Yeong Jun Koh · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
3d convolutional neural networks for human action recognition
Shuiwang Ji, Wei Xu, Ming Yang, and Kai Yu · 2012
Earlier work this paper cites.
Segmentation of moving objects by long term video analysis
Peter Ochs, Jitendra Malik, and Thomas Brox · 2013
Earlier work this paper cites.
Video saliency incorporating spatiotemporal cues and uncertainty weighting
Yuming Fang, Zhou Wang, Weisi Lin, and Zhijun Fang · 2014
Earlier work this paper cites.
Large-scale video classification with convolutional neural networks
Andrej Karpathy, George Toderici, Sanketh Shetty, Thomas Leung, Rahul Sukthankar, and Li Fei-Fei · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
Superpixel-based spatiotemporal saliency detection
Zhi Liu, Xiang Zhang, Shuhua Luo, and Olivier Le Meur · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Video object discovery and co-segmentation with extremely weak supervision
Le Wang, Gang Hua, Rahul Sukthankar, Jianru Xue, and Nanning Zheng · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Earlier work this paper cites.
Unsupervised object discovery and tracking in video collections
Suha Kwak, Minsu Cho, Ivan Laptev, Jean Ponce, and Cordelia Schmid · 2015
Earlier work this paper cites.
Learning spatiotemporal features with 3d convolutional networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Earlier work this paper cites.
Saliency-aware geodesic video object segmentation
Wenguan Wang, Jianbing Shen, and Fatih Porikli · 2015
Earlier work this paper cites.
Consistent video saliency using local gradient flow optimization and global refinement
Ling Shao Wenguan Wang, Jianbing Shen · 2015
Earlier work this paper cites.
Convolutional lstm network: A machine learning approach for precipitation nowcasting
SHI Xingjian, Zhourong Chen, Hao Wang, Dit-Yan Yeung, Wai-Kin Wong, and Wang-chun Woo · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Saliency detection for unconstrained videos using superpixel-level graph and spatiotemporal propagation
Zhi Liu, Junhao Li, Linwei Ye, Guangling Sun, and Liquan Shen · 2016
Earlier work this paper cites.
Learning to refine object segments
P.H.O. Pinheiro, T.-Y. Lin, R. Collobert, and P. Dollár · 2016
Cited alongside, same era.
A benchmark dataset and evaluation methodology for video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alex Sorkine-Hornung, and Luc Van Gool · 2016
Cited alongside, same era.
Track and segment: An iterative unsupervised approach for video object proposals
Fanyi Xiao and Yong Jae Lee · 2016
Cited alongside, same era.
One-shot video object segmentation
S. Caelles, K.-K. Maninis, J. Pont-Tuset, L. Leal-Taixé, D. Cremers, and L. Van Gool · 2017
Cited alongside, same era.
Quo vadis, action recognition? A new model and the kinetics dataset
João Carreira and Andrew Zisserman · 2017
Cited alongside, same era.
Rethinking atrous convolution for semantic image segmentation
Flow guided recurrent neural encoder for video salient object detection
Li Guanbin, Xie Yuan, Wei Tianhao, Wang Keze, and Lin Liang · 2018
Later among the works it cites.
Can spatiotemporal 3d cnns retrace the history of 2d cnns and imagenet?
Kensho Hara, Hirokatsu Kataoka, and Yutaka Satoh · 2018
Later among the works it cites.
Flow guided recurrent neural encoder for video salient object detection
Guanbin Li, Yuan Xie, Tianhao Wei, Keze Wang, and Liang Lin · 2018
Later among the works it cites.
Video segmentation using teacher-student adaptation in a human robot interaction (HRI) setting
Mennatullah Siam, Chen Jiang, Steven Weikai Lu, Laura Petrich, Mahmoud Gamal, Mohamed Elhoseiny, and Martin Jägersand · 2018
Later among the works it cites.
Pyramid dilated deeper convlstm for video salient object detection
Hongmei Song, Wenguan Wang, Sanyuan Zhao, Jianbing Shen, and Kin-Man Lam · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Liang-Chieh Chen, George Papandreou, Florian Schroff, and Hartwig Adam · 2017
Cited alongside, same era.
Xception: Deep learning with depthwise separable convolutions
François Chollet · 2017
Cited alongside, same era.
Mask R-CNN
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Cited alongside, same era.
Fusionseg: Learning to combine motion and appearance for fully automatic segmention of generic objects in videos
Suyog Jain, Bo Xiong, and Kristen Grauman · 2017
Cited alongside, same era.
The kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, et al · 2017
Cited alongside, same era.
Saliency detection for unconstrained videos using superpixel-level graph and spatiotemporal propagation
Zhi Liu, Junhao Li, Linwei Ye, Guangling Sun, and Liquan Shen · 2017
Cited alongside, same era.
Large kernel matters–improve semantic segmentation by global convolutional network
Chao Peng, Xiangyu Zhang, Gang Yu, Guiming Luo, and Jian Sun · 2017
Cited alongside, same era.
Du Tran, Heng Wang, Lorenzo Torresani, Jamie Ray, Yann LeCun, and Manohar Paluri · 2018
Later among the works it cites.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Later among the works it cites.
Video salient object detection via fully convolutional networks
Ling Shao Wenguan Wang, Jianbing Shen · 2018
Later among the works it cites.
Fast video object segmentation by reference-guided mask propagation
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Later among the works it cites.
Rethinking spatiotemporal feature learning: Speed-accuracy trade-offs in video classification
Saining Xie, Chen Sun, Jonathan Huang, Zhuowen Tu, and Kevin Murphy · 2018
Later among the works it cites.
YouTube-VOS: Sequence-to-sequence video object segmentation
Ning Xu, Linjie Yang, Yuchen Fan, Jianchao Yang, Dingcheng Yue, Yuchen Liang, Brian Price, Scott Cohen, and Thomas Huang · 2018
Later among the works it cites.
Large-scale weakly-supervised pre-training for video action recognition
Deepti Ghadiyaram, Matt Feiszli, Du Tran, Xueting Yan, Heng Wang, and Dhruv Mahajan · 2019
Later among the works it cites.
An efficient 3d CNN for action/object segmentation in video
Rui Hou, Chen Chen, Rahul Sukthankar, and Mubarak Shah · 2019
Later among the works it cites.
Large-scale object mining for object discovery from unlabeled video
Aljoša Ošep, Paul Voigtlaender, Jonathon Luiten, Stefan Breuers, and Bastian Leibe · 2019
Later among the works it cites.
Video classification with channel-separated convolutional networks
Du Tran, Heng Wang, Lorenzo Torresani, and Matt Feiszli · 2019
Later among the works it cites.
MOTS: Multi-object tracking and segmentation
Paul Voigtlaender, Michael Krause, Aljosa Osep, Jonathon Luiten, B.B.G Sekar, Andreas Geiger, and Bastian Leibe · 2019
Later among the works it cites.
Zero-shot video object segmentation via attentive graph neural networks
Wenguan Wang, Xiankai Lu, Jianbing Shen, David J. Crandall, and Ling Shao · 2019
Later among the works it cites.
Stem-seg: Spatio-temporal embeddings for instance segmentation in videos
Ali Athar, Sabarinath Mahadevan, Aljoša Ošep, Laura Leal-Taixé, and Bastian Leibe · 2020
Closest in time.