Fetching the paper…
Reading the bibliography…
Recent interest in self-supervised dense tracking has yielded rapid progress, but performance still remains far from supervised methods.
Visual responses in the newborn
T. Berry Brazelton, Mary Louise Scholl, and John S. Robey · 1966
Earlier work this paper cites.
Smooth-pursuit eye movements in the newborn infant
Janet P. Kremenitzer, Herbert G. Vaughan, Diane Kurtzberg, and Kathryn Dowling · 1979
Earlier work this paper cites.
Eye–hand coordination in the newborn
C. von Hofsten · 1982
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Color transfer between images
Erik Reinhard, Michael Adhikhmin, Bruce Gooch, and Peter Shirley · 2001
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Fei-Fei Li · 2009
Earlier work this paper cites.
The PASCAL Visual Object Classes Challenge 2009 (VOC2009) Results
Mark Everingham, Luc Van Gool, Christopher K. I. Williams, John Winn, and Andrew Zisserman · 2009
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Alex Graves, Greg Wayne, and Ivo Danihelka · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, Lubomir Bourdev, Ross Girshick, James Hays, Pietro Perona, Deva Ramanan, C. Lawrence Zitnick, and Piotr Dollar · 2014
Earlier work this paper cites.
Learning to see by moving
Pulkit Agrawal, Joao Carreira, and Jitendra Malik · 2015
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Learning visual groups from co-occurrences in space and time
Phillip Isola, Daniel Zoran, Dilip Krishnan, and Edward H. Adelson · 2015
Earlier work this paper cites.
Spatial transformer networks
Max Jaderberg, Karen Simonyan, Andrew Zisserman, et al · 2015
Earlier work this paper cites.
Learning image representations tied to ego-motion
Dinesh Jayaraman and Kristen Grauman · 2015
Earlier work this paper cites.
End-to-end memory networks
Sainbayar Sukhbaatar, Jason Weston, Rob Fergus, et al · 2015
Earlier work this paper cites.
Unsupervised learning of visual representations using videos
Xiaolong Wang and Abhinav Gupta · 2015
Earlier work this paper cites.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Earlier work this paper cites.
Describing videos by exploiting temporal structure
Li Yao, Atousa Torabi, Kyunghyun Cho, Nicolas Ballas, Christopher Pal, Hugo Larochelle, and Aaron Courville · 2015
Earlier work this paper cites.
Dynamic filter networks
Bert De Brabandere, Xu Jia, Tinne Tuytelaars, and Luc Van Gool · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Slow and steady feature analysis: higher order temporal coherence in video
Dinesh Jayaraman and Kristen Grauman · 2016
Earlier work this paper cites.
Learning video object segmentation from static images
Anna Khoreva, Federico Perazzi, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
A novel performance evaluation methodology for single-target trackers
Matej Kristan, Jiri Matas, Aleš Leonardis, Tomas Vojir, Roman Pflugfelder, Gustavo Fernandez, Georg Nebehay, Fatih Porikli, and Luka Čehovin · 2016
Earlier work this paper cites.
Ask me anything: Dynamic memory networks for natural language processing
Ankit Kumar, Ozan Irsoy, Peter Ondruska, Mohit Iyyer, James Bradbury, Ishaan Gulrajani, Victor Zhong, Romain Paulus, and Richard Socher · 2016
Earlier work this paper cites.
Shuffle and learn: Unsupervised learning using temporal order verification
Ishan Misra, C. Lawrence Zitnick, and Martial Hebert · 2016
Cited alongside, same era.
Pixel recurrent neural networks
Aaron van den Oord, Nal Kalchbrenner, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
A benchmark dataset and evaluation methodology for video object segmentation
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. Van Gool, M. Gross, and A. Sorkine-Hornung · 2016
Cited alongside, same era.
Richard Zhang, Phillip Isola, and Alexei Efros · 2016
Cited alongside, same era.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Cited alongside, same era.
Unsupervised learning of disentangled representations from video
Premvos: Proposal-generation, refinement and merging for video object segmentation
Jonathon Luiten, Paul Voigtlaender, and Bastian Leibe · 2018
Later among the works it cites.
Video object segmentation without temporal information
Kevis-Kokitsi Maninis, Sergi Caelles, Yuhua Chen, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2018
Later among the works it cites.
Fast video object segmentation by reference-guided mask propagation
Seoung Wug Oh, Joon-Young Lee, Kalyan Sunkavalli, and Seon Joo Kim · 2018
Later among the works it cites.
Long-term tracking in the wild: A benchmark
Jack Valmadre, Luca Bertinetto, Joao F. Henriques, Ran Tao, Andrea Vedaldi, Arnold Smeulders, Philip Torr, and Efstratios Gavves · 2018
Later among the works it cites.
Tracking emerges by colorizing videos
Carl Vondrick, Abhinav Shrivastava, Alireza Fathi, Sergio Guadarrama, and Kevin Murphy · 2018
Later among the works it cites.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Emily Denton and Vighnesh Birodkar · 2017
Cited alongside, same era.
Self-supervised video representation learning with odd-one-out networks
Basura Fernando, Hakan Bilen, Efstratios Gavves, and Stephen Gould · 2017
Cited alongside, same era.
Maskrnn: Instance level video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G. Schwing · 2017
Cited alongside, same era.
Lucid data dreaming for multiple object tracking
Anna Khoreva, Rodrigo Benenson, Eddy Ilg, Thomas Brox, and Bernt Schiele · 2017
Cited alongside, same era.
Unsupervised representation learning by sorting sequences
Hsin-Ying Lee, Jia-Bin Huang, Maneesh Singh, and Ming-Hsuan Yang · 2017
Cited alongside, same era.
The 2017 davis challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alex Sorkine-Hornung, and Luc Van Gool · 2017
Cited alongside, same era.
Get to the point: Summarization with pointer-generator networks
Abigail See, Peter J Liu, and Christopher D Manning · 2017
Cited alongside, same era.
Later among the works it cites.
Learning and using the arrow of time
Donglai Wei, Joseph Lim, Andrew Zisserman, and William T. Freeman · 2018
Later among the works it cites.
Self-supervised learning of a facial attribute embedding from video
Olivia Wiles, A. Sophia Koepke, and Andrew Zisserman · 2018
Later among the works it cites.
X2face: A network for controlling face generation using images, audio, and pose codes
Olivia Wiles, A. Sophia Koepke, and Andrew Zisserman · 2018
Later among the works it cites.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas Huang · 2018
Later among the works it cites.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos Katsaggelos · 2018
Later among the works it cites.
Long-Term Feature Banks for Detailed Video Understanding
Wu Chao-Yuan, Feichtenhofer Christoph, Fan Haoqi, He Kaiming, Krähenbühl Philipp, and Girshick Ross · 2019
Later among the works it cites.
Video representation learning by dense predictive coding
Tengda Han, Weidi Xie, and Andrew Zisserman · 2019
Later among the works it cites.
A generative appearance model for end-to-end video object segmentation
Joakim Johnander, Martin Danelljan, Emil Brissman, Fahad Shahbaz Khan, and Michael Felsberg · 2019
Later among the works it cites.
Multigrid predictive filter flow for unsupervised learning on videos
Shu Kong and Charless Fowlkes · 2019
Later among the works it cites.
Self-supervised learning for video correspondence flow
Zihang Lai and Weidi Xie · 2019
Later among the works it cites.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Later among the works it cites.
Rvos: End-to-end recurrent network for video object segmentation
Carles Ventura, Miriam Bellver, Andreu Girbau, Amaia Salvador, Ferran Marques, and Xavier Giro-i Nieto · 2019
Later among the works it cites.
Rvos: End-to-end recurrent network for video object segmentation
Carles Ventura, Miriam Bellver, Andreu Girbau, Amaia Salvador, Ferran Marques, and Xavier Giro-i Nieto · 2019
Later among the works it cites.
Feelvos: Fast end-to-end embedding learning for video object segmentation
Paul Voigtlaender, Yuning Chai, Florian Schroff, Hartwig Adam, Bastian Leibe, and Liang-Chieh Chen · 2019
Later among the works it cites.
Unsupervised deep tracking
Ning Wang, Yibing Song, Chao Ma, Wengang Zhou, Wei Liu, and Houqiang Li · 2019
Later among the works it cites.
Fast online object tracking and segmentation: A unifying approach
Qiang Wang, Li Zhang, Luca Bertinetto, Weiming Hu, and Philip H.S. Torr · 2019
Later among the works it cites.
Learning correspondence from the cycle-consistency of time
Xiaolong Wang, Allan Jabri, and Alexei A. Efros · 2019
Later among the works it cites.
Ranet: Ranking attention network for fast video object segmentation
Ziqin Wang, Jun Xu, Li Liu, Fan Zhu, and Ling Shao · 2019
Later among the works it cites.
Joint-task self-supervised learning for temporal correspondence
Li Xueting, Liu Sifei, De Mello Shalini, Wang Xiaolong, Kautz Jan, and Yang Ming-Hsuan · 2019
Later among the works it cites.