Fetching the paper…
Reading the bibliography…
In this paper, we consider the task of unsupervised object discovery in videos.
Untersuchungen zur lehre von der gestalt. ii
Max Wertheimer · 1923
Earlier work this paper cites.
The senses considered as perceptual systems
James Jerome Gibson and Leonard Carmichael · 1966
Earlier work this paper cites.
Conditional random fields: Probabilistic models for segmenting and labeling sequence data
John Lafferty, Andrew McCallum, and Fernando CN Pereira · 2001
Earlier work this paper cites.
Object segmentation by long term analysis of point trajectories
Thomas Brox and Jitendra Malik · 2010
Earlier work this paper cites.
Track to the future: Spatio-temporal video segmentation with long-range motion cues
José Lezama, Karteek Alahari, Josef Sivic, and Ivan Laptev · 2011
Earlier work this paper cites.
How to grow a mind: Statistics, structure, and abstraction
Joshua B Tenenbaum, Charles Kemp, Thomas L Griffiths, and Noah D Goodman · 2011
Earlier work this paper cites.
Video segmentation by tracing discontinuities in a trajectory embedding
Katerina Fragkiadaki, Geng Zhang, and Jianbo Shi · 2012
Earlier work this paper cites.
Higher order motion models and spectral clustering
Peter Ochs and Thomas Brox · 2012
Earlier work this paper cites.
Video segmentation by tracking many figure-ground segments
Fuxin Li, Taeyoung Kim, Ahmad Humayun, David Tsai, and James M Rehg · 2013
Earlier work this paper cites.
Segmentation of moving objects by long term video analysis
Peter Ochs, Jitendra Malik, and Thomas Brox · 2013
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Video segmentation by non-local consensus voting
Alon Faktor and Michal Irani · 2014
Earlier work this paper cites.
Motion trajectory segmentation via minimum cost multicuts
Margret Keuper, Bjoern Andres, and Thomas Brox · 2015
Earlier work this paper cites.
N. Mayer, E. Ilg, P. Häusser, P. Fischer, D. Cremers, A. Dosovitskiy, and T. Brox · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
One-shot video object segmentation
Sergi Caelles, Kevis-Kokitsi Maninis, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2017
Earlier work this paper cites.
Fusionseg: Learning to combine motion and appearance for fully automatic segmentation of generic objects in videos
Suyog Dutt Jain, Bo Xiong, and Kristen Grauman · 2017
Earlier work this paper cites.
Maskrnn: Instance level video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander Schwing · 2017
Earlier work this paper cites.
Learning video object segmentation from static images
Federico Perazzi, Anna Khoreva, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Earlier work this paper cites.
Learning motion patterns in videos
Pavel Tokmakov, Karteek Alahari, and Cordelia Schmid · 2017
Earlier work this paper cites.
Learning video object segmentation with visual memory
Pavel Tokmakov, Karteek Alahari, and Cordelia Schmid · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Saliency-aware video object segmentation
Wenguan Wang, Jianbing Shen, Ruigang Yang, and Fatih Porikli · 2017
Earlier work this paper cites.
Cnn in mrf: Video object segmentation via inference in a cnn-based higher-order spatio-temporal mrf
Linchao Bao, Baoyuan Wu, and Wei Liu · 2018
Earlier work this paper cites.
Videomatch: Matching based video object segmentation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G Schwing · 2018
Earlier work this paper cites.
Sequential attend, infer, repeat: Generative modelling of moving objects
Adam Kosiorek, Hyunjik Kim, Yee Whye Teh, and Ingmar Posner · 2018
Earlier work this paper cites.
Video object segmentation with joint re-identification and attention-aware mask propagation
Xiaoxiao Li and Chen Change Loy · 2018
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2018
Cited alongside, same era.
Video object segmentation without temporal information
K-K Maninis, Sergi Caelles, Yuhua Chen, Jordi Pont-Tuset, Laura Leal-Taixé, Daniel Cremers, and Luc Van Gool · 2018
Cited alongside, same era.
Pwc-net: Cnns for optical flow using pyramid, warping, and cost volume
Deqing Sun, Xiaodong Yang, Ming-Yu Liu, and Jan Kautz · 2018
Cited alongside, same era.
Tracking emerges by colorizing videos
Carl Vondrick, Abhinav Shrivastava, Alireza Fathi, Sergio Guadarrama, and Kevin Murphy · 2018
Cited alongside, same era.
Mast: A memory-augmented self-supervised tracker
Zihang Lai, Erika Lu, and Weidi Xie · 2020
Later among the works it cites.
Space: Unsupervised object-oriented scene representation via spatial attention and decomposition
Zhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri, Weihao Sun, Gautam Singh, Fei Deng, Jindong Jiang, and Sungjin Ahn · 2020
Later among the works it cites.
Object-centric learning with slot attention
Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran, Georg Heigold, Jakob Uszkoreit, Alexey Dosovitskiy, and Thomas Kipf · 2020
Later among the works it cites.
Learning video object segmentation from unlabeled videos
Xiankai Lu, Wenguan Wang, Jianbing Shen, Yu-Wing Tai, David J. Crandall, and Steven C. H. Hoi · 2020
Later among the works it cites.
Raft: Recurrent all-pairs field transforms for optical flow
Zachary Teed and Jia Deng · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Christopher P Burgess, Loic Matthey, Nicholas Watters, Rishabh Kabra, Irina Higgins, Matt Botvinick, and Alexander Lerchner · 2019
Cited alongside, same era.
Spatially invariant unsupervised object detection with convolutional neural networks
Eric Crawford and Joelle Pineau · 2019
Cited alongside, same era.
Genesis: Generative scene inference and sampling with object-centric latent representations
Martin Engelcke, Adam R Kosiorek, Oiwi Parker Jones, and Ingmar Posner · 2019
Cited alongside, same era.
Shifting more attention to video salient object detection
Deng-Ping Fan, Wenguan Wang, Ming-Ming Cheng, and Jianbing Shen · 2019
Cited alongside, same era.
Multi-object representation learning with iterative variational inference
Klaus Greff, Raphaël Lopez Kaufman, Rishabh Kabra, Nick Watters, Christopher Burgess, Daniel Zoran, Loic Matthey, Matthew Botvinick, and Alexander Lerchner · 2019
Cited alongside, same era.
Scalor: Generative world models with scalable object representations
Jindong Jiang, Sepehr Janghorbani, Gerard De Melo, and Sungjin Ahn · 2019
Cited alongside, same era.
A generative appearance model for end-to-end video object segmentation
Joakim Johnander, Martin Danelljan, Emil Brissman, Fahad Shahbaz Khan, and Michael Felsberg · 2019
Cited alongside, same era.
Polina Zablotskaia, Edoardo A Dominici, Leonid Sigal, and Andreas M Lehrmann · 2020
Later among the works it cites.
Motion-attentive transition for zero-shot video object segmentation
Tianfei Zhou, Shunzhou Wang, Yi Zhou, Yazhou Yao, Jianwu Li, and Ling Shao · 2020
Later among the works it cites.
Self-supervision by prediction for object discovery in videos
Beril Besbinar and Pascal Frossard · 2021
Later among the works it cites.
Swin-unet: Unet-like pure transformer for medical image segmentation
Hu Cao, Yueyue Wang, Joy Chen, Dongsheng Jiang, Xiaopeng Zhang, Qi Tian, and Manning Wang · 2021
Later among the works it cites.
Emerging properties in self-supervised vision transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin · 2021
Later among the works it cites.
Efficient iterative amortized inference for learning symmetric and disentangled multi-object representations
Patrick Emami, Pan He, Sanjay Ranka, and Anand Rangarajan · 2021
Later among the works it cites.
Genesis-v2: Inferring unordered object representations without iterative refinement
Martin Engelcke, Oiwi Parker Jones, and Ingmar Posner · 2021
Later among the works it cites.
Simone: View-invariant, temporally-abstracted object representations via unsupervised video decomposition
Rishabh Kabra, Daniel Zoran, Goker Erdogan, Loic Matthey, Antonia Creswell, Matt Botvinick, Alexander Lerchner, and Chris Burgess · 2021
Later among the works it cites.
Segmenting invisible moving objects
Hala Lamdouar, Weidi Xie, and Andrew Zisserman · 2021
Later among the works it cites.
The emergence of objectness: Learning zero-shot segmentation from videos
Runtao Liu, Zhirong Wu, Stella Yu, and Stephen Lin · 2021
Later among the works it cites.
Gatsbi: Generative agent-centric spatio-temporal object interaction
Cheol-Hui Min, Jinseok Bae, Junho Lee, and Young Min Kim · 2021
Later among the works it cites.
Rethinking self-supervised correspondence learning: A video frame-level similarity perspective
Jiarui Xu and Xiaolong Wang · 2021
Later among the works it cites.
Self-supervised video object segmentation by motion grouping
Charig Yang, Hala Lamdouar, Erika Lu, Andrew Zisserman, and Weidi Xie · 2021
Later among the works it cites.
Dystab: Unsupervised object segmentation via dynamic-static bootstrapping
Yanchao Yang, Brian Lai, and Stefano Soatto · 2021
Later among the works it cites.
Discovering objects that can move
Zhipeng Bao, Pavel Tokmakov, Allan Jabri, Yu-Xiong Wang, Adrien Gaidon, and Martial Hebert · 2022
Closest in time.
Guess what moves: Unsupervised video and image segmentation by anticipating motion
Subhabrata Choudhury, Laurynas Karazija, Iro Laina, Andrea Vedaldi, and Christian Rupprecht · 2022
Closest in time.
Conditional object-centric learning from video
Thomas Kipf, Gamaleldin Fathy Elsayed, Aravindh Mahendran, Austin Stone, Sara Sabour, Georg Heigold, Rico Jonschkowski, Alexey Dosovitskiy, and Klaus Greff · 2022
Closest in time.
D2conv3d: Dynamic dilated convolutions for object segmentation in videos
Christian Schmidt, Ali Athar, Sabarinath Mahadevan, and Bastian Leibe · 2022
Closest in time.
Segmenting moving objects via an object-centric layered representation
Junyu Xie, Weidi Xie, and Andrew Zisserman · 2022
Closest in time.
Deformable sprites for unsupervised video decomposition
Vickie Ye, Zhengqi Li, Richard Tucker, Angjoo Kanazawa, and Noah Snavely · 2022
Closest in time.