Fetching the paper…
Reading the bibliography…
We investigate the emergence of objects in visual perception in the absence of any semantic annotation.
Foundations of cyclopean perception
Bela Julesz · 1971
Earlier work this paper cites.
The ecological approach to the visual perception of pictures
James J Gibson · 1978
Earlier work this paper cites.
Determining optical flow
Berthold KP Horn and Brian G Schunck · 1981
Earlier work this paper cites.
Representing moving images with layers
John YA Wang and Edward H Adelson · 1994
Earlier work this paper cites.
A unified mixture framework for motion segmentation: Incorporating spatial coherence and estimating the number of models
Yair Weiss and Edward H Adelson · 1996
Earlier work this paper cites.
Hierarchical estimation and segmentation of dense motion fields
Etienne Mémin and Patrick Pérez · 2002
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Layered image motion with explicit occlusions, temporal consistency, and depth ordering
Deqing Sun, Erik Sudderth, and Michael Black · 2010
Earlier work this paper cites.
Detachable object detection: Segmentation and depth ordering from short-baseline video
Alper Ayvaci and Stefano Soatto · 2011
Earlier work this paper cites.
A naturalistic open source movie for optical flow evaluation
D. J. Butler, J. Wulff, G. B. Stanley, and M. J. Black · 2012
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Motion coherent tracking using multi-label mrf optimization
David Tsai, Matthew Flagg, Atsushi Nakazawa, and James M Rehg · 2012
Earlier work this paper cites.
Segmentation of moving objects by long term video analysis
Peter Ochs, Jitendra Malik, and Thomas Brox · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Flownet: Learning optical flow with convolutional networks
Alexey Dosovitskiy, Philipp Fischer, Eddy Ilg, Philip Hausser, Caner Hazirbas, Vladimir Golkov, Patrick Van Der Smagt, Daniel Cremers, and Thomas Brox · 2015
Earlier work this paper cites.
Motion trajectory segmentation via minimum cost multicuts
Margret Keuper, Bjoern Andres, and Thomas Brox · 2015
Earlier work this paper cites.
Learning deconvolution network for semantic segmentation
Hyeonwoo Noh, Seunghoon Hong, and Bohyung Han · 2015
Earlier work this paper cites.
Generalized regressive motion: a visual cue to collision
Krzysztof Chalupka, Michael Dickinson, and Pietro Perona · 2016
Earlier work this paper cites.
Attend, infer, repeat: Fast scene understanding with generative models
SM Eslami, Nicolas Heess, Theophane Weber, Yuval Tassa, David Szepesvari, Geoffrey E Hinton, et al · 2016
Earlier work this paper cites.
A multi-cut formulation for joint segmentation and tracking of multiple objects
Margret Keuper, Siyu Tang, Yu Zhongjie, Bjoern Andres, Thomas Brox, and Bernt Schiele · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille · 2017
Earlier work this paper cites.
Neural expectation maximization
Klaus Greff, Sjoerd Van Steenkiste, and Jürgen Schmidhuber · 2017
Earlier work this paper cites.
Suyog Jain, Bo Xiong, and Kristen Grauman · 2017
Earlier work this paper cites.
Primary object segmentation in videos based on region augmentation and reduction
Yeong Jun Koh and Chang-Su Kim · 2017
Earlier work this paper cites.
The 2017 davis challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alexander Sorkine-Hornung, and Luc Van Gool · 2017
Cited alongside, same era.
Learning video object segmentation with visual memory
P. Tokmakov, K. Alahari, and C. Schmid · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Pyramid scene parsing network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia · 2017
Cited alongside, same era.
Unsupervised video object segmentation using motion saliency-guided spatio-temporal propagation
Yuan-Ting Hu, Jia-Bin Huang, and Alexander G Schwing · 2018
Cited alongside, same era.
Sequential attend, infer, repeat: Generative modelling of moving objects
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Illiterate dall-e learns to compose
Gautam Singh, Fei Deng, and Sungjin Ahn · 2021
Later among the works it cites.
Smurf: Self-teaching multi-frame unsupervised raft with full-image warping
Austin Stone, Daniel Maurer, Alper Ayvaci, Anelia Angelova, and Rico Jonschkowski · 2021
Later among the works it cites.
Discovering objects that can move
Zhipeng Bao, Pavel Tokmakov, Allan Jabri, Yu-Xiong Wang, Adrien Gaidon, and Martial Hebert · 2022
Later among the works it cites.
Move: Unsupervised movable object segmentation and detection
Adam Bielski and Paolo Favaro · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam Kosiorek, Hyunjik Kim, Yee Whye Teh, and Ingmar Posner · 2018
Cited alongside, same era.
Extending layered models to 3d motion
Dong Lao and Ganesh Sundaramoorthi · 2018
Cited alongside, same era.
Pyramid dilated deeper convlstm for video salient object detection
Hongmei Song, Wenguan Wang, Sanyuan Zhao, Jianbing Shen, and Kin-Man Lam · 2018
Cited alongside, same era.
Youtube-vos: A large-scale video object segmentation benchmark
Ning Xu, Linjie Yang, Yuchen Fan, Dingcheng Yue, Yuchen Liang, Jianchao Yang, and Thomas S. Huang · 2018
Cited alongside, same era.
Conditional prior networks for optical flow
Yanchao Yang and Stefano Soatto · 2018
Cited alongside, same era.
Towards segmenting anything that moves
Achal Dave, Pavel Tokmakov, and Deva Ramanan · 2019
Cited alongside, same era.
Shifting more attention to video salient object detection
Deng-Ping Fan, Wenguan Wang, Ming-Ming Cheng, and Jianbing Shen · 2019
Cited alongside, same era.
Object representations as fixed points: Training iterative inference algorithms with implicit differentiation
Michael Chang, Thomas L Griffiths, and Sergey Levine · 2022
Later among the works it cites.
Motion-inductive self-supervised object discovery in videos
Shuangrui Ding, Weidi Xie, Yabo Chen, Rui Qian, Xiaopeng Zhang, Hongkai Xiong, and Qi Tian · 2022
Later among the works it cites.
Em-driven unsupervised learning for efficient motion segmentation
Etienne Meunier, Anaïs Badoual, and Patrick Bouthemy · 2022
Later among the works it cites.
Simple unsupervised object-centric learning for complex and naturalistic videos
Gautam Singh, Yi-Fu Wu, and Sungjin Ahn · 2022
Later among the works it cites.
Yangtao Wang, Xi Shen, Yuan Yuan, Yuming Du, Maomao Li, Shell Xu Hu, James L Crowley, and Dominique Vaufreydaz · 2022
Later among the works it cites.
Segmenting moving objects via an object-centric layered representation
Junyu Xie, Weidi Xie, and Andrew Zisserman · 2022
Later among the works it cites.
Deformable sprites for unsupervised video decomposition
Vickie Ye, Zhengqi Li, Richard Tucker, Angjoo Kanazawa, and Noah Snavely · 2022
Later among the works it cites.
Object discovery from motion-guided tokens
Zhipeng Bao, Pavel Tokmakov, Yu-Xiong Wang, Adrien Gaidon, and Martial Hebert · 2023
Closest in time.
MOSE: A new dataset for video object segmentation in complex scenes
Henghui Ding, Chang Liu, Shuting He, Xudong Jiang, Philip HS Torr, and Song Bai · 2023
Closest in time.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Closest in time.
Bootstrapping objectness from videos by relaxed common fate and visual grouping
Long Lian, Zhirong Wu, and Stella X Yu · 2023
Closest in time.
Unsupervised space-time network for temporally-consistent segmentation of multiple motions
Etienne Meunier and Patrick Bouthemy · 2023
Closest in time.
Dinov2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al · 2023
Closest in time.
A simple and powerful global optimization for unsupervised video object segmentation
Georgy Ponimatkin, Nermin Samet, Yang Xiao, Yuming Du, Renaud Marlet, and Vincent Lepetit · 2023
Closest in time.
Bridging the gap to real-world object-centric learning
Maximilian Seitzer, Max Horn, Andrii Zadaianchuk, Dominik Zietlow, Tianjun Xiao, Carl-Johann Simon-Gabriel, Tong He, Zheng Zhang, Bernhard Schölkopf, Thomas Brox, et al · 2023
Closest in time.
Locate: Self-supervised object discovery via flow-guided graph-cut and bootstrapped self-training
Silky Singh, Shripad Deshmukh, Mausoom Sarkar, and Balaji Krishnamurthy · 2023
Closest in time.
Cut and learn for unsupervised object detection and instance segmentation
Xudong Wang, Rohit Girdhar, Stella X Yu, and Ishan Misra · 2023
Closest in time.
Object-centric learning for real-world videos by predicting temporal feature similarities
Andrii Zadaianchuk, Maximilian Seitzer, and Georg Martius · 2023
Closest in time.
On the viability of monocular depth pre-training for semantic segmentation
Dong Lao, Alex Wong, Samuel Lu, and Stefano Soatto · 2024
Closest in time.
Sam 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, et al · 2024
Closest in time.