Fetching the paper…
Reading the bibliography…
The goal of this paper is to discover, segment, and track independently moving objects in complex visual scenes.
Schneider, G.E.: Two visual systems. Science 163
1969
Earlier work this paper cites.
Goodale, M.A., Milner, A.: Separate visual pathways for perception and action. Trends in Neurosciences 15
1992
Earlier work this paper cites.
Norman, J.: Two visual systems and two theories of perception: An attempt to reconcile the constructivist and ecological approaches. The Behavioral and brain sciences 25
2002
Earlier work this paper cites.
Martin, D., Fowlkes, C., Malik, J.: Learning to detect natural image boundaries using local brightness, color, and texture cues. TPAMI (2004)
2004
Earlier work this paper cites.
T.Brox, J.Malik: Object segmentation by long term analysis of point trajectories. In: ECCV (2010)
2010
Earlier work this paper cites.
Ochs, P., Brox, T.: Object segmentation in video: A hierarchical variational approach for turning point trajectories into dense regions. In: ICCV (2011)
2011
Earlier work this paper cites.
Li, F., Kim, T., Humayun, A., Tsai, D., Rehg, J.M.: Video segmentation by tracking many figure-ground segments. In: ICCV (2013)
2013
Earlier work this paper cites.
Papazoglou, A., Ferrari, V.: Fast object segmentation in unconstrained video. In: ICCV (2013)
2013
Earlier work this paper cites.
Raptis, M., Sigal, L.: Poselet key-framing: A model for human activity recognition. In: CVPR (2013)
2013
Earlier work this paper cites.
Faktor, A., Irani, M.: Video segmentation by non-local consensus voting. In: BMVC (2014)
2014
Earlier work this paper cites.
Ochs, P., Malik, J., Brox, T.: Segmentation of moving objects by long term video analysis. TPAMI (2014)
2014
Earlier work this paper cites.
Simonyan, K., Zisserman, A.: Two-stream convolutional networks for action recognition in videos. In: NeurIPS (2014)
2014
Earlier work this paper cites.
Keuper, M., Andres, B., Brox, T.: Motion trajectory segmentation via minimum cost multicuts. In: ICCV (2015)
2015
Earlier work this paper cites.
Keuper, M., Andres, B., Brox, T.: Motion trajectory segmentation via minimum cost multicuts. In: ICCV (2015)
2015
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A method for stochastic optimization. In: ICLR (2015)
2015
Earlier work this paper cites.
Bideau, P., Learned-Miller, E.: It’s moving! a probabilistic model for causal motion segmentation in moving camera videos. In: ECCV (2016)
2016
Earlier work this paper cites.
Eslami, S.M.A., Heess, N., Weber, T., Tassa, Y., Szepesvari, D., kavukcuoglu, k., Hinton, G.E.: Attend, infer, repeat: Fast scene understanding with generative models. In: NeurIPS (2016)
2016
Earlier work this paper cites.
Feichtenhofer, C., Pinz, A., Zisserman, A.: Convolutional two-stream network fusion for video action recognition. In: CVPR (2016)
2016
Earlier work this paper cites.
Greff, K., Rasmus, A., Berglund, M., Hao, T., Valpola, H., Schmidhuber, J.: Tagger: Deep unsupervised perceptual grouping. In: NeurIPS (2016)
2016
Earlier work this paper cites.
Perazzi, F., Pont-Tuset, J., McWilliams, B., Van Gool, L., Gross, M., Sorkine-Hornung, A.: A benchmark dataset and evaluation methodology for video object segmentation. In: CVPR (2016)
2016
Earlier work this paper cites.
Zhu, W., Hu, J., Sun, G., Cao, X., Qiao, Y.: A key volume mining deep framework for action recognition. In: CVPR (2016)
2016
Earlier work this paper cites.
He, K., Gkioxari, G., Dollar, P., Girshick, R.: Mask r-cnn. In: ICCV (2017)
2017
Earlier work this paper cites.
Jain, S.D., Xiong, B., Grauman, K.: Fusionseg: Learning to combine motion and appearance for fully automatic segmentation of generic objects in videos. In: CVPR (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Tokmakov, P., Alahari, K., Schmid, C.: Learning motion patterns in videos. In: CVPR (2017)
2017
Earlier work this paper cites.
Wang, W., Shen, J., Yang, R., Porikli, F.: Saliency-aware video object segmentation. IEEE Transactions on Pattern Analysis and Machine Intelligence 40
2017
Earlier work this paper cites.
Bideau, P., RoyChowdhury, A., Menon, R.R., Learned-Miller, E.: The best of both worlds: Combining cnns and geometric constraints for hierarchical motion segmentation. In: CVPR (2018)
2018
Earlier work this paper cites.
Kosiorek, A., Kim, H., Teh, Y.W., Posner, I.: Sequential attend, infer, repeat: Generative modelling of moving objects. In: NeurIPS (2018)
2018
Earlier work this paper cites.
Li, S., Seybold, B., Vorobyov, A., Fathi, A., Huang, Q., Kuo, C.C.J.: Instance embedding transfer to unsupervised video object segmentation. In: CVPR (2018)
2018
Earlier work this paper cites.
Ma, C.Y., Chen, M.H., Kira, Z., AlRegib, G.: Ts-lstm and temporal-inception: Exploiting spatiotemporal dynamics for activity recognition. Signal Processing: Image Communication (2018)
2018
Earlier work this paper cites.
Mahendran, A., Thewlis, J., Vedaldi, A.: Self-supervised segmentation by grouping optical-flow. In: ECCV (2018)
2018
Earlier work this paper cites.
Song, H., Wang, W., Zhao, S., Shen, J., Lam, K.M.: Pyramid dilated deeper convlstm for video salient object detection. In: ECCV (2018)
2018
Earlier work this paper cites.
Vondrick, C., Shrivastava, A., Fathi, A., Guadarrama, S., Murphy, K.: Tracking emerges by colorizing videos. In: ECCV (2018)
2018
Earlier work this paper cites.
Xu, N., Yang, L., Fan, Y., Yue, D., Liang, Y., Yang, J., Huang, T.S.: Youtube-vos: A large-scale video object segmentation benchmark. CoRR (2018)
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Dave, A., Tokmakov, P., Ramanan, D.: Towards segmenting anything that moves. ICCV workshops (2019)
2019
Earlier work this paper cites.
Feichtenhofer, C., Fan, H., Malik, J., He, K.: Slowfast networks for video recognition. In: ICCV (2019)
2019
Earlier work this paper cites.
Greff, K., Kaufman, R.L., Kabra, R., Watters, N., Burgess, C., Zoran, D., Matthey, L., Botvinick, M., Lerchner, A.: Multi-object representation learning with iterative variational inference. In: ICML (2019)
2019
Cited alongside, same era.
Griffin, B.A., Corso, J.J.: Bubblenets: Learning to select the guidance frame in video object segmentation by deep sorting frames. In: CVPR (2019)
2019
Cited alongside, same era.
He, Z., Li, J., Liu, D., He, H., Barber, D.: Tracking by animation: Unsupervised learning of multi-object attentive trackers. In: CVPR (2019)
2019
Cited alongside, same era.
Lai, Z., Xie, W.: Self-supervised learning for video correspondence flow. In: BMVC (2019)
2019
Cited alongside, same era.
Li, X., Liu, S., De Mello, S., Wang, X., Kautz, J., Yang, M.H.: Joint-task self-supervised learning for temporal correspondence. In: NeurIPS (2019)
2019
Yin, Z., Zheng, J., Luo, W., Qian, S., Zhang, H., Gao, S.: Learning to recommend frame for interactive video object segmentation in the wild. In: CVPR (2021)
2021
Later among the works it cites.
Zhang, K., Zhao, Z., Liu, D., Liu, Q., Liu, B.: Deep transport network for unsupervised video object segmentation. In: ICCV (2021)
2021
Later among the works it cites.
Bao, Z., Tokmakov, P., Jabri, A., Wang, Y.X., Gaidon, A., Hebert, M.: Discovering objects that can move. In: CVPR (2022)
2022
Later among the works it cites.
Buch, S., Eyzaguirre, C., Gaidon, A., Wu, J., Fei-Fei, L., Niebles, J.C.: Revisiting the ”video” in video-language understanding. In: CVPR (2022)
2022
Later among the works it cites.
Choudhury, S., Karazija, L., Laina, I., Vedaldi, A., Rupprecht, C.: Guess What Moves: Unsupervised Video and Image Segmentation by Anticipating Motion. In: BMVC (2022)
2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Lu, X., Wang, W., Ma, C., Shen, J., Shao, L., Porikli, F.: See more, know more: Unsupervised video object segmentation with co-attention siamese networks. In: CVPR (2019)
2019
Cited alongside, same era.
Tokmakov, P., Schmid, C., Alahari, K.: Learning to segment moving objects. IJCV (2019)
2019
Cited alongside, same era.
Tokmakov, P., Schmid, C., Alahari, K.: Learning to segment moving objects. IJCV (2019)
2019
Cited alongside, same era.
Veerapaneni, R., Co-Reyes, J.D., Chang, M., Janner, M., Finn, C., Wu, J., Tenenbaum, J., Levine, S.: Entity abstraction in visual model-based reinforcement learning. In: CoRL (2019)
2019
Cited alongside, same era.
Ventura, C., Bellver, M., Girbau, A., Salvador, A., Marques, F., Giro-i Nieto, X.: RVOS: End-to-end recurrent network for video object segmentation. In: CVPR (2019)
2019
Cited alongside, same era.
Wang, X., Jabri, A., Efros, A.A.: Learning correspondence from the cycle-consistency of time. In: CVPR (2019)
2019
Cited alongside, same era.
Xie, C., Xiang, Y., Harchaoui, Z., Fox, D.: Object discovery in videos as foreground motion clustering. In: CVPR (2019)
2019
Cited alongside, same era.
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
He, K., Chen, X., Xie, S., Li, Y., Dollár, P., Girshick, R.: Masked autoencoders are scalable vision learners. In: CVPR (2022)
2022
Later among the works it cites.
Kipf, T., Elsayed, G.F., Mahendran, A., Stone, A., Sabour, S., Heigold, G., Jonschkowski, R., Dosovitskiy, A., Greff, K.: Conditional Object-Centric Learning from Video. In: ICLR (2022)
2022
Later among the works it cites.
Melas-Kyriazi, L., Laina, I., Rupprecht, C., Vedaldi, A.: Deep spectral methods: A surprisingly strong baseline for unsupervised semantic segmentation and localization. In: CVPR (2022)
2022
Later among the works it cites.
Miao, B., Bennamoun, M., Gao, Y., Mian, A.: Self-supervised video object segmentation by motion-aware mask propagation. In: ICME (2022)
2022
Later among the works it cites.
Shin, G., Albanie, S., Wie, W.: Unsupervised salient object detection with spectral cluster voting. In: CVPR workshops (2022)
2022
Later among the works it cites.
Wang, Y., Shen, X., Hu, S.X., Yuan, Y., Crowley, J.L., Vaufreydaz, D.: Self-supervised transformers for unsupervised object discovery using normalized cut. In: CVPR (2022)
2022
Later among the works it cites.
Xie, J., Xie, W., Zisserman, A.: Segmenting moving objects via an object-centric layered representation. In: NeurIPS (2022)
2022
Later among the works it cites.
Ye, V., Li, Z., Tucker, R., Kanazawa, A., Snavely, N.: Deformable sprites for unsupervised video decomposition. In: CVPR (2022)
2022
Later among the works it cites.
Yu, Y., Yuan, J., Mittal, G., Fuxin, L., Chen, M.: Batman: Bilateral attention transformer in motion-appearance neighboring space for video object segmentation. In: ECCV (2022)
2022
Later among the works it cites.
Zhao, S., Zhu, L., Wang, X., Yang, Y.: Centerclip: Token clustering for efficient text-video retrieval. In: The 45th International ACM SIGIR Conference on Research and Development in Information Retrieval (2022)
2022
Later among the works it cites.
Aydemir, G., Xie, W., Güney, F.: Self-supervised object-centric learning for videos. In: NeurIPS (2023)
2023
Closest in time.
Bekuzarov, M., Bermudez, A., Lee, J.Y., Li, H.: Xmem++: Production-level video segmentation from few annotated frames. In: ICCV (2023)
2023
Closest in time.
Cho, S., Lee, M., Lee, S., Park, C., Kim, D., Lee, S.: Treating motion as option to reduce motion dependency in unsupervised video object segmentation. In: WACV (2023)
2023
Closest in time.
2023
Closest in time.
Fan, K., Bai, Z., Xiao, T., Zietlow, D., Horn, M., Zhao, Z., Simon-Gabriel, C.J., Shou, M.Z., Locatello, F., Schiele, B., Brox, T., Zhang, Z., Fu, Y., He, T.: Unsupervised open-vocabulary object localization in videos. In: ICCV (2023)
2023
Closest in time.
Ke, L., Ye, M., Danelljan, M., Liu, Y., Tai, Y.W., Tang, C.K., Yu, F.: Segment anything in high quality. In: NeurIPS (2023)
2023
Closest in time.
2023
Closest in time.
Lee, M., Cho, S., Lee, S., Park, C., Lee, S.: Unsupervised video object segmentation via prototype memory network. In: WACV (2023)
2023
Closest in time.
2023
Closest in time.
Meunier, E., Badoual, A., Bouthemy, P.: Em-driven unsupervised learning for efficient motion segmentation. TPAMI (2023)
2023
Closest in time.
Meunier, E., Bouthemy, P.: Unsupervised space-time network for temporally-consistent segmentation of multiple motions. In: CVPR (2023)
2023
Closest in time.
2023
Closest in time.
Ponimatkin, G., Samet, N., Xiao, Y., Du, Y., Marlet, R., Lepetit, V.: A simple and powerful global optimization for unsupervised video object segmentation. In: WACV (2023)
2023
Closest in time.
Ponimatkin, G., Samet, N., Xiao, Y., Du, Y., Marlet, R., Lepetit, V.: A simple and powerful global optimization for unsupervised video object segmentation. In: WACV (2023)
2023
Closest in time.
Safadoust, S., Güney, F.: Multi-object discovery by low-dimensional object motion. In: ICCV (2023)
2023
Closest in time.
Seitzer, M., Horn, M., Zadaianchuk, A., Zietlow, D., Xiao, T., Simon-Gabriel, C.J., He, T., Zhang, Z., Schölkopf, B., Brox, T., Locatello, F.: Bridging the gap to real-world object-centric learning. In: ICLR (2023)
2023
Closest in time.
Singh, S., Deshmukh, S., Sarkar, M., Jain, R., Hemani, M., Krishnamurthy, B.: Fodvid: Flow-guided object discovery in videos. In: CVPR workshops (2023)
2023
Closest in time.
Singh, S., Deshmukh, S., Sarkar, M., Krishnamurthy, B.: Locate: Self-supervised object discovery via flow-guided graph-cut and bootstrapped self-training. In: BMVC (2023)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Delatolas, T., Kalogeiton, V., Papadopoulos, D.P.: Learning the what and how of annotation in video object segmentation. In: WACV (2024)
2024
Closest in time.