Fetching the paper…
Reading the bibliography…
Visual scenes are extremely rich in diversity, not only because there are infinite combinations of objects and background, but also because the observations of the same scene may vary greatly with the change of viewpoints.
MONet: Unsupervised scene decomposition and representation
Burgess, C. P.; Matthey, L.; Watters, N.; Kabra, R.; Higgins, I.; Botvinick, M.; and Lerchner, A. 2019 · 1901
Earlier work this paper cites.
Mental rotation of three-dimensional objects
Shepard, R.; and Metzler, J. 1971 · 1971
Earlier work this paper cites.
Vision: A computational investigation into the human representation and processing of visual information
Marr, D. 1982 · 1982
Earlier work this paper cites.
Comparing partitions
Hubert, L.; and Arabie, P. 1985 · 1985
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Fodor, J.; and Pylyshyn, Z. 1988 · 1988
Earlier work this paper cites.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson, J.; Hariharan, B.; van der Maaten, L.; Fei-Fei, L.; Zitnick, C. L.; and Girshick, R. B. 2017 · 1997
Earlier work this paper cites.
The neuropsychology of object constancy
Turnbull, O.; Carey, D.; and McCarthy, R. 1997 · 1997
Earlier work this paper cites.
ROOTS: Object-centric representation and rendering of 3D scenes
Chen, C.; Deng, F.; and Ahn, S. 2020 · 2006
Earlier work this paper cites.
How infants learn about the visual world
Johnson, S. 2010 · 2010
Earlier work this paper cites.
Information theoretic measures for clusterings comparison: Variants, properties, normalization and correction for chance
Nguyen, X.; Epps, J.; and Bailey, J. 2010 · 2010
Earlier work this paper cites.
Fixed-form variational posterior approximation through stochastic linear regression
Salimans, T.; and Knowles, D. A. 2013 · 2013
Earlier work this paper cites.
Bringing semantics into focus using visual abstraction
Zitnick, C. L.; and Parikh, D. 2013 · 2013
Earlier work this paper cites.
Auto-encoding variational Bayes
Kingma, D. P.; and Welling, M. 2014 · 2014
Earlier work this paper cites.
Neural variational inference and learning in belief networks
Mnih, A.; and Gregor, K. 2014 · 2014
Earlier work this paper cites.
Attend, infer, repeat: Fast scene understanding with generative models
Eslami, S.; Heess, N.; Weber, T.; Tassa, Y.; Szepesvari, D.; Kavukcuoglu, K.; and Hinton, G. E. 2016 · 2016
Earlier work this paper cites.
Efficient inference in occlusion-aware generative models of images
Huang, J.; and Murphy, K. 2016 · 2016
Cited alongside, same era.
Neural Expectation Maximization
Greff, K.; van Steenkiste, S.; and Schmidhuber, J. 2017 · 2017
Cited alongside, same era.
Categorical reparameterization with Gumbel-softmax
Jang, E.; Gu, S.; and Poole, B. 2017 · 2017
Cited alongside, same era.
Building machines that learn and think like people
Lake, B.; Ullman, T. D.; Tenenbaum, J.; and Gershman, S. 2017 · 2017
Cited alongside, same era.
The concrete distribution: A continuous relaxation of discrete random variables
Maddison, C. J.; Mnih, A.; and Teh, Y. 2017 · 2017
Cited alongside, same era.
dSprites: Disentanglement testing Sprites dataset
Matthey, L.; Higgins, I.; Hassabis, D.; and Lerchner, A. 2017 · 2017
Cited alongside, same era.
Amodal instance segmentation with KINS dataset
Qi, L.; Jiang, L.; Liu, S.; Shen, X.; and Jia, J. 2019 · 2019
Later among the works it cites.
R-SQAIR: Relational sequential attend, infer, repeat
Stanic, A.; and Schmidhuber, J. 2019 · 2019
Later among the works it cites.
Exploiting spatial invariance for scalable unsupervised object tracking
Crawford, E.; and Pineau, J. 2020 · 2020
Later among the works it cites.
GENESIS: Generative scene inference and sampling with object-centric latent representations
Engelcke, M.; Kosiorek, A. R.; Jones, O. P.; and Posner, I. 2020 · 2020
Later among the works it cites.
Generative neurosymbolic machines
Jiang, J.; and Ahn, S.-J. 2020 · 2020
Later among the works it cites.
SCALOR: Generative world models with scalable object representations
Jiang, J.; Janghorbani, S.; de Melo, G.; and Ahn, S. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural scene representation and rendering
Eslami, S.; Rezende, D. J.; Besse, F.; Viola, F.; Morcos, A. S.; Garnelo, M.; Ruderman, A.; Rusu, A. A.; Danihelka, I.; Gregor, K.; Reichert, D. P.; Buesing, L.; Weber, T.; Vinyals, O.; Rosenbaum, D.; Rabinowitz, N. C.; King, H.; Hillier, C.; Botvinick, M.; Wierstra, D.; Kavukcuoglu, K.; and Hassabis, D. 2018 · 2018
Cited alongside, same era.
Sequential attend, infer, repeat: Generative modelling of moving objects
Kosiorek, A. R.; Kim, H.; Posner, I.; and Teh, Y. 2018 · 2018
Cited alongside, same era.
Iterative amortized inference
Marino, J.; Yue, Y.; and Mandt, S. 2018 · 2018
Cited alongside, same era.
Relational neural expectation maximization: Unsupervised discovery of objects and their interactions
van Steenkiste, S.; Chang, M.; Greff, K.; and Schmidhuber, J. 2018 · 2018
Cited alongside, same era.
Spatially invariant unsupervised object detection with convolutional neural networks
Crawford, E.; and Pineau, J. 2019 · 2019
Cited alongside, same era.
Multi-object representation learning with iterative variational inference
Greff, K.; Kaufman, R. L.; Kabra, R.; Watters, N.; Burgess, C. P.; Zoran, D.; Matthey, L.; Botvinick, M.; and Lerchner, A. 2019 · 2019
Cited alongside, same era.
Learning object-centric representations of multi-object scenes from multiple views
Li, N.; Eastwood, C.; and Fisher, R. B. 2020 · 2020
Later among the works it cites.
SPACE: Unsupervised object-oriented scene representation via spatial attention and decomposition
Lin, Z.; Wu, Y.-F.; Peri, S.; Sun, W.; Singh, G.; Deng, F.; Jiang, J.; and Ahn, S. 2020 · 2020
Later among the works it cites.
Object-centric learning with slot attention
Locatello, F.; Weissenborn, D.; Unterthiner, T.; Mahendran, A.; Heigold, G.; Uszkoreit, J.; Dosovitskiy, A.; and Kipf, T. 2020 · 2020
Later among the works it cites.
Entity abstraction in visual model-based reinforcement learning
Veerapaneni, R.; Co-Reyes, J. D.; Chang, M.; Janner, M.; Finn, C.; Wu, J.; Tenenbaum, J.; and Levine, S. 2020 · 2020
Later among the works it cites.
Efficient iterative amortized inference for learning symmetric and disentangled multi-object representations
Emami, P.; He, P.; Ranka, S.; and Rangarajan, A. 2021 · 2021
Closest in time.
Benchmarking unsupervised object representations for video sequences
Weis, M. A.; Chitta, K.; Sharma, Y.; Brendel, W.; Bethge, M.; Geiger, A.; and Ecker, A. S. 2021 · 2021
Closest in time.
Knowledge-Guided Object Discovery with Acquired Deep Impressions
Yuan, J.; Li, B.; and Xue, X. 2021 · 2021
Closest in time.
PROVIDE: A probabilistic framework for unsupervised video decomposition
Zablotskaia, P.; Dominici, E. A.; Sigal, L.; and Lehrmann, A. M. 2021 · 2021
Closest in time.