Fetching the paper…
Reading the bibliography…
We present ObSuRF, a method which turns a single image of a scene into a 3D model represented as a set of Neural Radiance Fields (NeRFs), with each NeRF corresponding to a different object.
Objective criteria for the evaluation of clustering methods
Rand, W. M · 1971
Earlier work this paper cites.
Pattern synthesis , volume 1 of Lectures in pattern theory
Grenander, U · 1976
Earlier work this paper cites.
Pattern analysis , volume 2 of Lectures in pattern theory
Grenander, U · 1978
Earlier work this paper cites.
Light reflection functions for simulation of clouds and dusty surfaces
Blinn, J. F · 1982
Earlier work this paper cites.
Ray tracing volume densities
Kajiya, J. and von Herzen, B · 1984
Earlier work this paper cites.
Compositing digital images
Porter, T. and Duff, T · 1984
Earlier work this paper cites.
Comparing partitions
Hubert, L. and Arabie, P · 1985
Earlier work this paper cites.
Statistical inference and simulation for spatial point processes
Møller, J. and Waagepetersen, R. P · 2003
Earlier work this paper cites.
Simulation as an engine of physical scene understanding
Battaglia, P. W., Hamrick, J. B., and Tenenbaum, J. B · 2013
Earlier work this paper cites.
ShapeNet: An Information-Rich 3D Model Repository
Chang, A. X., Funkhouser, T., Guibas, L., Hanrahan, P., Huang, Q., Li, Z., Savarese, S., Savva, M., Song, S., Su, H., Xiao, J., Yi, L., and Yu, F · 2015
Earlier work this paper cites.
Interaction networks for learning about objects, relations and physics
Battaglia, P. W., Pascanu, R., Lai, M., Rezende, D., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
Monocular 3d object detection for autonomous driving
Chen, X., Kundu, K., Zhang, Z., Ma, H., Fidler, S., and Urtasun, R · 2016
Earlier work this paper cites.
Attend, infer, repeat: Fast scene understanding with generative models
Eslami, S. M. A., Heess, N., Weber, T., Tassa, Y., Szepesvari, D., Kavukcuoglu, K., and Hinton, G. E · 2016
Earlier work this paper cites.
Tagger: Deep unsupervised perceptual grouping
Greff, K., Rasmus, A., Berglund, M., Hao, T., Valpola, H., and Schmidhuber, J · 2016
Earlier work this paper cites.
An overview of depth cameras and range scanners based on time-of-flight technologies
Horaud, R., Hansard, M., Evangelidis, G., and Clément, M · 2016
Earlier work this paper cites.
A learned representation for artistic style
Dumoulin, V., Shlens, J., and Kudlur, M · 2017
Earlier work this paper cites.
Neural expectation maximization
Greff, K., van Steenkiste, S., and Schmidhuber, J · 2017
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson, J., Hariharan, B., Van Der Maaten, L., Fei-Fei, L., Lawrence Zitnick, C., and Girshick, R · 2017
Earlier work this paper cites.
Pointnet: Deep learning on point sets for 3d classification and segmentation
Qi, C. R., Su, H., Mo, K., and Guibas, L. J · 2017
Earlier work this paper cites.
Vision-as-inverse-graphics: Obtaining a rich 3d explanation of a scene from a single image
Romaszko, L., Williams, C. K. I., Moreno, P., and Kohli, P · 2017
Earlier work this paper cites.
A simple neural network module for relational reasoning
Santoro, A., Raposo, D., Barrett, D. G., Malinowski, M., Pascanu, R., Battaglia, P., and Lillicrap, T · 2017
Earlier work this paper cites.
Relational inductive biases, deep learning, and graph networks
Battaglia, P. W., Hamrick, J. B., Bapst, V., Sanchez-Gonzalez, A., Zambaldi, V., Malinowski, M., Tacchetti, A., Raposo, D., Santoro, A., Faulkner, R., Gulcehre, C., Song, F., Ballard, A., Gilmer, J., Dahl, G., Vaswani, A., Allen, K., Nash, C., Langston, V., Dyer, C., Heess, N., Wierstra, D., Kohli, P., Botvinick, M., Vinyals, O., Li, Y., and Pascanu, R · 2018
Cited alongside, same era.
Pyramid stereo matching network
Chang, J.-R. and Chen, Y.-S · 2018
Cited alongside, same era.
Sequential attend, infer, repeat: Generative modelling of moving objects
Kosiorek, A., Kim, H., Teh, Y. W., and Posner, I · 2018
Cited alongside, same era.
On nesting monte carlo estimators
Rainforth, T., Cornish, R., Yang, H., and Warrington, A · 2018
Cited alongside, same era.
Relational neural expectation maximization: Unsupervised discovery of objects and their interactions
van Steenkiste, S., Chang, M., Greff, K., and Schmidhuber, J · 2018
Cited alongside, same era.
Graspnet-1billion: A large-scale benchmark for general object grasping
Fang, H.-S., Wang, C., Gou, M., and Lu, C · 2020
Later among the works it cites.
Multi-object representation learning with iterative variational inference
Greff, K., Kaufman, R. L., Kabra, R., Watters, N., Burgess, C., Zoran, D., Matthey, L., Botvinick, M., and Lerchner, A · 2020
Later among the works it cites.
Object-centric neural scene rendering
Guo, M., Fathi, A., Wu, J., and Funkhouser, T · 2020
Later among the works it cites.
Ucsg-net – unsupervised discovering of constructive solid geometry tree
Kania, K., Zięba, M., and Kajdanowicz, T · 2020
Later among the works it cites.
Conditional set generation with transformers
Kosiorek, A. R., Kim, H., and Rezende, D. J · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pixel2mesh: Generating 3d mesh models from single rgb images
Wang, N., Zhang, Y., Li, Z., Fu, Y., Liu, W., and Jiang, Y.-G · 2018
Cited alongside, same era.
Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes
Xiang, Y., Schmidt, T., Narayanan, V., and Fox, D · 2018
Cited alongside, same era.
Large scale GAN training for high fidelity natural image synthesis
Brock, A., Donahue, J., and Simonyan, K · 2019
Cited alongside, same era.
Monet: Unsupervised scene decomposition and representation
Burgess, C. P., Matthey, L., Watters, N., Kabra, R., Higgins, I., Botvinick, M., and Lerchner, A · 2019
Cited alongside, same era.
Bae-net: Branched autoencoder for shape co-segmentation
Chen, Z., Yin, K., Fisher, M., Chaudhuri, S., and Zhang, H · 2019
Cited alongside, same era.
Spatially invariant unsupervised object detection with convolutional neural networks
Crawford, E. and Pineau, J · 2019
Cited alongside, same era.
Multi-object datasets
Kabra, R., Burgess, C., Matthey, L., Kaufman, R. L., Greff, K., Reynolds, M., and Lerchner, A · 2019
Cited alongside, same era.
Perspective plane program induction from a single image
Li, Y., Mao, J., Zhang, X., Freeman, W. T., Tenenbaum, J. B., and Wu, J · 2020
Later among the works it cites.
Space: Unsupervised object-oriented scene representation via spatial attention and decomposition
Lin, Z., Wu, Y.-F., Peri, S. V., Sun, W., Singh, G., Deng, F., Jiang, J., and Ahn, S · 2020
Later among the works it cites.
Object-centric learning with slot attention
Locatello, F., Weissenborn, D., Unterthiner, T., Mahendran, A., Heigold, G., Uszkoreit, J., Dosovitskiy, A., and Kipf, T · 2020
Later among the works it cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall, B., Srinivasan, P. P., Tancik, M., Barron, J. T., Ramamoorthi, R., and Ng, R · 2020
Later among the works it cites.
Blockgan: Learning 3d object-aware scene representations from unlabelled images
Nguyen-Phuoc, T., Richardt, C., Mai, L., Yang, Y.-L., and Mitra, N · 2020
Later among the works it cites.
Giraffe: Representing scenes as compositional generative neural feature fields
Niemeyer, M. and Geiger, A · 2020
Later among the works it cites.
Generative adversarial set transformers
Stelzner, K., Kersting, K., and Kosiorek, A. R · 2020
Later among the works it cites.
Deepv2d: Video to depth with differentiable structure from motion
Teed, Z. and Deng, J · 2020
Later among the works it cites.
Grf: Learning a general radiance field for 3d scene representation and rendering
Trevithick, A. and Yang, B · 2020
Later among the works it cites.
Unmasking the inductive biases of unsupervised object representations for video sequences
Weis, M. A., Chitta, K., Sharma, Y., Brendel, W., Bethge, M., Geiger, A., and Ecker, A. S · 2020
Later among the works it cites.
Fspool: Learning set representations with featurewise sort pooling
Zhang, Y., Hare, J., and Prügel-Bennett, A · 2020
Later among the works it cites.
Semi-supervised learning of multi-object 3d scene representations
Elich, C., Oswald, M. R., Pollefeys, M., and Stueckler, J · 2021
Closest in time.
Nerf-vae: A geometry aware 3d scene generative model
Kosiorek, A. R., Strathmann, H., Zoran, D., Moreno, P., Schneider, R., Mokrá, S., and Rezende, D. J · 2021
Closest in time.
Nerf in the wild: Neural radiance fields for unconstrained photo collections
Martin-Brualla, R., Radwan, N., Sajjadi, M. S. M., Barron, J. T., Dosovitskiy, A., and Duckworth, D · 2021
Closest in time.
Neural scene graphs for dynamic scenes
Ost, J., Mannan, F., Thuerey, N., Knodt, J., and Heide, F · 2021
Closest in time.