Fetching the paper…
Reading the bibliography…
Scene synthesis is a challenging problem with several industrial applications.
Method for registration of 3-d shapes
P. J. Besl and N. D. McKay · 1992
Earlier work this paper cites.
Efficient variants of the icp algorithm
S. Rusinkiewicz and M. Levoy · 2001
Earlier work this paper cites.
Convex optimization
S. P. Boyd and L. Vandenberghe · 2004
Earlier work this paper cites.
Towards 3d object maps for autonomous household robots
R. B. Rusu, N. Blodow, Z. Marton, A. Soos, and M. Beetz · 2007
Earlier work this paper cites.
Probability and statistics
M. H. DeGroot and M. J. Schervish · 2012
Earlier work this paper cites.
Example-based synthesis of 3d object arrangements
M. Fisher, D. Ritchie, M. Savva, T. Funkhouser, and P. Hanrahan · 2012
Earlier work this paper cites.
Learning spatial knowledge for text to 3d scene generation
A. Chang, M. Savva, and C. D. Manning · 2014
Earlier work this paper cites.
Text to 3d scene generation with rich lexical grounding
A. Chang, W. Monroe, M. Savva, C. Potts, and C. D. Manning · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli · 2015
Earlier work this paper cites.
Sceneseer: 3d scene design with natural language
A. X. Chang, M. Eric, M. Savva, and C. D. Manning · 2017
Earlier work this paper cites.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner · 2017
Earlier work this paper cites.
Colored point cloud registration revisited
J. Park, Q.-Y. Zhou, and V. Koltun · 2017
Earlier work this paper cites.
Scenesuggest: Context-driven 3d scene design
M. Savva, A. X. Chang, and M. Agrawala · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Learning representations and generative models for 3d point clouds
P. Achlioptas, O. Diamanti, I. Mitliagkas, and L. Guibas · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Earlier work this paper cites.
Configurable 3d scene synthesis and 2d image rendering with per-pixel ground truth using stochastic grammars
C. Jiang, S. Qi, Y. Zhu, S. Huang, J. Lin, L.-F. Yu, D. Terzopoulos, and S.-C. Zhu · 2018
Earlier work this paper cites.
Language-driven synthesis of 3d scenes from scene databases
R. Ma, A. G. Patil, M. Fisher, M. Li, S. Pirk, B.-S. Hua, S.-K. Yeung, X. Tong, L. Guibas, and H. Zhang · 2018
Earlier work this paper cites.
Human-centric indoor scene synthesis using stochastic grammar
S. Qi, Y. Zhu, S. Huang, C. Jiang, and S.-C. Zhu · 2018
Earlier work this paper cites.
Deep convolutional priors for indoor scene synthesis
K. Wang, M. Savva, A. X. Chang, and D. Ritchie · 2018
Earlier work this paper cites.
Resolving 3d human pose ambiguities with 3d scene constraints
M. Hassan, V. Choutas, D. Tzionas, and M. J. Black · 2019
Earlier work this paper cites.
Grains: Generative recursive autoencoders for indoor scenes
M. Li, A. G. Patil, K. Xu, S. Chaudhuri, O. Khan, A. Shamir, C. Tu, B. Chen, D. Cohen-Or, and H. Zhang · 2019
Earlier work this paper cites.
Expressive body capture: 3d hands, face, and body from a single image
G. Pavlakos, V. Choutas, N. Ghorbani, T. Bolkart, A. A. Osman, D. Tzionas, and M. J. Black · 2019
Cited alongside, same era.
Fast and flexible indoor scene synthesis via deep convolutional generative models
D. Ritchie, K. Wang, and Y.-a. Lin · 2019
Cited alongside, same era.
Pointflow: 3d point cloud generation with continuous normalizing flows
G. Yang, X. Huang, Z. Hao, M.-Y. Liu, S. Belongie, and B. Hariharan · 2019
Cited alongside, same era.
Gspn: Generative shape proposal network for 3d instance segmentation in point cloud
L. Yi, W. Zhao, H. Wang, M. Sung, and L. J. Guibas · 2019
Cited alongside, same era.
Learning gradient fields for shape generation
R. Cai, G. Yang, H. Averbuch-Elor, Z. Hao, S. Belongie, N. Snavely, and B. Hariharan · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
3d shape generation and completion through point-voxel diffusion
L. Zhou, Y. Du, and J. Wu · 2021
Later among the works it cites.
Continuous scene representations for embodied ai
S. Y. Gadre, K. Ehsani, S. Song, and R. Mottaghi · 2022
Later among the works it cites.
J. Ho, T. Salimans, A. Gritsenko, W. Chan, M. Norouzi, and D. J. Fleet · 2022
Later among the works it cites.
Reconstruct from top view: A 3d lane detection approach based on geometry structure prior
C. Li, J. Shi, Y. Wang, and G. Cheng · 2022
Later among the works it cites.
Pose2room: understanding 3d scenes from human activities
Y. Nie, A. Dai, X. Han, and M. Nießner · 2022
Later among the works it cites.
Mm-diffusion: Learning multi-modal diffusion models for joint audio and video generation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Ho, A. Jain, and P. Abbeel · 2020
Cited alongside, same era.
Denoising diffusion implicit models
J. Song, C. Meng, and S. Ermon · 2020
Cited alongside, same era.
Improved techniques for training score-based generative models
Y. Song and S. Ermon · 2020
Cited alongside, same era.
Learning 3d semantic scene graphs from 3d indoor reconstructions
J. Wald, H. Dhamo, N. Navab, and F. Tombari · 2020
Cited alongside, same era.
Generating 3d people in scenes without people
Y. Zhang, M. Hassan, H. Neumann, M. J. Black, and S. Tang · 2020
Cited alongside, same era.
Graph-to-3d: End-to-end generation and manipulation of 3d scenes using scene graphs
H. Dhamo, F. Manhardt, N. Navab, and F. Tombari · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
P. Dhariwal and A. Nichol · 2021
Cited alongside, same era.
L. Ruan, Y. Ma, H. Yang, H. He, B. Liu, J. Fu, N. J. Yuan, Q. Jin, and B. Guo · 2022
Later among the works it cites.
Palette: Image-to-image diffusion models
C. Saharia, W. Chan, H. Chang, C. Lee, J. Ho, T. Salimans, D. Fleet, and M. Norouzi · 2022
Later among the works it cites.
G. Tevet, S. Raab, B. Gordon, Y. Shafir, D. Cohen-Or, and A. H. Bermano · 2022
Later among the works it cites.
Edge: Editable dance generation from music
J. Tseng, R. Castellon, and C. K. Liu · 2022
Later among the works it cites.
Humanise: Language-conditioned human motion generation in 3d scenes
Z. Wang, Y. Chen, T. Liu, Y. Zhu, W. Liang, and S. Huang · 2022
Later among the works it cites.
Scene synthesis from human motion
S. Ye, Y. Wang, J. Li, D. Park, C. K. Liu, H. Xu, and J. Wu · 2022
Later among the works it cites.
Open3d: A modern library for 3d data processing
Q.-Y. Zhou, J. Park, and V. Koltun · 2022
Later among the works it cites.
Diffusion-based generation, optimization, and planning in 3d scenes
S. Huang, Z. Wang, P. Li, B. Jia, T. Liu, Y. Zhu, W. Liang, and S.-C. Zhu · 2023
Closest in time.
Music-driven group choreography
N. Le, T. Pham, T. Do, E. Tjiputra, Q. D. Tran, and A. Nguyen · 2023
Closest in time.
More control for free! image synthesis with semantic diffusion guidance
X. Liu, D. H. Park, S. Azadi, G. Zhang, A. Chopikyan, Y. Hu, H. Shi, A. Rohrbach, and T. Darrell · 2023
Closest in time.
Unified multi-modal latent diffusion for joint subject and text conditional image generation
Y. Ma, H. Yang, W. Wang, J. Fu, and J. Liu · 2023
Closest in time.
Language-conditioned affordance-pose detection in 3d point clouds
T. Nguyen, M. N. Vu, B. Huang, T. Van Vo, V. Truong, N. Le, T. Vo, B. Le, and A. Nguyen · 2023
Closest in time.
Gluegen: Plug and play multi-modal encoders for x-to-image generation
C. Qin, N. Yu, C. Xing, S. Zhang, Z. Chen, S. Ermon, Y. Fu, C. Xiong, and R. Xu · 2023
Closest in time.
Grasp-anything: Large-scale grasp dataset from foundation models
A. D. Vuong, M. N. Vu, H. Le, B. Huang, B. Huynh, T. Vo, A. Kugi, and A. Nguyen · 2023
Closest in time.
Mime: Human-aware 3d scene generation
H. Yi, C.-H. P. Huang, S. Tripathi, L. Hering, J. Thies, and M. J. Black · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
L. Zhang and M. Agrawala · 2023
Closest in time.
Ddfm: Denoising diffusion model for multi-modality image fusion
Z. Zhao, H. Bai, Y. Zhu, J. Zhang, S. Xu, Y. Zhang, K. Zhang, D. Meng, R. Timofte, and L. Van Gool · 2023
Closest in time.