Fetching the paper…
Reading the bibliography…
3D layout tasks have traditionally concentrated on geometric constraints, but many practical applications demand richer contextual understanding that spans social interactions, cultural traditions, and usage conventions.
Understanding natural language
Terry Winograd. 1972 · 1972
Earlier work this paper cites.
Put: Language-Based Interactive Manipulation of Objects
Sharon Rose Clay and Jane Wilhelms. 1996a · 1996
Earlier work this paper cites.
Put: Language-based interactive manipulation of objects
Sharon Rose Clay and Jane Wilhelms. 1996b · 1996
Earlier work this paper cites.
WordsEye: an automatic text-to-scene conversion system. In Proceedings of the 28th Annual Conference on Computer Graphics and Interactive Techniques, SIGGRAPH 2001, Los Angeles, California, USA, August 12-17, 2001 , Lynn Pocock (Ed.). ACM, 487–496
Robert Coyne and Richard Sproat. 2001b · 2001
Earlier work this paper cites.
Interactive furniture layout using interior design guidelines. In ACM SIGGRAPH 2011 Papers (Vancouver, British Columbia, Canada) (SIGGRAPH ’11) . Association for Computing Machinery, New York, NY, USA, Article 87, 10 pages
Paul Merrell, Eric Schkufza, Zeyang Li, Maneesh Agrawala, and Vladlen Koltun. 2011 · 2011
Earlier work this paper cites.
Make it home: automatic optimization of furniture arrangement
Lap-Fai Yu, Sai-Kit Yeung, Chi-Keung Tang, Demetri Terzopoulos, Tony F. Chan, and Stanley J. Osher. 2011 · 2011
Earlier work this paper cites.
Learning the Visual Interpretation of Sentences. In IEEE International Conference on Computer Vision, ICCV 2013, Sydney, Australia, December 1-8, 2013 . IEEE Computer Society, 1681–1688
C. Lawrence Zitnick, Devi Parikh, and Lucy Vanderwende. 2013 · 2013
Earlier work this paper cites.
Beyond pascal: A benchmark for 3d object detection in the wild. In IEEE winter conference on applications of computer vision . IEEE, 75–82
Yu Xiang, Roozbeh Mottaghi, and Silvio Savarese. 2014 · 2014
Earlier work this paper cites.
Text to 3D Scene Generation with Rich Lexical Grounding. In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing of the Asian Federation of Natural Language Processing, ACL 2015, July 26-31, 2015, Beijing, China, Volume 1: Long Papers . The Association for Computer Linguistics, 53–62
Angel X. Chang, Will Monroe, Manolis Savva, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Relationship templates for creating scene variations
Xi Zhao, Ruizhen Hu, Paul Guerrero, Niloy Mitra, and Taku Komura. 2016 · 2016
Earlier work this paper cites.
SceneSeer: 3D Scene Design with Natural Language
Angel X. Chang, Mihail Eric, Manolis Savva, and Christopher D. Manning. 2017b · 2017
Earlier work this paper cites.
Designing deep convolutional neural networks for continuous object orientation estimation
Kota Hara, Raviteja Vemulapalli, and Rama Chellappa. 2017 · 2017
Earlier work this paper cites.
An interactive system for efficient 3D furniture arrangement. In Proceedings of the Computer Graphics International Conference (Yokohama, Japan) (CGI ’17) . Association for Computing Machinery, New York, NY, USA, Article 29, 6 pages
Meng Yan, Xuejin Chen, and Jie Zhou. 2017 · 2017
Cited alongside, same era.
Automatic Furniture Arrangement Using Greedy Cost Minimization. In 2018 IEEE Conference on Virtual Reality and 3D User Interfaces (VR) . 491–498
Peter Kán and Hannes Kaufmann. 2018 · 2018
Cited alongside, same era.
Deep convolutional priors for indoor scene synthesis
Kai Wang, Manolis Savva, Angel X. Chang, and Daniel Ritchie. 2018 · 2018
Cited alongside, same era.
GRAINS: Generative Recursive Autoencoders for INdoor Scenes
Manyi Li, Akshay Gadi Patil, Kai Xu, Siddhartha Chaudhuri, Owais Khan, Ariel Shamir, Changhe Tu, Baoquan Chen, Daniel Cohen-Or, and Hao Zhang. 2019 · 2019
Cited alongside, same era.
Fast and Flexible Indoor Scene Synthesis via Deep Convolutional Generative Models. In 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 6175–6183
LEGO-Net: Learning Regular Rearrangements of Objects in Rooms. In IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2023, Vancouver, BC, Canada, June 17-24, 2023 . IEEE, 19037–19047
Qiuhong Anna Wei, Sijie Ding, Jeong Joon Park, Rahul Sajnani, Adrien Poulenard, Srinath Sridhar, and Leonidas J. Guibas. 2023 · 2023
Later among the works it cites.
Towards Text-guided 3D Scene Composition
Qihang Zhang, Chaoyang Wang, Aliaksandr Siarohin, Peiye Zhuang, Yinghao Xu, Ceyuan Yang, Dahua Lin, Bolei Zhou, S. Tulyakov, and Hsin-Ying Lee. 2023 · 2023
Later among the works it cites.
3d neural embedding likelihood: Probabilistic inverse graphics for robust 6d pose estimation. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 21625–21636
Guangyao Zhou, Nishad Gothoskar, Lirui Wang, Joshua B Tenenbaum, Dan Gutfreund, Miguel Lázaro-Gredilla, Dileep George, and Vikash K Mansinghka. 2023 · 2023
Later among the works it cites.
LLMR: Real-time Prompting of Interactive Worlds using Large Language Models. In Proceedings of the CHI Conference on Human Factors in Computing Systems (CHI ’24) . ACM, 1–22
Fernanda De La Torre, Cathy Mengying Fang, Han Huang, Andrzej Banburski-Fahey, Judith Amores Fernandez, and Jaron Lanier. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Daniel Ritchie, Kai Wang, and Yu-An Lin. 2019 · 2019
Cited alongside, same era.
SceneGraphNet: Neural Message Passing for 3D Indoor Scene Augmentation. In 2019 IEEE/CVF International Conference on Computer Vision (ICCV) . 7383–7391
Yang Zhou, Zachary While, and Evangelos Kalogerakis. 2019 · 2019
Cited alongside, same era.
House-GAN: Relational Generative Adversarial Networks for Graph-Constrained House Layout Generation. In Computer Vision – ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part I (Glasgow, United Kingdom). Springer-Verlag, Berlin, Heidelberg, 162–177
Nelson Nauata, Kai-Hung Chang, Chin-Yi Cheng, Greg Mori, and Yasutaka Furukawa. 2020 · 2020
Cited alongside, same era.
Gdr-net: Geometry-guided direct regression network for monocular 6d object pose estimation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 16611–16621
Gu Wang, Fabian Manhardt, Federico Tombari, and Xiangyang Ji. 2021 · 2021
Cited alongside, same era.
Diffusion-based Generation, Optimization, and Planning in 3D Scenes. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 16750–16761
Siyuan Huang, Zan Wang, Puhao Li, Baoxiong Jia, Tengyu Liu, Yixin Zhu, Wei Liang, and Song-Chun Zhu. 2023 · 2023
Cited alongside, same era.
Advances in Data-Driven Analysis and Synthesis of 3D Indoor Scenes
Akshay Gadi Patil, Supriya Gadi Patil, Manyi Li, Matthew Fisher, Manolis Savva, and Hao Zhang. 2023 · 2023
Cited alongside, same era.
RoomDreamer: Text-Driven 3D Indoor Scene Synthesis with Coherent Geometry and Texture. In Proceedings of the 31st ACM International Conference on Multimedia, MM 2023, Ottawa, ON, Canada, 29 October 2023- 3 November 2023 , Abdulmotaleb El-Saddik, Tao Mei, Rita Cucchiara, Marco Bertini, Diana Patricia Tobon Vallejo, Pradeep K. Atrey, and M. Shamim Hossain (Eds.). ACM, 6898–6906
Liangchen Song, Liangliang Cao, Hongyu Xu, Kai Kang, Feng Tang, Junsong Yuan, and Zhao Yang. 2023 · 2023
Cited alongside, same era.
3D-GPT: Procedural 3D Modeling with Large Language Models
Chunyi Sun, Junlin Han, Weijian Deng, Xinlong Wang, Zishan Qin, and Stephen Gould. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
LayoutGPT: compositional visual planning and generation with large language models. In Proceedings of the 37th International Conference on Neural Information Processing Systems (New Orleans, LA, USA) (NIPS ’23) . Curran Associates Inc., Red Hook, NY, USA, Article 802, 26 pages
Weixi Feng, Wanrong Zhu, Tsu-jui Fu, Varun Jampani, Arjun Akula, Xuehai He, Sugato Basu, Xin Eric Wang, and William Yang Wang. 2024 · 2024
Later among the works it cites.
SceneCraft: An LLM Agent for Synthesizing 3D Scenes as Blender Code. In Forty-first International Conference on Machine Learning
Ziniu Hu, Ahmet Iscen, Aashi Jain, Thomas Kipf, Yisong Yue, David A Ross, Cordelia Schmid, and Alireza Fathi. 2024 · 2024
Later among the works it cites.
ATISS: autoregressive transformers for indoor scene synthesis. In Proceedings of the 35th International Conference on Neural Information Processing Systems (NIPS ’21) . Curran Associates Inc., Red Hook, NY, USA, Article 919, 14 pages
Despoina Paschalidou, Amlan Kar, Maria Shugrina, Karsten Kreis, Andreas Geiger, and Sanja Fidler. 2024 · 2024
Later among the works it cites.
DiffuScene: Denoising Diffusion Models for Generative Indoor Scene Synthesis. In 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 20507–20518
Jiapeng Tang, Yinyu Nie, Lev Markhasin, Angela Dai, Justus Thies, and Matthias Nießner. 2024 · 2024
Later among the works it cites.
The Scene Language: Representing Scenes with Programs, Words, and Embeddings
Yunzhi Zhang, Zizhang Li, Matt Zhou, Shangzhe Wu, and Jiajun Wu. 2024 · 2024
Later among the works it cites.
Scenescript: Reconstructing scenes with an autoregressive structured language model. In European Conference on Computer Vision . Springer, 247–263
Armen Avetisyan, Christopher Xie, Henry Howard-Jenkins, Tsun-Yi Yang, Samir Aroudj, Suvam Patra, Fuyang Zhang, Duncan Frost, Luke Holland, Campbell Orme, et al · 2025
Closest in time.
Learning Spatial Knowledge for Text to 3D Scene Generation. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing, EMNLP 2014, October 25-29, 2014, Doha, Qatar, A meeting of SIGDAT, a Special Interest Group of the ACL , Alessandro Moschitti, Bo Pang, and Walter Daelemans (Eds.). ACL, 2028–2038
Angel X. Chang, Manolis Savva, and Christopher D. Manning. 2014b · 2038
Closest in time.