Fetching the paper…
Reading the bibliography…
Shape primitive abstraction, which decomposes complex 3D shapes into simple geometric elements, plays a crucial role in human visual cognition and has broad applications in computer vision and graphics.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Machine perception of three-dimensional solids
Lawrence G Roberts. 1963 · 1963
Earlier work this paper cites.
Visual perception by computer. In Proc. IEEE Conf. on Systems and Control, 1975
Thomas Binford. 1975 · 1975
Earlier work this paper cites.
Human image understanding: Recent research and a theory
Irving Biederman. 1985 · 1985
Earlier work this paper cites.
Parts: Structured Descriptions of Shape.. In AAAI . 695–701
Alex Pentland. 1986 · 1986
Earlier work this paper cites.
Surface simplification using quadric error metrics. In Proceedings of the 24th annual conference on Computer graphics and interactive techniques . 209–216
Michael Garland and Paul S Heckbert. 1997 · 1997
Earlier work this paper cites.
Superquadrics for segmenting and modeling range data
Ales Leonardis, Ales Jaklic, and Franc Solina. 1997 · 1997
Earlier work this paper cites.
Segmentation and superquadric modeling of 3D objects
Laurent Chevalier, Fabrice Jaillet, and Atilla Baskurt. 2003 · 2003
Earlier work this paper cites.
Recognition-by-components: A theory of human image understanding
Irving Biederman. 2005 · 2005
Earlier work this paper cites.
Blocks world revisited: Image understanding using qualitative geometry and mechanics. In Computer Vision–ECCV 2010: 11th European Conference on Computer Vision, Heraklion, Crete, Greece, September 5-11, 2010, Proceedings, Part IV 11 . Springer, 482–496
Abhinav Gupta, Alexei A Efros, and Martial Hebert. 2010 · 2010
Earlier work this paper cites.
ShapeNet: An Information-Rich 3D Model Repository
Angel X. Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, Li Yi, and Fisher Yu. 2015 · 2015
Earlier work this paper cites.
A point set generation network for 3d object reconstruction from a single image. In Proceedings of the IEEE conference on computer vision and pattern recognition . 605–613
Haoqiang Fan, Hao Su, and Leonidas J Guibas. 2017 · 2017
Earlier work this paper cites.
Categorical Reparameterization with Gumbel-Softmax. In International Conference on Learning Representations
Eric Jang, Shixiang Gu, and Ben Poole. 2017 · 2017
Earlier work this paper cites.
Grass: Generative recursive autoencoders for shape structures
Jun Li, Kai Xu, Siddhartha Chaudhuri, Ersin Yumer, Hao Zhang, and Leonidas Guibas. 2017 · 2017
Earlier work this paper cites.
Learning shape abstractions by assembling volumetric primitives. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 2635–2643
Shubham Tulsiani, Hao Su, Leonidas J Guibas, Alexei A Efros, and Jitendra Malik. 2017 · 2017
Earlier work this paper cites.
3d-prnn: Generating shape primitives with recurrent neural networks. In Proceedings of the IEEE International Conference on Computer Vision . 900–909
Chuhang Zou, Ersin Yumer, Jimei Yang, Duygu Ceylan, and Derek Hoiem. 2017 · 2017
Earlier work this paper cites.
Bae-net: Branched autoencoder for shape co-segmentation. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 8490–8499
Zhiqin Chen, Kangxue Yin, Matthew Fisher, Siddhartha Chaudhuri, and Hao Zhang. 2019 · 2019
Earlier work this paper cites.
Learning shape templates with structured implicit functions. In Proceedings of the IEEE/CVF international conference on computer vision . 7154–7164
Kyle Genova, Forrester Cole, Daniel Vlasic, Aaron Sarna, William T Freeman, and Thomas Funkhouser. 2019 · 2019
Earlier work this paper cites.
Supervised fitting of geometric primitives to 3d point clouds. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2652–2660
Lingxiao Li, Minhyuk Sung, Anastasia Dubrovina, Li Yi, and Leonidas J Guibas. 2019 · 2019
Earlier work this paper cites.
StructureNet: Hierarchical Graph Networks for 3D Shape Generation
Kaichun Mo, Paul Guerrero, Li Yi, Hao Su, Peter Wonka, Niloy Mitra, and Leonidas Guibas. 2019 · 2019
Earlier work this paper cites.
Superquadrics Revisited: Learning 3D Shape Parsing beyond Cuboids. In Proceedings IEEE Conf. on Computer Vision and Pattern Recognition (CVPR)
Despoina Paschalidou, Ali Osman Ulusoy, and Andreas Geiger. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Fast and flexible indoor scene synthesis via deep convolutional generative models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6182–6190
Daniel Ritchie, Kai Wang, and Yu-an Lin. 2019 · 2019
Earlier work this paper cites.
Bsp-net: Generating compact meshes via binary space partitioning. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 45–54
Zhiqin Chen, Andrea Tagliasacchi, and Hao Zhang. 2020 · 2020
Earlier work this paper cites.
Cvxnet: Learnable convex decomposition. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 31–44
Boyang Deng, Kyle Genova, Soroosh Yazdani, Sofien Bouaziz, Geoffrey Hinton, and Andrea Tagliasacchi. 2020 · 2020
Earlier work this paper cites.
Learning generative models of shape handles. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 402–411
Matheus Gadelha, Giorgio Gori, Duygu Ceylan, Radomir Mech, Nathan Carr, Tamy Boubekeur, Rui Wang, and Subhransu Maji. 2020 · 2020
Earlier work this paper cites.
Taming transformers for high-resolution image synthesis. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 12873–12883
Patrick Esser, Robin Rombach, and Bjorn Ommer. 2021 · 2021
Cited alongside, same era.
Cpfn: Cascaded primitive fitting networks for high-resolution point clouds. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 7457–7466
Eric-Tuan Lê, Minhyuk Sung, Duygu Ceylan, Radomir Mech, Tamy Boubekeur, and Niloy J Mitra. 2021 · 2021
Cited alongside, same era.
Atiss: Autoregressive transformers for indoor scene synthesis
Despoina Paschalidou, Amlan Kar, Maria Shugrina, Karsten Kreis, Andreas Geiger, and Sanja Fidler. 2021 · 2021
Cited alongside, same era.
Zero-shot text-to-image generation. In International conference on machine learning . Pmlr, 8821–8831
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021 · 2021
Cited alongside, same era.
Sceneformer: Indoor scene generation with transformers. In 2021 International Conference on 3D Vision (3DV) . IEEE, 106–115
Coin3d: Controllable and interactive 3d assets generation with proxy-guided conditioning. In ACM SIGGRAPH 2024 Conference Papers . 1–10
Wenqi Dong, Bangbang Yang, Lin Ma, Xiao Liu, Liyuan Cui, Hujun Bao, Yuewen Ma, and Zhaopeng Cui. 2024 · 2024
Later among the works it cites.
Learning Part-aware 3D Representations by Fusing 2D Gaussians and Superquadrics
Zhirui Gao, Renjiao Yi, Yuhang Huang, Wei Chen, Chenyang Zhu, and Kai Xu. 2024 · 2024
Later among the works it cites.
Texsliders: Diffusion-based texture editing in clip space. In ACM SIGGRAPH 2024 Conference Papers . 1–11
Julia Guerrero-Viu, Milos Hasan, Arthur Roullier, Midhun Harikumar, Yiwei Hu, Paul Guerrero, Diego Gutierrez, Belen Masia, and Valentin Deschaintre. 2024 · 2024
Later among the works it cites.
Medial Skeletal Diagram: A Generalized Medial Axis Approach for Compact 3D Shape Representation
Minghao Guo, Bohan Wang, and Wojciech Matusik. 2024 · 2024
Later among the works it cites.
AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xinpeng Wang, Chandan Yeshwanth, and Matthias Nießner. 2021 · 2021
Cited alongside, same era.
Robust and accurate superquadric recovery: A probabilistic approach. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2676–2685
Weixiao Liu, Yuwei Wu, Sipu Ruan, and Gregory S Chirikjian. 2022 · 2022
Cited alongside, same era.
Point-e: A system for generating 3d point clouds from complex prompts
Alex Nichol, Heewoo Jun, Prafulla Dhariwal, Pamela Mishkin, and Mark Chen. 2022 · 2022
Cited alongside, same era.
Scalable Diffusion Models with Transformers
William Peebles and Saining Xie. 2022 · 2022
Cited alongside, same era.
Lion: Latent point diffusion models for 3d shape generation
Arash Vahdat, Francis Williams, Zan Gojcic, Or Litany, Sanja Fidler, Karsten Kreis, et al · 2022
Cited alongside, same era.
Planarrecon: Real-time 3d plane detection and reconstruction from posed monocular videos. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6219–6228
Yiming Xie, Matheus Gadelha, Fengting Yang, Xiaowei Zhou, and Huaizu Jiang. 2022 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Objaverse: A universe of annotated 3d objects. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 13142–13153
Matt Deitke, Dustin Schwenk, Jordi Salvador, Luca Weihs, Oscar Michel, Eli VanderBilt, Ludwig Schmidt, Kiana Ehsani, Aniruddha Kembhavi, and Ali Farhadi. 2023 · 2023
Cited alongside, same era.
Yuze He, Wang Zhao, Shaohui Liu, Yubin Hu, Yushi Bai, Yu-Hui Wen, and Yong-Jin Liu. 2024 · 2024
Later among the works it cites.
LRM: Large Reconstruction Model for Single Image to 3D. In The Twelfth International Conference on Learning Representations
Yicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi, Yang Zhou, Difan Liu, Feng Liu, Kalyan Sunkavalli, Trung Bui, and Hao Tan. 2024 · 2024
Later among the works it cites.
Diffusion texture painting. In ACM SIGGRAPH 2024 Conference Papers . 1–12
Anita Hu, Nishkrit Desai, Hassan Abu Alhaija, Seung Wook Kim, and Maria Shugrina. 2024 · 2024
Later among the works it cites.
Make-a-shape: a ten-million-scale 3d shape model. In Forty-first International Conference on Machine Learning
Ka-Hei Hui, Aditya Sanghi, Arianna Rampini, Kamal Rahimi Malekshan, Zhengzhe Liu, Hooman Shayani, and Chi-Wing Fu. 2024 · 2024
Later among the works it cites.
PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers
Songlin Li, Despoina Paschalidou, and Leonidas Guibas. 2024b · 2024
Later among the works it cites.
Weiyu Li, Jiarui Liu, Hongyu Yan, Rui Chen, Yixun Liang, Xuelin Chen, Ping Tan, and Xiaoxiao Long. 2024a · 2024
Later among the works it cites.
Gem3d: Generative medial abstractions for 3d shape synthesis. In ACM SIGGRAPH 2024 Conference Papers . 1–11
Dmitry Petrov, Pradyumn Goyal, Vikas Thamizharasan, Vladimir Kim, Matheus Gadelha, Melinos Averkiou, Siddhartha Chaudhuri, and Evangelos Kalogerakis. 2024 · 2024
Later among the works it cites.
Meshgpt: Generating triangle meshes with decoder-only transformers. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 19615–19625
Yawar Siddiqui, Antonio Alliegro, Alexey Artemov, Tatiana Tommasi, Daniele Sirigatti, Vladislav Rosov, Angela Dai, and Matthias Nießner. 2024 · 2024
Later among the works it cites.
Edgerunner: Auto-regressive auto-encoder for artistic mesh generation
Jiaxiang Tang, Zhaoshuo Li, Zekun Hao, Xian Liu, Gang Zeng, Ming-Yu Liu, and Qinsheng Zhang. 2024 · 2024
Later among the works it cites.
Visual autoregressive modeling: Scalable image generation via next-scale prediction
Keyu Tian, Yi Jiang, Zehuan Yuan, Bingyue Peng, and Liwei Wang. 2024 · 2024
Later among the works it cites.
TripoSR: Fast 3D Object Reconstruction from a Single Image
Dmitry Tochilkin, David Pankratz, Zexiang Liu, Zixuan Huang, , Adam Letts, Yangguang Li, Ding Liang, Christian Laforte, Varun Jampani, and Yan-Pei Cao. 2024 · 2024
Later among the works it cites.
CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model
Zhengyi Wang, Yikai Wang, Yifei Chen, Chendong Xiang, Shuo Chen, Dajiang Yu, Chongxuan Li, Hang Su, and Jun Zhu. 2024b · 2024
Later among the works it cites.
Structured 3D Latents for Scalable and Versatile 3D Generation
Jianfeng Xiang, Zelong Lv, Sicheng Xu, Yu Deng, Ruicheng Wang, Bowen Zhang, Dong Chen, Xin Tong, and Jiaolong Yang. 2024 · 2024
Later among the works it cites.
Jiale Xu, Weihao Cheng, Yiming Gao, Xintao Wang, Shenghua Gao, and Ying Shan. 2024 · 2024
Later among the works it cites.
Tencent Hunyuan3D-1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation
Xianghui Yang, Huiwen Shi, Bowen Zhang, Fan Yang, Jiacheng Wang, Hongxu Zhao, Xinhai Liu, Xinzhou Wang, Qingxiang Lin, Jiaao Yu, Lifu Wang, Zhuo Chen, Sicong Liu, Yuhong Liu, Yong Yang, Di Wang, Jie Jiang, and Chunchao Guo. 2024 · 2024
Later among the works it cites.
TEXGen: a Generative Diffusion Model for Mesh Textures
Xin Yu, Ze Yuan, Yuan-Chen Guo, Ying-Tian Liu, Jianhui Liu, Yangguang Li, Yan-Pei Cao, Ding Liang, and Xiaojuan Qi. 2024 · 2024
Later among the works it cites.
3D representation in 512-Byte: Variational tokenizer is the key for autoregressive 3D generation
Jinzhi Zhang, Feng Xiong, and Mu Xu. 2024d · 2024
Later among the works it cites.
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
Kai Zhang, Sai Bi, Hao Tan, Yuanbo Xiangli, Nanxuan Zhao, Kalyan Sunkavalli, and Zexiang Xu. 2024a · 2024
Later among the works it cites.
CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets
Longwen Zhang, Ziyu Wang, Qixuan Zhang, Qiwei Qiu, Anqi Pang, Haoran Jiang, Wei Yang, Lan Xu, and Jingyi Yu. 2024c · 2024
Later among the works it cites.
Mapa: Text-driven photorealistic material painting for 3d shapes. In ACM SIGGRAPH 2024 Conference Papers . 1–12
Shangzhan Zhang, Sida Peng, Tao Xu, Yuanbo Yang, Tianrun Chen, Nan Xue, Yujun Shen, Hujun Bao, Ruizhen Hu, and Xiaowei Zhou. 2024b · 2024
Later among the works it cites.
Roomdesigner: Encoding anchor-latents for style-consistent and shape-compatible indoor scene generation. In 2024 International Conference on 3D Vision (3DV) . IEEE, 1413–1423
Yiqun Zhao, Zibo Zhao, Jing Li, Sixun Dong, and Shenghua Gao. 2024 · 2024
Later among the works it cites.