Fetching the paper…
Reading the bibliography…
Autoregressive models have achieved remarkable success across various domains, yet their performance in 3D shape generation lags significantly behind that of diffusion models.
Marching Cubes: A High Resolution 3D Surface Construction Algorithm. In SIGGRAPH
William E. Lorensen and Harvey E. Cline. 1987 · 1987
Earlier work this paper cites.
Multi-level partition of unity implicits
Yutaka Ohtake, Alexander Belyaev, Marc Alexa, Greg Turk, and Hans-Peter Seidel. 2003 · 2003
Earlier work this paper cites.
ShapeNet: An information-rich 3D model repository
Angel X. Chang, Thomas Funkhouser, Leonidas J. Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, Li Yi, and Fisher Yu. 2015 · 2015
Earlier work this paper cites.
Generative adversarial networks. In NeurIPS
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2016 · 2016
Earlier work this paper cites.
Learning a probabilistic latent space of object shapes via 3D generative-adversarial modeling. In NeurIPS
Jiajun Wu, Chengkai Zhang, Tianfan Xue, William T. Freeman, and Joshua B. Tenenbaum. 2016 · 2016
Earlier work this paper cites.
Neural Discrete Representation Learning. In NeurIPS
Aaron van den Oord, Oriol Vinyals, and Koray Kavukcuoglu. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In NeurIPS
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
O-CNN: Octree-based convolutional neural networks for 3D shape analysis
Peng-Shuai Wang, Yang Liu, Yu-Xiao Guo, Chun-Yu Sun, and Xin Tong. 2017 · 2017
Earlier work this paper cites.
Text2Shape: Generating Shapes from Natural Language by Learning Joint Embeddings
Kevin Chen, Christopher B. Choy, Manolis Savva, Angel X Chang, Thomas Funkhouser, and Silvio Savarese. 2018 · 2018
Earlier work this paper cites.
Adaptive O-CNN: A patch-based deep representation of 3D shapes
Peng-Shuai Wang, Chun-Yu Sun, Yang Liu, and Xin Tong. 2018 · 2018
Earlier work this paper cites.
Learning implicit fields for generative shape modeling. In CVPR
Zhiqin Chen and Hao Zhang. 2019 · 2019
Earlier work this paper cites.
Decoupled weight decay regularization. In ICLR
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners. In NeurIPS
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models. In NeurIPS
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
PolyGen: An autoregressive generative model of 3D meshes. In ICML
Charlie Nash, Yaroslav Ganin, SM Ali Eslami, and Peter Battaglia. 2020 · 2020
Earlier work this paper cites.
Convolutional occupancy networks. In ECCV
Songyou Peng, Michael Niemeyer, Lars Mescheder, Marc Pollefeys, and Andreas Geiger. 2020 · 2020
Earlier work this paper cites.
PointGrow: Autoregressively Learned Point Cloud Generation with Self-Attention. In WACV
Yongbin Sun, Yue Wang, Ziwei Liu, Joshua E Siegel, and Sanjay E. Sarma. 2020 · 2020
Earlier work this paper cites.
PCT: Point cloud transformer
Meng-Hao Guo, Jun-Xiong Cai, Zheng-Ning Liu, Tai-Jiang Mu, Ralph R Martin, and Shi-Min Hu. 2021 · 2021
Earlier work this paper cites.
Diffusion Probabilistic Models for 3D Point Cloud Generation. In CVPR
Shitong Luo and Wei Hu. 2021 · 2021
Earlier work this paper cites.
Atiss: Autoregressive transformers for indoor scene synthesis. In NeurIPS
Despoina Paschalidou, Amlan Kar, Maria Shugrina, Karsten Kreis, Andreas Geiger, and Sanja Fidler. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision. In ICML
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
SceneFormer: Indoor scene generation with transformers. In 3DV
Xinpeng Wang, Chandan Yeshwanth, and Matthias Nießner. 2021 · 2021
Earlier work this paper cites.
Point transformer. In ICCV
Hengshuang Zhao, Li Jiang, Jiaya Jia, Philip Torr, and Vladlen Koltun. 2021 · 2021
Earlier work this paper cites.
Autoregressive 3D shape generation via canonical mapping. In ECCV
An-Chieh Cheng, Xueting Li, Sifei Liu, Min Sun, and Ming-Hsuan Yang. 2022 · 2022
Earlier work this paper cites.
Embracing single stride 3D object detector with sparse transformer. In CVPR
Lue Fan, Ziqi Pang, Tianyuan Zhang, Yu-Xiong Wang, Hang Zhao, Feng Wang, Naiyan Wang, and Zhaoxiang Zhang. 2022 · 2022
Earlier work this paper cites.
SPAGHETTI: Editing Implicit Shapes Through Part Aware Generation
Amir Hertz, Or Perel, Raja Giryes, Olga Sorkine-Hornung, and Daniel Cohen-Or. 2022 · 2022
Earlier work this paper cites.
Neural Wavelet-Domain Diffusion for 3D Shape Generation. In SIGGRAPH Asia
Ka-Hei Hui, Ruihui Li, Jingyu Hu, and Chi-Wing Fu. 2022 · 2022
Earlier work this paper cites.
Stratified Transformer for 3D Point Cloud Segmentation. In CVPR
Xin Lai, Jianhui Liu, Li Jiang, Liwei Wang, Hengshuang Zhao, Shu Liu, Xiaojuan Qi, and Jiaya Jia. 2022 · 2022
Earlier work this paper cites.
AutoSDF: Shape Priors for 3D Completion, Reconstruction and Generation. In CVPR
Paritosh Mittal, Yen-Chi Cheng, Maneesh Singh, and Shubham Tulsiani. 2022 · 2022
Earlier work this paper cites.
Point-E: A system for generating 3D point clouds from complex prompts
Alex Nichol, Heewoo Jun, Prafulla Dhariwal, Pamela Mishkin, and Mark Chen. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback. In NeurIPS
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Masked autoencoders for point cloud self-supervised learning. In ECCV
Yatian Pang, Wenxiao Wang, Francis EH Tay, Wei Liu, Yonghong Tian, and Li Yuan. 2022 · 2022
Cited alongside, same era.
SWFormer: Sparse Window Transformer for 3D Object Detection in Point Clouds. In ECCV
Pei Sun, Mingxing Tan, Weiyue Wang, Chenxi Liu, Fei Xia, Zhaoqi Leng, and Dragomir Anguelov. 2022 · 2022
Cited alongside, same era.
Dual Octree Graph Networks for Learning Adaptive Volumetric Shape Representations
Peng-Shuai Wang, Yang Liu, and Xin Tong. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models. In NeurIPS
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
SAR3D: Autoregressive 3D object generation and understanding via multi-scale 3D VQVAE
Yongwei Chen, Yushi Lan, Shangchen Zhou, Tengfei Wang, and XIngang Pan. 2024b · 2024
Later among the works it cites.
Meshanything v2: Artist-created mesh generation with adjacent mesh tokenization
Yiwen Chen, Yikai Wang, Yihao Luo, Zhengyi Wang, Zilong Chen, Jun Zhu, Chi Zhang, and Guosheng Lin. 2024c · 2024
Later among the works it cites.
Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale
Zekun Hao, David W Romero, Tsung-Yi Lin, and Ming-Yu Liu. 2024 · 2024
Later among the works it cites.
LN3Diff: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation. In ECCV
Yushi Lan, Fangzhou Hong, Shuai Yang, Shangchen Zhou, Xuyi Meng, Bo Dai, Xingang Pan, and Chen Change Loy. 2024 · 2024
Later among the works it cites.
Autoregressive Image Generation without Vector Quantization. In NeurIPS
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Point Transformer V2: Grouped Vector Attention and Partition-based Pooling. In NeurIPS
Xiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu, and Hengshuang Zhao. 2022 · 2022
Cited alongside, same era.
ShapeFormer: Transformer-based shape completion via sparse representation. In CVPR
Xingguang Yan, Liqiang Lin, Niloy J. Mitra, Dani Lischinski, Daniel Cohen-Or, and Hui Huang. 2022 · 2022
Cited alongside, same era.
Point-BERT: Pre-training 3D point cloud transformers with masked point modeling. In CVPR
Xumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang, Jie Zhou, and Jiwen Lu. 2022 · 2022
Cited alongside, same era.
LION: Latent Point Diffusion Models for 3D Shape Generation. In NeurIPS
Xiaohui Zeng, Arash Vahdat, Francis Williams, Zan Gojcic, Or Litany, Sanja Fidler, and Karsten Kreis. 2022 · 2022
Cited alongside, same era.
3DILG: Irregular latent grids for 3D generative modeling. In NeurIPS
Biao Zhang, Matthias Nießner, and Peter Wonka. 2022 · 2022
Cited alongside, same era.
SDF-StyleGAN: Implicit SDF-Based StyleGAN for 3D Shape Generation
Xin-Yang Zheng, Yang Liu, Peng-Shuai Wang, and Xin Tong. 2022 · 2022
Cited alongside, same era.
SDFusion: Multimodal 3D Shape Completion, Reconstruction, and Generation. In CVPR
Yen-Chi Cheng, Hsin-Ying Lee, Sergey Tuyakov, Alex Schwing, and Liangyan Gui. 2023 · 2023
Cited alongside, same era.
Tianhong Li, Yonglong Tian, He Li, Mingyang Deng, and Kaiming He. 2024 · 2024
Later among the works it cites.
Visual instruction tuning. In NeurIPS
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2024 · 2024
Later among the works it cites.
Pushing Auto-regressive Models for 3D Shape Generation at Capacity and Scalability
Xuelin Qian, Yu Wang, Simian Luo, Yinda Zhang, Ying Tai, Zhenyu Zhang, Chengjie Wang, Xiangyang Xue, Bo Zhao, Tiejun Huang, Yunsheng Wu, and Yanwei Fu. 2024 · 2024
Later among the works it cites.
XCube ( 𝒳 3 \mathcal{X}^{3} ): Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies. In CVPR
Xuanchi Ren, Jiahui Huang, Xiaohui Zeng, Ken Museth, Sanja Fidler, and Francis Williams. 2024 · 2024
Later among the works it cites.
L3DG: Latent 3D Gaussian Diffusion. In SIGGRAPH Asia
Barbara Roessle, Norman Müller, Lorenzo Porzi, Samuel Rota Bulò, Peter Kontschieder, Angela Dai, and Matthias Nießner. 2024 · 2024
Later among the works it cites.
MeshGPT: Generating triangle meshes with decoder-only transformers. In CVPR
Yawar Siddiqui, Antonio Alliegro, Alexey Artemov, Tatiana Tommasi, Daniele Sirigatti, Vladislav Rosov, Angela Dai, and Matthias Nießner. 2024 · 2024
Later among the works it cites.
RoFormer: Enhanced transformer with rotary position embedding
Jianlin Su, Murtadha Ahmed, Yu Lu, Shengfeng Pan, Wen Bo, and Yunfeng Liu. 2024 · 2024
Later among the works it cites.
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Peize Sun, Yi Jiang, Shoufa Chen, Shilong Zhang, Bingyue Peng, Ping Luo, and Zehuan Yuan. 2024 · 2024
Later among the works it cites.
LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation
Jiaxiang Tang, Zhaoxi Chen, Xiaokang Chen, Tengfei Wang, Gang Zeng, and Ziwei Liu. 2024a · 2024
Later among the works it cites.
EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation
Jiaxiang Tang, Zhaoshuo Li, Zekun Hao, Xian Liu, Gang Zeng, Ming-Yu Liu, and Qinsheng Zhang. 2024b · 2024
Later among the works it cites.
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction. In NeurIPS
Keyu Tian, Yi Jiang, Zehuan Yuan, Bingyue Peng, and Liwei Wang. 2024 · 2024
Later among the works it cites.
Emu3: Next-token prediction is all you need
Xinlong Wang, Xiaosong Zhang, Zhengxiong Luo, Quan Sun, Yufeng Cui, Jinsheng Wang, Fan Zhang, Yueze Wang, Zhen Li, Qiying Yu, et al · 2024
Later among the works it cites.
LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models
Zhengyi Wang, Jonathan Lorraine, Yikai Wang, Hang Su, Jun Zhu, Sanja Fidler, and Xiaohui Zeng. 2024a · 2024
Later among the works it cites.
Scaling Mesh Generation via Compressive Tokenization
Haohan Weng, Zibo Zhao, Biwen Lei, Xianghui Yang, Jian Liu, Zeqiang Lai, Zhuo Chen, Yuhong Liu, Jie Jiang, Chunchao Guo, Tong Zhang, Shenghua Gao, and C. L. Philip Chen. 2024 · 2024
Later among the works it cites.
Point Transformer V3: Simpler Faster Stronger. In CVPR
Xiaoyang Wu, Li Jiang, Peng-Shuai Wang, Zhijian Liu, Xihui Liu, Yu Qiao, Wanli Ouyang, Tong He, and Hengshuang Zhao. 2024 · 2024
Later among the works it cites.
Structured 3D Latents for Scalable and Versatile 3D Generation
Jianfeng Xiang, Zelong Lv, Sicheng Xu, Yu Deng, Ruicheng Wang, Bowen Zhang, Dong Chen, Xin Tong, and Jiaolong Yang. 2024 · 2024
Later among the works it cites.
TexGaussian: Generating High-quality PBR Material via Octree-based 3D Gaussian Splatting
Bojun Xiong, Jialun Liu, Jiakui Hu, Chenming Wu, Jinbo Wu, Xing Liu, Chen Zhao, Errui Ding, and Zhouhui Lian. 2024a · 2024
Later among the works it cites.
OctFusion: Octree-based Diffusion Models for 3D Shape Generation
Bojun Xiong, Si-Tong Wei, Xin-Yang Zheng, Yan-Pei Cao, Zhouhui Lian, and Peng-Shuai Wang. 2024b · 2024
Later among the works it cites.
Swin3D: A pretrained transformer backbone for 3D indoor scene understanding
Yu-Qi Yang, Yu-Xiao Guo, Jian-Yu Xiong, Yang Liu, Hao Pan, Peng-Shuai Wang, Xin Tong, and Baining Guo. 2024 · 2024
Later among the works it cites.
GaussianCube: Structuring Gaussian Splatting using Optimal Transport for 3D Generative Modeling
Bowen Zhang, Yiji Cheng, Jiaolong Yang, Chunyu Wang, Feng Zhao, Yansong Tang, Dong Chen, and Baining Guo. 2024a · 2024
Later among the works it cites.
3D representation in 512-Byte: Variational tokenizer is the key for autoregressive 3D generation
Jinzhi Zhang, Feng Xiong, and Mu Xu. 2024c · 2024
Later among the works it cites.
Jinzhi Zhang, Feng Xiong, and Mu Xu. 2024d · 2024
Later among the works it cites.
CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets
Longwen Zhang, Ziyu Wang, Qixuan Zhang, Qiwei Qiu, Anqi Pang, Haoran Jiang, Wei Yang, Lan Xu, and Jingyi Yu. 2024b · 2024
Later among the works it cites.
Image and Video Tokenization with Binary Spherical Quantization
Yue Zhao, Yuanjun Xiong, and Philipp Krähenbühl. 2024 · 2024
Later among the works it cites.
Rotary position embedding for vision transformer. In ECCV
Byeongho Heo, Song Park, Dongyoon Han, and Sangdoo Yun. 2025 · 2025
Closest in time.
Data-parallel octrees for surface reconstruction
Kun Zhou, Minmin Gong, Xin Huang, and Baining Guo. 2011 · 2025
Closest in time.