Fetching the paper…
Reading the bibliography…
We propose octree-based transformers, named OctFormer, for 3D point cloud learning.
Octrees for faster isosurface generation
Jane Wilhelms and Allen Van Gelder. 1992 · 1992
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In CVPR
Jia Deng, Wei Dong, Richard Socher, Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
Data-parallel octrees for surface reconstruction
Kun Zhou, Minmin Gong, Xin Huang, and Baining Guo. 2011 · 2011
Earlier work this paper cites.
Batch Normalization: Accelerating deep network training by reducing internal covariate shift. In ICML
Sergey Loffe and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
VoxNet: A 3D convolutional neural network for real-time object recognition. In IROS
Daniel Maturana and Sebastian Scherer. 2015 · 2015
Earlier work this paper cites.
SUN RGB-D: A RGB-D scene understanding benchmark suite. In CVPR
Shuran Song, Samuel P Lichtenberg, and Jianxiong Xiao. 2015 · 2015
Earlier work this paper cites.
Multi-view convolutional neural networks for 3D shape recognition. In ICCV
Hang Su, Subhransu Maji, Evangelos Kalogerakis, and Erik Learned-Miller. 2015 · 2015
Earlier work this paper cites.
3D ShapeNets: A deep representation for volumetric shape modeling. In CVPR
Zhirong Wu, Shuran Song, Aditya Khosla, Fisher Yu, Linguang Zhang, Xiaoou Tang, and Jianxiong Xiao. 2015 · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016 · 2016
Earlier work this paper cites.
Volumetric and multi-view CNNs for object classification on 3D data. In CVPR
Charles R. Qi, Hao Su, Matthias Nießner, Angela Dai, Mengyuan Yan, and Leonidas J. Guibas. 2016 · 2016
Earlier work this paper cites.
ScanNet: Richly-annotated 3D Reconstructions of Indoor Scenes. In CVPR
Angela Dai, Angel X. Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner. 2017 · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection. In CVPR
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie. 2017 · 2017
Earlier work this paper cites.
Dynamic edge-conditioned filters in convolutional neural networks on graphs. In CVPR
Martin Simonovsky and Nikos Komodakis. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In NeurIPS
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
O-CNN: Octree-based convolutional neural networks for 3D shape analysis
Peng-Shuai Wang, Yang Liu, Yu-Xiao Guo, Chun-Yu Sun, and Xin Tong. 2017 · 2017
Earlier work this paper cites.
Point Convolutional Neural Networks by Extension Operators
Matan Atzmon, Haggai Maron, and Yaron Lipman. 2018 · 2018
Earlier work this paper cites.
3DMV: Joint 3D-multi-view prediction for 3D semantic scene segmentation. In ECCV
Angela Dai and Matthias Nießner. 2018 · 2018
Earlier work this paper cites.
SplineCNN: Fast Geometric Deep Learning with Continuous B-Spline Kernels. In CVPR
Matthias Fey, Jan Eric Lenssen, Frank Weichert, and Heinrich Müller. 2018 · 2018
Earlier work this paper cites.
3D semantic segmentation with submanifold sparse convolutional networks. In CVPR
Benjamin Graham, Martin Engelcke, and Laurens van der Maaten. 2018 · 2018
Earlier work this paper cites.
PointCNN: Convolution on X-transformed points. In NeurIPS
Yangyan Li, Rui Bu, Mingchao Sun, Wei Wu, Xinhan Di, and Baoquan Chen. 2018 · 2018
Earlier work this paper cites.
H-CNN: spatial hashing based CNN for 3D shape analysis
Tianjia Shao, Yin Yang, Yanlin Weng, Qiming Hou, and Kun Zhou. 2018 · 2018
Earlier work this paper cites.
Adaptive O-CNN: A patch-based deep representation of 3D shapes
Peng-Shuai Wang, Chun-Yu Sun, Yang Liu, and Xin Tong. 2018 · 2018
Earlier work this paper cites.
SpiderCNN: Deep Learning on Point Sets with Parameterized Convolutional Filters. In ECCV
Yifan Xu, Tianqi Fan, Mingye Xu, Long Zeng, and Yu Qiao. 2018 · 2018
Earlier work this paper cites.
A unified point-based framework for 3D segmentation. In 3DV
Hung-Yueh Chiang, Yen-Liang Lin, Yueh-Cheng Liu, and Winston H Hsu. 2019 · 2019
Earlier work this paper cites.
4D spatio-temporal convnets: Minkowski convolutional neural networks. In CVPR
Christopher Choy, JunYoung Gwak, and Silvio Savarese. 2019 · 2019
Cited alongside, same era.
Panoptic feature pyramid networks. In CVPR
Alexander Kirillov, Ross Girshick, Kaiming He, and Piotr Dollár. 2019 · 2019
Cited alongside, same era.
Octree guided CNN with spherical kernels for 3D point clouds. In CVPR
Huan Lei, Naveed Akhtar, and Ajmal Mian. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization. In ICLR
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
PanopticFusion: Online volumetric semantic mapping at the level of stuff and things. In IROS
Gaku Narita, Takashi Seno, Tomoya Ishikawa, and Yohsuke Kaji. 2019 · 2019
Cited alongside, same era.
PyTorch: An Imperative Style, High-Performance Deep Learning Library. In NeurIPS
Mix3D: Out-of-Context Data Augmentation for 3D Scenes. In 3DV
Alexey Nekrasov, Jonas Schult, Or Litany, Bastian Leibe, and Francis Engelmann. 2021 · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision. In ICML
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Pyramid vision transformer: A versatile backbone for dense prediction without convolutions. In ICCV
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. 2021 · 2021
Later among the works it cites.
Early convolutions help transformers see better. In NeurIPS
Tete Xiao, Mannat Singh, Eric Mintun, Trevor Darrell, Piotr Dollár, and Ross Girshick. 2021 · 2021
Later among the works it cites.
VENet: Voting Enhancement Network for 3D Object Detection. In ICCV
Qian Xie, Yu-Kun Lai, Jing Wu, Zhoutao Wang, Dening Lu, Mingqiang Wei, and Jun Wang. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Cited alongside, same era.
Deep Hough Voting for 3D Object Detection in Point Clouds. In CVPR
Charles R. Qi, Or Litany, Kaiming He, and Leonidas J. Guibas. 2019 · 2019
Cited alongside, same era.
KPConv: Flexible and deformable convolution for point clouds. In ICCV
Hugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui, François Goulette, and Leonidas J. Guibas. 2019 · 2019
Cited alongside, same era.
PyTorch Image Models
Ross Wightman. 2019 · 2019
Cited alongside, same era.
PointConv: Deep Convolutional Networks on 3D Point Clouds. In CVPR
Wenxuan Wu, Zhongang Qi, and Li Fuxin. 2019 · 2019
Cited alongside, same era.
A hierarchical graph network for 3D object detection on point clouds. In CVPR
Jintai Chen, Biwen Lei, Qingyu Song, Haochao Ying, Danny Z Chen, and Jian Wu. 2020 · 2020
Cited alongside, same era.
On the relationship between self-attention and convolutional layers. In ICLR
Jean-Baptiste Cordonnier, Andreas Loukas, and Martin Jaggi. 2020 · 2020
Cited alongside, same era.
Focal Self-attention for Local-Global Interactions in Vision Transformers. In NeurIPS
Jianwei Yang, Chunyuan Li, Pengchuan Zhang, Xiyang Dai, Bin Xiao, Lu Yuan, and Jianfeng Gao. 2021 · 2021
Later among the works it cites.
Point transformer. In ICCV
Hengshuang Zhao, Li Jiang, Jiaya Jia, Philip Torr, and Vladlen Koltun. 2021 · 2021
Later among the works it cites.
Scaling up Kernels in 3D CNNs. In NeurIPS
Yukang Chen, Jianhui Liu, Xiaojuan Qi, X. Zhang, Jian Sun, and Jiaya Jia. 2022 · 2022
Later among the works it cites.
CSWin Transformer: A general vision transformer backbone with cross-shaped windows. In CVPR
Xiaoyi Dong, Jianmin Bao, Dongdong Chen, Weiming Zhang, Nenghai Yu, Lu Yuan, Dong Chen, and Baining Guo. 2022 · 2022
Later among the works it cites.
Embracing single stride 3D object detector with sparse transformer. In CVPR
Lue Fan, Ziqi Pang, Tianyuan Zhang, Yu-Xiong Wang, Hang Zhao, Feng Wang, Naiyan Wang, and Zhaoxiang Zhang. 2022 · 2022
Later among the works it cites.
Dilated neighborhood attention transformer
Ali Hassani and Humphrey Shi. 2022 · 2022
Later among the works it cites.
Stratified Transformer for 3D Point Cloud Segmentation. In CVPR
Xin Lai, Jianhui Liu, Li Jiang, Liwei Wang, Hengshuang Zhao, Shu Liu, Xiaojuan Qi, and Jiaya Jia. 2022 · 2022
Later among the works it cites.
Rethinking network design and local geometry in point cloud: A simple residual MLP framework. In ICLR
Xu Ma, Can Qin, Haoxuan You, Haoxi Ran, and Yun Fu. 2022 · 2022
Later among the works it cites.
Masked autoencoders for point cloud self-supervised learning. In ECCV
Yatian Pang, Wenxiao Wang, Francis EH Tay, Wei Liu, Yonghong Tian, and Li Yuan. 2022 · 2022
Later among the works it cites.
Fast Point Transformer. In CVPR
Chunghyun Park, Yoonwoo Jeong, Minsu Cho, and Jaesik Park. 2022 · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with CLIP latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Later among the works it cites.
Language-Grounded Indoor 3D Semantic Segmentation in the Wild. In ECCV
David Rozenberszki, Or Litany, and Angela Dai. 2022 · 2022
Later among the works it cites.
FCAF3D: fully convolutional anchor-free 3D object detection. In ECCV
Danila Rukhovich, Anna Vorontsova, and Anton Konushin. 2022 · 2022
Later among the works it cites.
SWFormer: Sparse Window Transformer for 3D Object Detection in Point Clouds. In ECCV
Pei Sun, Mingxing Tan, Weiyue Wang, Chenxi Liu, Fei Xia, Zhaoqi Leng, and Dragomir Anguelov. 2022 · 2022
Later among the works it cites.
MaxViT: Multi-Axis Vision Transformer. In ECCV
Zhengzhong Tu, Hossein Talebi, Han Zhang, Feng Yang, Peyman Milanfar, Alan Bovik, and Yinxiao Li. 2022 · 2022
Later among the works it cites.
PVT v2: Improved baselines with pyramid vision transformer
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. 2022b · 2022
Later among the works it cites.
Point Transformer V2: Grouped Vector Attention and Partition-based Pooling. In NeurIPS
Xiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu, and Hengshuang Zhao. 2022 · 2022
Later among the works it cites.
Point-BERT: Pre-training 3D point cloud transformers with masked point modeling. In CVPR
Xumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang, Jie Zhou, and Jiwen Lu. 2022 · 2022
Later among the works it cites.