Fetching the paper…
Reading the bibliography…
This paper investigates the potential of enhancing Neural Radiance Fields (NeRF) with semantics to expand their applications.
Adam: A Method for Stochastic Optimization. In ICLR
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation. In CVPR . 3431–3440
Jonathan Long, Evan Shelhamer, and Trevor Darrell. 2015 · 2015
Earlier work this paper cites.
Deep Interactive Object Selection. In CVPR . 373–381
Ning Xu, Brian L. Price, Scott Cohen, Jimei Yang, and Thomas S. Huang. 2016 · 2016
Earlier work this paper cites.
Feature pyramid networks for object detection. In CVPR . 2117–2125
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie. 2017 · 2017
Earlier work this paper cites.
Attention is All You Need. In NeurIPS . 6000–6010
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Recurrent Slice Networks for 3D Segmentation of Point Clouds. In CVPR . 2626–2635
Qiangui Huang, Weiyue Wang, and Ulrich Neumann. 2018 · 2018
Earlier work this paper cites.
Local Light Field Fusion: Practical View Synthesis with Prescriptive Sampling Guidelines
Ben Mildenhall, Pratul P. Srinivasan, Rodrigo Ortiz-Cayon, Nima Khademi Kalantari, Ravi Ramamoorthi, Ren Ng, and Abhishek Kar. 2019 · 2019
Earlier work this paper cites.
Learning Object Bounding Boxes for 3D Instance Segmentation on Point Clouds. In NeurIPS . 6737–6746
Bo Yang, Jianan Wang, Ronald Clark, Qingyong Hu, Sen Wang, Andrew Markham, and Niki Trigoni. 2019 · 2019
Earlier work this paper cites.
End-to-End Object Detection with Transformers. In ECCV . 213–229
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko. 2020 · 2020
Earlier work this paper cites.
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis. In ECCV . 405–421
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng. 2020 · 2020
Earlier work this paper cites.
Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields. In ICCV . 5855–5864
Jonathan T Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman, Ricardo Martin-Brualla, and Pratul P Srinivasan. 2021 · 2021
Earlier work this paper cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In ICLR
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby. 2021 · 2021
Earlier work this paper cites.
CodeNeRF: Disentangled Neural Radiance Fields for Object Categories. In ICCV . 12929–12938
Wonbong Jang and Lourdes Agapito. 2021 · 2021
Earlier work this paper cites.
Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision. In ICML . 4904–4916
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc V. Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig. 2021 · 2021
Earlier work this paper cites.
BARF: Bundle-Adjusting Neural Radiance Fields. In ICCV . 5721–5731
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, and Simon Lucey. 2021 · 2021
Earlier work this paper cites.
Conditional detr for fast training convergence. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 3651–3660
Depu Meng, Xiaokang Chen, Zejia Fan, Gang Zeng, Houqiang Li, Yuhui Yuan, Lei Sun, and Jingdong Wang. 2021 · 2021
Earlier work this paper cites.
GIRAFFE: Representing Scenes As Compositional Generative Neural Feature Fields. In CVPR . 11453–11464
Michael Niemeyer and Andreas Geiger. 2021 · 2021
Earlier work this paper cites.
D-NeRF: Neural Radiance Fields for Dynamic Scenes. In CVPR . 10318–10327
Albert Pumarola, Enric Corona, Gerard Pons-Moll, and Francesc Moreno-Noguer. 2021 · 2021
Earlier work this paper cites.
DenseCLIP: Language-Guided Dense Prediction with Context-Aware Prompting. In CVPR . 18061–18070
Yongming Rao, Wenliang Zhao, Guangyi Chen, Yansong Tang, Zheng Zhu, Guan Huang, Jie Zhou, and Jiwen Lu. 2021 · 2021
Earlier work this paper cites.
NeuralDiff: Segmenting 3D objects that move in egocentric videos. In 3DV
Vadim Tschernezki, Diane Larlus, and Andrea Vedaldi. 2021 · 2021
Earlier work this paper cites.
pixelNeRF: Neural Radiance Fields From One or Few Images. In CVPR . 4578–4587
Alex Yu, Vickie Ye, Matthew Tancik, and Angjoo Kanazawa. 2021 · 2021
Earlier work this paper cites.
In-Place Scene Labelling and Understanding with Implicit Scene Representation. In ICCV . 15818–15827
Shuaifeng Zhi, Tristan Laidlow, Stefan Leutenegger, and Andrew J. Davison. 2021 · 2021
Earlier work this paper cites.
Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance Fields. In CVPR . 5460–5469
Jonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan, and Peter Hedman. 2022 · 2022
Earlier work this paper cites.
Real-Time Neural Light Field on Mobile Devices
Junli Cao, Huan Wang, Pavlo Chemerys, Vladislav Shakhrai, Ju Hu, Yun Fu, Denys Makoviichuk, Sergey Tulyakov, and Jian Ren. 2022 · 2022
Cited alongside, same era.
Group detr: Fast detr training with group-wise one-to-many assignment
Qiang Chen, Xiaokang Chen, Jian Wang, Haocheng Feng, Junyu Han, Errui Ding, Gang Zeng, and Jingdong Wang. 2022b · 2022
Cited alongside, same era.
Group detr v2: Strong object detector with encoder-decoder pretraining
Qiang Chen, Jian Wang, Chuchu Han, Shan Zhang, Zexian Li, Xiaokang Chen, Jiahui Chen, Xiaodi Wang, Shuming Han, Gang Zhang, et al · 2022
Cited alongside, same era.
D 3 ETR: Decoder Distillation for Detection Transformer
Xiaokang Chen, Jiahui Chen, Yan Liu, and Gang Zeng. 2022a · 2022
Cited alongside, same era.
Panoptic Lifting for 3D Scene Understanding with Neural Fields
Yawar Siddiqui, Lorenzo Porzi, Samuel Rota Bulò, Norman Müller, Matthias Nießner, Angela Dai, and Peter Kontschieder. 2022 · 2022
Later among the works it cites.
Direct Voxel Grid Optimization: Super-fast Convergence for Radiance Fields Reconstruction. In CVPR . 5449–5459
Cheng Sun, Min Sun, and Hwann-Tzong Chen. 2022 · 2022
Later among the works it cites.
Point scene understanding via disentangled instance mesh reconstruction. In Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXXII . Springer, 684–701
Jiaxiang Tang, Xiaokang Chen, Jingbo Wang, and Gang Zeng. 2022c · 2022
Later among the works it cites.
Real-time Neural Radiance Talking Portrait Synthesis via Audio-spatial Decomposition
Jiaxiang Tang, Kaisiyuan Wang, Hang Zhou, Xiaokang Chen, Dongliang He, Tianshu Hu, Jingtuo Liu, Gang Zeng, and Jingdong Wang. 2022d · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xiaokang Chen, Mingyu Ding, Xiaodi Wang, Ying Xin, Shentong Mo, Yunhao Wang, Shumin Han, Ping Luo, Gang Zeng, and Jingdong Wang. 2022c · 2022
Cited alongside, same era.
Conditional detr v2: Efficient detection transformer with box queries
Xiaokang Chen, Fangyun Wei, Gang Zeng, and Jingdong Wang. 2022f · 2022
Cited alongside, same era.
Zhiqin Chen, Thomas Funkhouser, Peter Hedman, and Andrea Tagliasacchi. 2022d · 2022
Cited alongside, same era.
Masked-attention Mask Transformer for Universal Image Segmentation. In CVPR . 1280–1289
Bowen Cheng, Ishan Misra, Alexander G. Schwing, Alexander Kirillov, and Rohit Girdhar. 2022 · 2022
Cited alongside, same era.
NeRF-SOS: Any-View Self-supervised Object Segmentation on Complex Scenes
Zhiwen Fan, Peihao Wang, Yifan Jiang, Xinyu Gong, Dejia Xu, and Zhangyang Wang. 2022 · 2022
Cited alongside, same era.
Panoptic NeRF: 3D-to-2D Label Transfer for Panoptic Urban Scene Segmentation. In 3DV
Xiao Fu, Shangzhan Zhang, Tianrun Chen, Yichong Lu, Lanyun Zhu, Xiaowei Zhou, Andreas Geiger, and Yiyi Liao. 2022 · 2022
Cited alongside, same era.
Scaling Open-Vocabulary Image Segmentation with Image-Level Labels. In ECCV . 540–557
Golnaz Ghiasi, Xiuye Gu, Yin Cui, and Tsung-Yi Lin. 2022 · 2022
Cited alongside, same era.
Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-Labeling. In CVPR . 7010–7021
Dat Huynh, Jason Kuen, Zhe Lin, Jiuxiang Gu, and Ehsan Elhamifar. 2022 · 2022
Cited alongside, same era.
SoftGroup for 3D Instance Segmentation on Point Clouds. In CVPR . 2698–2707
Thang Vu, Kookhoi Kim, Tung Minh Luu, Thanh Nguyen, and Chang D. Yoo. 2022 · 2022
Later among the works it cites.
GroupViT: Semantic Segmentation Emerges from Text Supervision. In CVPR . 18113–18123
Jiarui Xu, Shalini De Mello, Sifei Liu, Wonmin Byeon, Thomas M. Breuel, Jan Kautz, and Xiaolong Wang. 2022 · 2022
Later among the works it cites.
GIRAFFE HD: A High-Resolution 3D-aware Generative Model. In CVPR . 18419–18428
Yang Xue, Yuheng Li, Krishna Kumar Singh, and Yong Jae Lee. 2022 · 2022
Later among the works it cites.
SegNeRF: 3D Part Segmentation with Neural Radiance Fields
Jesus Zarzar, Sara Rojas, Silvio Giancola, and Bernard Ghanem. 2022 · 2022
Later among the works it cites.
CAE v2: Context Autoencoder with CLIP Target
Xinyu Zhang, Jiahui Chen, Junkun Yuan, Qiang Chen, Jian Wang, Xiaodi Wang, Shumin Han, Xiaokang Chen, Jimin Pi, Kun Yao, et al · 2022
Later among the works it cites.
Generalized Decoding for Pixel, Image, and Language
Xueyan Zou, Zi-Yi Dou, Jianwei Yang, Zhe Gan, Linjie Li, Chunyuan Li, Xiyang Dai, Harkirat Singh Behl, Jianfeng Wang, Lu Yuan, Nanyun Peng, Lijuan Wang, Yong Jae Lee, and Jianfeng Gao. 2022 · 2022
Later among the works it cites.
Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields
Jonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan, and Peter Hedman. 2023 · 2023
Closest in time.
NoPe-NeRF: Optimising Neural Radiance Field with No Pose Prior. In CVPR
Wenjing Bian, Zirui Wang, Kejie Li, Jiawang Bian, and Victor Adrian Prisacariu. 2023 · 2023
Closest in time.
PLA: Language-Driven Open-Vocabulary 3D Scene Understanding. In CVPR . 7010–7019
Runyu Ding, Jihan Yang, Chuhui Xue, Wenqing Zhang, Song Bai, and Xiaojuan Qi. 2023 · 2023
Closest in time.
Interactive Segmentation of Radiance Fields. In CVPR
Rahul Goel, Dhawal Sirikonda, Saurabh Saini, and P.J. Narayanan. 2023 · 2023
Closest in time.
LERF: Language Embedded Radiance Fields
Justin Kerr, Chung Min Kim, Ken Goldberg, Angjoo Kanazawa, and Matthew Tancik. 2023 · 2023
Closest in time.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollár, and Ross B. Girshick. 2023 · 2023
Closest in time.
MERF: Memory-Efficient Radiance Fields for Real-time View Synthesis in Unbounded Scenes
Christian Reiser, Richard Szeliski, Dor Verbin, Pratul P. Srinivasan, Ben Mildenhall, Andreas Geiger, Jonathan T. Barron, and Peter Hedman. 2023 · 2023
Closest in time.
Delicate Textured Mesh Recovery from NeRF via Adaptive Surface Refinement
Jiaxiang Tang, Hang Zhou, Xiaokang Chen, Tianshu Hu, Errui Ding, Jingdong Wang, and Gang Zeng. 2023 · 2023
Closest in time.
VisionLLM: Large Language Model is also an Open-Ended Decoder for Vision-Centric Tasks
Wenhai Wang, Zhe Chen, Xiaokang Chen, Jiannan Wu, Xizhou Zhu, Gang Zeng, Ping Luo, Tong Lu, Jie Zhou, Yu Qiao, et al · 2023
Closest in time.
FreeNeRF: Improving Few-shot Neural Rendering with Free Frequency Regularization
Jiawei Yang, Marco Pavone, and Yue Wang. 2023 · 2023
Closest in time.
BakedSDF: Meshing Neural SDFs for Real-Time View Synthesis
Lior Yariv, Peter Hedman, Christian Reiser, Dor Verbin, Pratul P.Srinivasan, Richard Szeliski, Jonathan T.Barron, and Ben Mildenhall. 2023 · 2023
Closest in time.
MaskGroup: Hierarchical Point Grouping and Masking for 3D Instance Segmentation. In ICME
Min Zhong, Xinghao Chen, Xiaokang Chen, Gang Zeng, and Yunhe Wang. 2023 · 2023
Closest in time.
Segment Everything Everywhere All at Once
Xueyan Zou, Jianwei Yang, Hao Zhang, Feng Li, Linjie Li, Jianfeng Gao, and Yong Jae Lee. 2023 · 2023
Closest in time.