Fetching the paper…
Reading the bibliography…
We propose DrivingForward, a feed-forward Gaussian Splatting model that reconstructs driving scenes from flexible surround-view input.
Image quality assessment: from error visibility to structural similarity
Wang, Z.; Bovik, A. C.; Sheikh, H. R.; and Simoncelli, E. P. 2004 · 2004
Earlier work this paper cites.
Spatial transformer networks
Jaderberg, M.; Simonyan, K.; Zisserman, A.; et al. 2015 · 2015
Earlier work this paper cites.
Unsupervised monocular depth estimation with left-right consistency
Godard, C.; Mac Aodha, O.; and Brostow, G. J. 2017 · 2017
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Zhang, R.; Isola, P.; Efros, A. A.; Shechtman, E.; and Wang, O. 2018 · 2018
Earlier work this paper cites.
Digging into self-supervised monocular depth estimation
Godard, C.; Mac Aodha, O.; Firman, M.; and Brostow, G. J. 2019 · 2019
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
Caesar, H.; Bankiti, V.; Lang, A. H.; Vora, S.; Liong, V. E.; Xu, Q.; Krishnan, A.; Pan, Y.; Baldan, G.; and Beijbom, O. 2020 · 2020
Earlier work this paper cites.
Drivinggaussian: Composite gaussian splatting for surrounding dynamic autonomous driving scenes
Zhou, X.; Lin, Z.; Shan, X.; Wang, Y.; Sun, D.; and Yang, M.-H. 2024 · 2020
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2021 · 2021
Earlier work this paper cites.
pixelnerf: Neural radiance fields from one or few images
Yu, A.; Ye, V.; Tancik, M.; and Kanazawa, A. 2021 · 2021
Earlier work this paper cites.
Full surround monodepth from multiple cameras
Guizilini, V.; Vasiljevic, I.; Ambrus, R.; Shakhnarovich, G.; and Gaidon, A. 2022 · 2022
Earlier work this paper cites.
Self-supervised surround-view depth estimation with volumetric feature fusion
Kim, J.-H.; Hur, J.; Nguyen, T. P.; and Jeong, S.-G. 2022 · 2022
Earlier work this paper cites.
Bevfusion: A simple and robust lidar-camera fusion framework
Liang, T.; Xie, H.; Yu, K.; Xia, Z.; Lin, Z.; Wang, Y.; Tang, T.; Wang, B.; and Tang, Z. 2022 · 2022
Cited alongside, same era.
Generalizable patch-based neural rendering
Suhail, M.; Esteves, C.; Sigal, L.; and Makadia, A. 2022 · 2022
Cited alongside, same era.
Objectfusion: Multi-modal 3d object detection with object-centric fusion
Cai, Q.; Pan, Y.; Yao, T.; Ngo, C.-W.; and Mei, T. 2023 · 2023
Cited alongside, same era.
Futr3d: A unified sensor fusion framework for 3d detection
Chen, X.; Zhang, T.; Wang, Y.; Wang, Y.; and Zhao, H. 2023 · 2023
Cited alongside, same era.
Learning to render novel views from wide-baseline stereo pairs
Du, Y.; Smith, C.; Tewari, A.; and Sitzmann, V. 2023 · 2023
Cited alongside, same era.
Streetsurf: Extending multi-view implicit surface reconstruction to street views
pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction
Charatan, D.; Li, S. L.; Tagliasacchi, A.; and Sitzmann, V. 2024 · 2024
Closest in time.
Mvsplat: Efficient 3d gaussian splatting from sparse multi-view images
Chen, Y.; Xu, H.; Zheng, C.; Zhuang, B.; Pollefeys, M.; Geiger, A.; Cham, T.-J.; and Cai, J. 2024 · 2024
Closest in time.
Benchmarking micro-action recognition: dataset, method, and application
Guo, D.; Li, K.; Hu, B.; Zhang, Y.; and Wang, M. 2024 · 2024
Closest in time.
FastLGS: Speeding up Language Embedded Gaussians with Feature Grid Mapping
Ji, Y.; Zhu, H.; Tang, J.; Liu, W.; Zhang, Z.; Xie, Y.; and Tan, X. 2024 · 2024
Closest in time.
AutoSplat: Constrained Gaussian Splatting for Autonomous Driving Scene Reconstruction
Khan, M.; Fazlali, H.; Sharma, D.; Cao, T.; Bai, D.; Ren, Y.; and Liu, B. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Guo, J.; Deng, N.; Li, X.; Bai, Y.; Shi, B.; Wang, C.; Ding, C.; Wang, D.; and Li, Y. 2023 · 2023
Cited alongside, same era.
3D Gaussian Splatting for Real-Time Radiance Field Rendering
Kerbl, B.; Kopanas, G.; Leimkühler, T.; and Drettakis, G. 2023 · 2023
Cited alongside, same era.
Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation
Liu, Z.; Tang, H.; Amini, A.; Yang, X.; Mao, H.; Rus, D. L.; and Han, S. 2023 · 2023
Cited alongside, same era.
Positive-negative receptive field reasoning for omni-supervised 3d segmentation
Tan, X.; Ma, Q.; Gong, J.; Xu, J.; Zhang, Z.; Song, H.; Qu, Y.; Xie, Y.; and Ma, L. 2023 · 2023
Cited alongside, same era.
Surrounddepth: Entangling surrounding views for self-supervised multi-camera depth estimation
Wei, Y.; Zhao, L.; Zheng, W.; Zhu, Z.; Rao, Y.; Huang, G.; Lu, J.; and Zhou, J. 2023 · 2023
Cited alongside, same era.
Sparsegs: Real-time 360 ° sparse view synthesis using gaussian splatting
Xiong, H.; Muttukuru, S.; Upadhyay, R.; Chari, P.; and Kadambi, A. 2023 · 2023
Cited alongside, same era.
S3Gaussian: Self-Supervised Street Gaussians for Autonomous Driving
Huang, N.; Wei, X.; Zheng, W.; An, P.; Lu, M.; Zhan, W.; Tomizuka, M.; Keutzer, K.; and Zhang, S. 2024a
Cited in the paper.
Closest in time.
Dngaussian: Optimizing sparse-view 3d gaussian radiance fields with global-local depth normalization
Li, J.; Zhang, J.; Bai, X.; Zheng, J.; Ning, X.; Zhou, J.; and Gu, L. 2024 · 2024
Closest in time.
Cotr: Compact occupancy transformer for vision-based 3d occupancy prediction
Ma, Q.; Tan, X.; Qu, Y.; Ma, L.; Zhang, Z.; and Xie, Y. 2024 · 2024
Closest in time.
Splatter image: Ultra-fast single-view 3d reconstruction
Szymanowicz, S.; Rupprecht, C.; and Vedaldi, A. 2024 · 2024
Closest in time.
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features
Wang, L.; Kim, S. W.; Yang, J.; Yu, C.; Ivanovic, B.; Waslander, S. L.; Wang, Y.; Fidler, S.; Pavone, M.; and Karkus, P. 2024 · 2024
Closest in time.
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
Yan, Y.; Lin, H.; Zhou, C.; Wang, W.; Sun, H.; Zhan, K.; Lang, X.; Zhou, X.; and Peng, S. 2024 · 2024
Closest in time.
Gps-gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis
Zheng, S.; Zhou, B.; Shao, R.; Liu, B.; Zhang, S.; Nie, L.; and Liu, Y. 2024 · 2024
Closest in time.