Fetching the paper…
Reading the bibliography…
We propose DOME, a diffusion-based world model that predicts future occupancy frames based on past occupancy observations.
A formal basis for the heuristic determination of minimum cost paths
Peter Hart, Nils Nilsson, and Bertram Raphael · 1968
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Diederik P. Kingma and Max Welling · 2013
Earlier work this paper cites.
The Lovász-Softmax loss: A tractable surrogate for the optimization of the intersection-over-union measure in neural networks
Maxim Berman, Amal Rannen Triki, and Matthew B Blaschko · 2018
Earlier work this paper cites.
David Ha and Jürgen Schmidhuber · 2018
Earlier work this paper cites.
nuScenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom · 2019
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Earlier work this paper cites.
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Earlier work this paper cites.
Lift, Splat, Shoot: Encoding Images From Arbitrary Camera Rigs by Implicitly Unprojecting to 3D, August 2020
Jonah Philion and Sanja Fidler · 2020
Earlier work this paper cites.
LidarDM: Generative LiDAR Simulation in a Generated World, April 2024
Vlas Zyrianov, Henry Che, Zhijian Liu, and Shenlong Wang · 2020
Earlier work this paper cites.
Einops: Clear and reliable tensor manipulations with einstein-like notation
Alex Rogozhnikov · 2022
Cited alongside, same era.
Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models, April 2023
Andreas Blattmann, Robin Rombach, Huan Ling, Tim Dockhorn, Seung Wook Kim, Sanja Fidler, and Karsten Kreis · 2023
Cited alongside, same era.
Tri-Perspective View for Vision-Based 3D Semantic Occupancy Prediction, March 2023
Yuanhui Huang, Wenzhao Zheng, Yunpeng Zhang, Jie Zhou, and Jiwen Lu · 2023
Cited alongside, same era.
Cam4DOcc: Benchmark for Camera-Only 4D Occupancy Forecasting in Autonomous Driving Applications, December 2023
Junyi Ma, Xieyuanli Chen, Jiawei Huang, Jingyi Xu, Zhen Luo, Jintao Xu, Weihao Gu, Rui Ai, and Hesheng Wang · 2023
Cited alongside, same era.
Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving, December 2023
Xiaoyu Tian, Tao Jiang, Longfei Yun, Yucheng Mao, Huitong Yang, Yue Wang, Yilun Wang, and Hang Zhao · 2023
Cited alongside, same era.
End-to-end autonomous driving: Challenges and frontiers, 2024
Li Chen, Penghao Wu, Kashyap Chitta, Bernhard Jaeger, Andreas Geiger, and Hongyang Li · 2024
Closest in time.
Latte: Latent Diffusion Transformer for Video Generation, January 2024
Xin Ma, Yaohui Wang, Gengyun Jia, Xinyuan Chen, Ziwei Liu, Yuan-Fang Li, Cunjian Chen, and Yu Qiao · 2024
Closest in time.
Text2Street: Controllable Text-to-image Generation for Street Views, February 2024
Jinming Su, Songen Gu, Yiting Duan, Xingyue Chen, and Junfeng Luo · 2024
Closest in time.
OccSora: 4D Occupancy Generation Models as World Simulators for Autonomous Driving, May 2024
Lening Wang, Wenzhao Zheng, Yilong Ren, Han Jiang, Zhiyong Cui, Haiyang Yu, and Jiwen Lu · 2024
Closest in time.
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving, September 2024
Julong Wei, Shanshuai Yuan, Pengfei Li, Qingda Hu, Zhongxue Gan, and Wenchao Ding · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Visual Point Cloud Forecasting enables Scalable Autonomous Driving, December 2023
Zetong Yang, Li Chen, Yanan Sun, and Hongyang Li · 2023
Cited alongside, same era.
OccFormer: Dual-path Transformer for Vision-based 3D Semantic Occupancy Prediction, April 2023
Yunpeng Zhang, Zheng Zhu, and Dalong Du · 2023
Cited alongside, same era.
OccWorld: Learning a 3D Occupancy World Model for Autonomous Driving, November 2023
Wenzhao Zheng, Weiliang Chen, Yuanhui Huang, Borui Zhang, Yueqi Duan, and Jiwen Lu · 2023
Cited alongside, same era.
PointOcc: Cylindrical Tri-Perspective View for Point-based 3D Semantic Occupancy Prediction, August 2023
Sicheng Zuo, Wenzhao Zheng, Yuanhui Huang, Jie Zhou, and Jiwen Lu · 2023
Cited alongside, same era.
GAIA-1: A Generative World Model for Autonomous Driving, September 2023a
Anthony Hu, Lloyd Russell, Hudson Yeo, Zak Murez, George Fedoseev, Alex Kendall, Jamie Shotton, and Gianluca Corrado
Cited in the paper.
Planning-oriented autonomous driving, 2023b
Yihan Hu, Jiazhi Yang, Li Chen, Keyu Li, Chonghao Sima, Xizhou Zhu, Siqi Chai, Senyao Du, Tianwei Lin, Wenhai Wang, Lewei Lu, Xiaosong Jia, Qiang Liu, Jifeng Dai, Yu Qiao, and Hongyang Li
Cited in the paper.
VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion, March 2023a
Yiming Li, Zhiding Yu, Christopher Choy, Chaowei Xiao, Jose M. Alvarez, Sanja Fidler, Chen Feng, and Anima Anandkumar
Cited in the paper.
Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion, April 2024
Lunjun Zhang, Yuwen Xiong, Ze Yang, Sergio Casas, Rui Hu, and Raquel Urtasun · 2024
Closest in time.
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation, March 2024
Guosheng Zhao, Xiaofeng Wang, Zheng Zhu, Xinze Chen, Guan Huang, Xiaoyi Bao, and Xingang Wang · 2024
Closest in time.
MonoOcc: Digging into Monocular Semantic Occupancy Prediction, March 2024
Yupeng Zheng, Xiang Li, Pengfei Li, Yuhang Zheng, Bu Jin, Chengliang Zhong, Xiaoxiao Long, Hao Zhao, and Qichao Zhang · 2024
Closest in time.