Fetching the paper…
Reading the bibliography…
We introduce UniOcc, a comprehensive, unified benchmark and toolkit for occupancy forecasting (i.e., predicting future occupancies based on historical information) and occupancy prediction (i.e., predicting current-frame occupancy from camera images.
Solving geometric problems with the rotating calipers
G. T. Toussaint · 1983
Earlier work this paper cites.
Carla: An open urban driving simulator
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom · 2019
Earlier work this paper cites.
Scaling laws for neural language models
J. Kaplan, S. McCandlish, T. Henighan, T. B. Brown, B. Chess, R. Child, S. Gray, A. Radford, J. Wu, and D. Amodei · 2020
Earlier work this paper cites.
Evolvegraph: Multi-agent trajectory prediction with dynamic relational reasoning
J. Li, F. Yang, M. Tomizuka, and C. Choi · 2020
Earlier work this paper cites.
Human motion trajectory prediction: A survey
A. Rudenko, L. Palmieri, M. Herman, K. M. Kitani, D. M. Gavrila, and K. O. Arras · 2020
Earlier work this paper cites.
Trajectron++: Dynamically-feasible trajectory forecasting with heterogeneous data
T. Salzmann, B. Ivanovic, P. Chakravarty, and M. Pavone · 2020
Earlier work this paper cites.
Scalability in perception for autonomous driving: Waymo open dataset
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, V. Vasudevan, W. Han, J. Ngiam, H. Zhao, A. Timofeev, S. Ettinger, M. Krivokon, A. Gao, A. Joshi, Y. Zhang, J. Shlens, Z. Chen, and D. Anguelov · 2020
Earlier work this paper cites.
Spectral temporal graph neural network for trajectory prediction
D. Cao, J. Li, H. Ma, and M. Tomizuka · 2021
Earlier work this paper cites.
Shared cross-modal trajectory prediction for autonomous driving
C. Choi, J. H. Choi, J. Li, and S. Malla · 2021
Earlier work this paper cites.
Loki: Long term and key intentions for trajectory prediction
H. Girase, H. Gang, S. Malla, J. Li, A. Kanehara, K. Mangalam, and C. Choi · 2021
Earlier work this paper cites.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
J. Huang, G. Huang, Z. Zhu, Y. Ye, and D. Du · 2021
Earlier work this paper cites.
Multi-agent driving behavior prediction across different scenarios with self-supervised domain knowledge
H. Ma, Y. Sun, J. Li, and M. Tomizuka · 2021
Earlier work this paper cites.
Continual multi-agent interaction behavior prediction with conditional generative memory
H. Ma, Y. Sun, J. Li, M. Tomizuka, and C. Choi · 2021
Earlier work this paper cites.
Autonomous driving strategies at intersections: Scenarios, state-of-the-art, and future outlooks
L. Wei, Z. Li, J. Gong, C. Gong, and J. Li · 2021
Earlier work this paper cites.
Point set voting for partial point cloud analysis
J. Zhang, W. Chen, Y. Wang, R. Vasudevan, and M. Johnson-Roberson · 2021
Earlier work this paper cites.
A survey on trajectory-prediction methods for autonomous driving
Y. Huang, J. Du, Z. Yang, Z. Zhou, L. Zhang, and H. Chen · 2022
Earlier work this paper cites.
Z. Li, W. Wang, H. Li, E. Xie, C. Sima, T. Lu, Y. Qiao, and J. Dai · 2022
Earlier work this paper cites.
Strajnet: Occupancy flow prediction via multi-modal swin transformer
H. Liu, Z. Huang, and C. Lv · 2022
Earlier work this paper cites.
Multi-objective diverse human motion prediction with knowledge distillation
H. Ma, J. Li, R. Hosseini, M. Tomizuka, and C. Choi · 2022
Earlier work this paper cites.
Occupancy flow fields for motion forecasting in autonomous driving
R. Mahjourian, J. Kim, Y. Chai, M. Tan, B. Sapp, and D. Anguelov · 2022
Earlier work this paper cites.
Interaction modeling with multiplex attention
F.-Y. Sun, I. Kauvar, R. Zhang, J. Li, M. J. Kochenderfer, J. Wu, and N. Haber · 2022
Earlier work this paper cites.
Dynamics-aware spatiotemporal occupancy prediction in urban environments
M. Toyungyernsub, E. Yel, J. Li, and M. J. Kochenderfer · 2022
Earlier work this paper cites.
Motionsc: Data set and network for real-time semantic mapping in dynamic environments
J. Wilson, J. Song, Y. Fu, A. Zhang, A. Capodieci, P. Jayakumar, K. Barton, and M. Ghaffari · 2022
Earlier work this paper cites.
Opv2v: An open benchmark dataset and fusion pipeline for perception with vehicle-to-vehicle communication
R. Xu, H. Xiang, X. Xia, X. Han, J. Li, and J. Ma · 2022
Earlier work this paper cites.
Cobevt: Cooperative bird’s eye view semantic segmentation with sparse transformers
R. X. Xu, Z. Tu, H. Xiang, W. Shao, B. Zhou, and J. Ma · 2022
Earlier work this paper cites.
Grouptron: Dynamic multi-scale graph convolutional networks for group-aware dense crowd trajectory forecasting
R. Zhou, H. Zhou, H. Gao, M. Tomizuka, J. Li, and Z. Xu · 2022
Earlier work this paper cites.
Hivt: Hierarchical vector transformer for multi-agent motion prediction
Z. Zhou, L. Ye, J. Wang, K. Wu, and K. Lu · 2022
Earlier work this paper cites.
Disentangled neural relational inference for interpretable motion prediction
V. M. Dax, J. Li, E. Sachdeva, N. Agarwal, and M. J. Kochenderfer · 2023
Earlier work this paper cites.
Novel view synthesis from a single rgbd image for indoor scenes
C. Hetang and Y. Wang · 2023
Cited alongside, same era.
Tri-perspective view for vision-based 3d semantic occupancy prediction
Y. Huang, W. Zheng, Y. Zhang, J. Zhou, and J. Lu · 2023
Cited alongside, same era.
Pedestrian crossing action recognition and trajectory prediction with 3d human keypoints
J. Li, X. Shi, F. Chen, J. Stroud, Z. Zhang, T. Lan, J. Mao, J. Kang, K. S. Refaat, W. Yang, et al · 2023
Cited alongside, same era.
Game theory-based simultaneous prediction and planning for autonomous vehicle navigation in crowded environments
K. Li, Y. Chen, M. Shan, J. Li, S. Worrall, and E. Nebot · 2023
Cited alongside, same era.
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Y. Li, Z. Ge, G. Yu, J. Yang, Z. Wang, Y. Shi, J. Sun, and Z. Li · 2023
Cited alongside, same era.
Occsora: 4d occupancy generation models as world simulators for autonomous driving
L. Wang, W. Zheng, Y. Ren, H. Jiang, Z. Cui, H. Yu, and J. Lu · 2024
Later among the works it cites.
Occllama: An occupancy-language-action generative world model for autonomous driving
J. Wei, S. Yuan, P. Li, Q. Hu, Z. Gan, and W. Ding · 2024
Later among the works it cites.
Autotrust: Benchmarking trustworthiness in large vision language models for autonomous driving
S. Xing, H. Hua, X. Gao, S. Zhu, R. Li, K. Tian, X. Li, H. Huang, T. Yang, Z. Wang, et al · 2024
Later among the works it cites.
Matrix: Multi-agent trajectory generation with diverse contexts
Z. Xu, R. Zhou, Y. Yin, H. Gao, M. Tomizuka, and J. Li · 2024
Later among the works it cites.
Cvt-occ: Cost volume temporal fusion for 3d occupancy prediction
Z. Ye, T. Jiang, C. Xu, Y. Li, and H. Zhao · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Lin, Y. Yue, S. Hou, X. Yu, Y. Xu, K. D. Yamada, and Z. Zhang · 2023
Cited alongside, same era.
Infocd: A contrastive chamfer distance loss for point cloud completion
F. Lin, Y. Yue, Z. Zhang, S. Hou, K. Yamada, V. Kolachalama, and V. Saligrama · 2023
Cited alongside, same era.
Multi-modal hierarchical transformer for occupancy flow field prediction in autonomous driving
H. Liu, Z. Huang, and C. Lv · 2023
Cited alongside, same era.
Scene as occupancy
W. Tong, C. Sima, T. Wang, L. Chen, S. Wu, H. Deng, Y. Gu, L. Lu, P. Luo, D. Lin, et al · 2023
Cited alongside, same era.
Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception
X. Wang, Z. Zhu, W. Xu, Y. Zhang, Y. Wei, X. Chi, Y. Ye, D. Du, J. Lu, and X. Wang · 2023
Cited alongside, same era.
Eqdrive: Efficient equivariant motion forecasting with multi-modality for autonomous driving
Y. Wang and J. Chen · 2023
Cited alongside, same era.
Equivariant map and agent geometry for autonomous driving motion prediction
Y. Wang and J. Chen · 2023
Cited alongside, same era.
Later among the works it cites.
Vision-based 3d occupancy prediction in autonomous driving: a review and outlook
Y. Zhang, J. Zhang, Z. Wang, J. Xu, and D. Huang · 2024
Later among the works it cites.
Lowrankocc: tensor decomposition and low-rank recovery for vision-based 3d semantic occupancy prediction
L. Zhao, X. Xu, Z. Wang, Y. Zhang, B. Zhang, W. Zheng, D. Du, J. Zhou, and J. Lu · 2024
Later among the works it cites.
Dynamiccity: Large-scale occupancy generation from dynamic scenes
H. Bian, L. Kong, H. Xie, L. Pan, Y. Qiao, and Z. Liu · 2025
Closest in time.
Unitraj: A unified framework for scalable vehicle trajectory prediction
L. Feng, M. Bahari, K. M. B. Amor, É. Zablocki, M. Cord, and A. Alahi · 2025
Closest in time.
Airv2x: Unified air-ground vehicle-to-everything collaboration
X. Gao, Y. Wu, X. Luo, K. Wu, X. Chen, Y. Wang, C. Liu, Y. Zhou, and Z. Tu · 2025
Closest in time.
Langcoop: Collaborative driving with language
X. Gao, Y. Wu, R. Wang, C. Liu, Y. Zhou, and Z. Tu · 2025
Closest in time.
Lp-detr: Layer-wise progressive relation for object detection
Z. Kang, Y. Zhang, X. Deng, X. Li, and Y. Zhang · 2025
Closest in time.
Self-supervised multi-future occupancy forecasting for autonomous driving
B. Lange, M. Itkina, J. Li, and M. J. Kochenderfer · 2025
Closest in time.
Safeflow: A principled protocol for trustworthy and transactional autonomous agent systems
P. Li, X. Zou, Z. Wu, R. Li, S. Xing, H. Zheng, Z. Hu, Y. Wang, H. Li, Q. Yuan, et al · 2025
Closest in time.
A strong view-free baseline approach for single-view image guided point cloud completion
F. Lin, Z. Dai, R. Sanku, S. Hou, K. D. Yamada, H. K. Zhang, and Z. Zhang · 2025
Closest in time.
Let occ flow: Self-supervised 3d occupancy flow prediction
Y. Liu, L. Mou, X. Yu, C. Han, S. Mao, R. Xiong, and Y. Wang · 2025
Closest in time.
Position: Prospective of autonomous driving-multimodal llms world models embodied intelligence ai alignment and mamba
Y. Ma, W. Ye, C. Cui, H. Zhang, S. Xing, F. Ke, J. Wang, C. Miao, J. Chen, H. Rezatofighi, et al · 2025
Closest in time.
Trends in motion prediction toward deployable and generalizable autonomy: A revisit and perspectives
L. Wang, M.-A. Lavoie, S. Papais, B. Nisar, Y. Chen, W. Ding, B. Ivanovic, H. Shao, A. Abuduweili, E. Cook, et al · 2025
Closest in time.
Generative ai for autonomous driving: Frontiers and opportunities
Y. Wang, S. Xing, C. Can, R. Li, H. Hua, K. Tian, Z. Mo, X. Gao, K. Wu, S. Zhou, et al · 2025
Closest in time.
Cmp: Cooperative motion prediction with multi-agent communication
Z. Wang, Y. Wang, Z. Wu, H. Ma, Z. Li, H. Qiu, and J. Li · 2025
Closest in time.
Demystifying the visual quality paradox in multimodal large language models
S. Xing, L. Guo, H. Hua, S. Lee, P. Li, Y. Wang, Z. Wang, and Z. Tu · 2025
Closest in time.
Openemma: Open-source multimodal model for end-to-end autonomous driving
S. Xing, C. Qian, Y. Wang, H. Hua, K. Tian, Y. Zhou, and Z. Tu · 2025
Closest in time.
Can large vision language models read maps like a human?
S. Xing, Z. Sun, S. Xie, K. Chen, Y. Huang, Y. Wang, J. Li, D. Song, and Z. Tu · 2025
Closest in time.
Re-align: Aligning vision language models via retrieval-augmented direct preference optimization
S. Xing, Y. Wang, P. Li, R. Bai, Y. Wang, C. Qian, H. Yao, and Z. Tu · 2025
Closest in time.
Towards generalizable safety in crowd navigation via conformal uncertainty handling
J. Yao, X. Zhang, Y. Xia, A. K. Roy-Chowdhury, and J. Li · 2025
Closest in time.
Gps: A probabilistic distributional similarity with gumbel priors for set-to-set matching
Z. Zhang, F. Lin, H. Liu, J. Morales, H. Zhang, K. Yamada, V. B. Kolachalama, and V. Saligrama · 2025
Closest in time.
Trajevo: Trajectory prediction heuristics design via llm-driven evolution
Z. Zhao, C. Hua, F. Berto, K. Lee, Z. Ma, J. Li, and J. Park · 2025
Closest in time.
Occworld: Learning a 3d occupancy world model for autonomous driving
W. Zheng, W. Chen, Y. Huang, B. Zhang, Y. Duan, and J. Lu · 2025
Closest in time.
J. Zhuang, H. Jin, Y. Zhang, Z. Kang, W. Zhang, G. G. Dagher, and H. Wang · 2025
Closest in time.