Fetching the paper…
Reading the bibliography…
End-to-end architectures in autonomous driving (AD) face a significant challenge in interpretability, impeding human-AI trust.
Alvinn: An autonomous land vehicle in a neural network
D. A. Pomerleau · 1988
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
C.-Y. Lin · 2004
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
S. Banerjee and A. Lavie · 2005
Earlier work this paper cites.
The structure and function of explanations
T. Lombrozo · 2006
Earlier work this paper cites.
Explanation and abductive inference
T. Lombrozo · 2012
Earlier work this paper cites.
A review of motion planning techniques for automated vehicles
D. González, J. Pérez, V. Milanés, and F. Nashashibi · 2015
Earlier work this paper cites.
Cider: Consensus-based image description evaluation
R. Vedantam, C. Lawrence Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Interpretable learning for self-driving cars by visualizing causal attention
J. Kim and J. Canny · 2017
Earlier work this paper cites.
Textual explanations for self-driving vehicles
J. Kim, A. Rohrbach, T. Darrell, J. Canny, and Z. Akata · 2018
Earlier work this paper cites.
A tale of two explanations: Enhancing human trust by explaining robot behavior
M. Edmonds, F. Gao, H. Liu, X. Xie, S. Qi, B. Rothrock, Y. Zhu, Y. N. Wu, H. Lu, and S.-C. Zhu · 2019
Earlier work this paper cites.
End-to-end interpretable neural motion planner
W. Zeng, W. Luo, S. Suo, A. Sadat, B. Yang, S. Casas, and R. Urtasun · 2019
Earlier work this paper cites.
Learning to drive in a day
A. Kendall, J. Hawke, D. Janz, P. Mazur, D. Reda, J.-M. Allen, V.-D. Lam, A. Bewley, and A. Shah · 2019
Earlier work this paper cites.
Deep modular co-attention networks for visual question answering
Z. Yu, J. Yu, Y. Cui, D. Tao, and Q. Tian · 2019
Earlier work this paper cites.
Learning by cheating
D. Chen, B. Zhou, V. Koltun, and P. Krähenbühl · 2020
Earlier work this paper cites.
Perceive, predict, and plan: Safe motion planning through interpretable semantic representations
A. Sadat, S. Casas, M. Ren, X. Wu, P. Dhawan, and R. Urtasun · 2020
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom · 2020
Earlier work this paper cites.
Bdd100k: A diverse driving dataset for heterogeneous multitask learning
F. Yu, H. Chen, X. Wang, W. Xian, Y. Chen, F. Liu, V. Madhavan, and T. Darrell · 2020
Cited alongside, same era.
Learning interpretable end-to-end vision-based motion planning for autonomous driving with optical flow distillation
H. Wang, P. Cai, Y. Sun, L. Wang, and M. Liu · 2021
Cited alongside, same era.
Lookout: Diverse multi-future prediction and planning for self-driving
A. Cui, S. Casas, A. Sadat, R. Liao, and R. Urtasun · 2021
Cited alongside, same era.
Autonomous vehicle motion planning via recurrent spline optimization
W. Xu, Q. Wang, and J. M. Dolan · 2021
Cited alongside, same era.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
J. Huang, G. Huang, Z. Zhu, Y. Ye, and D. Du · 2021
Cited alongside, same era.
Steps: Joint self-supervised nighttime image enhancement and depth estimation
Y. Zheng, C. Zhong, P. Li, H.-a. Gao, Y. Zheng, B. Jin, L. Wang, H. Zhao, G. Zhou, Q. Zhang, et al · 2023
Later among the works it cites.
Dqs3d: Densely-matched quantization-aware semi-supervised 3d detection
H.-a. Gao, B. Tian, P. Li, H. Zhao, and G. Zhou · 2023
Later among the works it cites.
Unsupervised road anomaly detection with language anchors
B. Tian, M. Liu, H.-a. Gao, P. Li, H. Zhao, and G. Zhou · 2023
Later among the works it cites.
Nle-dm: Natural-language explanations for decision making of autonomous driving based on semantic scene understanding
Y. Feng, W. Hua, and Y. Sun · 2023
Later among the works it cites.
W. Wang, J. Xie, C. Hu, H. Zou, J. Fan, W. Tong, Y. Wen, S. Wu, H. Deng, Z. Li, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Cited alongside, same era.
In situ bidirectional human-robot value alignment
L. Yuan, X. Gao, Z. Zheng, M. Edmonds, Y. N. Wu, F. Rossano, H. Lu, Y. Zhu, and S.-C. Zhu · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou, et al · 2022
Cited alongside, same era.
Video swin transformer
Z. Liu, J. Ning, Y. Cao, Y. Wei, Z. Zhang, S. Lin, and H. Hu · 2022
Cited alongside, same era.
Planning-oriented autonomous driving
Y. Hu, J. Yang, L. Chen, K. Li, C. Sima, X. Zhu, S. Chai, S. Du, T. Lin, W. Wang, et al · 2023
Cited alongside, same era.
Palm-e: An embodied multimodal language model
D. Driess, F. Xia, M. S. Sajjadi, C. Lynch, A. Chowdhery, B. Ichter, A. Wahid, J. Tompson, Q. Vuong, T. Yu, et al · 2023
Cited alongside, same era.
Adapt: Action-aware driving caption transformer
B. Jin, X. Liu, Y. Zheng, P. Li, H. Zhao, T. Zhang, Y. Zheng, G. Zhou, and J. Liu · 2023
Cited alongside, same era.
Drivegpt4: Interpretable end-to-end autonomous driving via large language model
Z. Xu, Y. Zhang, E. Xie, Z. Zhao, Y. Guo, K. K. Wong, Z. Li, and H. Zhao · 2023
Later among the works it cites.
Drive anywhere: Generalizable end-to-end autonomous driving with multi-modal foundation models
T.-H. Wang, A. Maalouf, W. Xiao, Y. Ban, A. Amini, G. Rosman, S. Karaman, and D. Rus · 2023
Later among the works it cites.
Llama-adapter v2: Parameter-efficient visual instruction model
P. Gao, J. Han, R. Zhang, Z. Lin, S. Geng, A. Zhou, W. Zhang, P. Lu, C. He, X. Yue, et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
H. Touvron, L. Martin, K. Stone, P. Albert, A. Almahairi, Y. Babaei, N. Bashlykov, S. Batra, P. Bhargava, S. Bhosale, et al · 2023
Later among the works it cites.
Llama-adapter: Efficient fine-tuning of language models with zero-init attention
R. Zhang, J. Han, C. Liu, P. Gao, A. Zhou, X. Hu, S. Yan, P. Lu, H. Li, and Y. Qiao · 2023
Later among the works it cites.
Tod3cap: Towards 3d dense captioning in outdoor scenes
B. Jin, Y. Zheng, P. Li, W. Li, Y. Zheng, S. Hu, X. Liu, J. Zhu, Z. Yan, H. Sun, et al · 2024
Closest in time.
Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario
T. Qian, J. Chen, L. Zhuo, Y. Jiao, and Y.-G. Jiang · 2024
Closest in time.
Drivevlm: The convergence of autonomous driving and large vision-language models
X. Tian, J. Gu, B. Li, Y. Liu, C. Hu, Y. Wang, K. Zhan, P. Jia, X. Lang, and H. Zhao · 2024
Closest in time.
A survey on multimodal large language models for autonomous driving
C. Cui, Y. Ma, X. Cao, W. Ye, Y. Zhou, K. Liang, J. Chen, J. Lu, Z. Yang, K.-D. Liao, et al · 2024
Closest in time.
P-mapnet: Far-seeing map generator enhanced by both sdmap and hdmap priors
Z. Jiang, Z. Zhu, P. Li, H.-a. Gao, T. Yuan, Y. Shi, H. Zhao, and H. Zhao · 2024
Closest in time.
Training-free model merging for multi-target domain adaptation
W. Li, H.-a. Gao, M. Gao, B. Tian, R. Zhi, and H. Zhao · 2024
Closest in time.
Vote2cap-detr++: Decoupling localization and describing for end-to-end 3d dense captioning
S. Chen, H. Zhu, M. Li, X. Chen, P. Guo, Y. Lei, Y. Gang, T. Li, and T. Chen · 2024
Closest in time.