Fetching the paper…
Reading the bibliography…
Language models uncover unprecedented abilities in analyzing driving scenarios, owing to their limitless knowledge accumulated from text-based pre-training.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y · 2004
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
Banerjee, S. and Lavie, A · 2005
Earlier work this paper cites.
Cider: Consensus-based image description evaluation
Vedantam, R., Lawrence Zitnick, C., and Parikh, D · 2015
Earlier work this paper cites.
Spice: Semantic propositional image caption evaluation
Anderson, P., He, X., Batra, D., Parikh, D., Lee, M., Ehsan, A., and Zitnick, C. L · 2016
Earlier work this paper cites.
End-to-end learning of driving models from large-scale video datasets, 2017
Xu, H., Gao, Y., Yu, F., and Darrell, T · 2017
Earlier work this paper cites.
Textual explanations for self-driving vehicles, 2018
Kim, J., Rohrbach, A., Darrell, T., Canny, J., and Akata, Z · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Radford, A., Narasimhan, K., Salimans, T., and Sutskever, I · 2018
Earlier work this paper cites.
Touchdown: Natural language navigation and spatial reasoning in visual street environments
Chen, H., Suhr, A., Misra, D., Snavely, N., and Artzi, Y · 2019
Earlier work this paper cites.
Talk2car: Taking control of your self-driving car
Deruyttere, T., Vandenhende, S., Grujicic, D., Van Gool, L., and Moens, M.-F · 2019
Earlier work this paper cites.
Talk to the vehicle: Language conditioned autonomous navigation of self driving cars
Sriram, N., Maniar, T., Kalyanasundaram, J., Gandhi, V., Bhowmick, B., and Krishna, K. M · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving, 2020
Caesar, H., Bankiti, V., Lang, A. H., Vora, S., Liong, V. E., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., and Beijbom, O · 2020
Earlier work this paper cites.
Conditional driving from natural language instructions
Roh, J., Paxton, C., Pronobis, A., Farhadi, A., and Fox, D · 2020
Earlier work this paper cites.
Large Scale Interactive Motion Forecasting for Autonomous Driving: The Waymo Open Motion Dataset, 2021
Ettinger, S., Cheng, S., Caine, B., Liu, C., Zhao, H., Pradhan, S., Chai, Y., Sapp, B., Qi, C., Zhou, Y., Yang, Z., Chouard, A., Sun, P., Ngiam, J., Vasudevan, V., McCauley, A., Shlens, J., and Anguelov, D · 2021
Earlier work this paper cites.
Multipath++: Efficient information fusion and trajectory aggregation for behavior prediction, 2021
Varadarajan, B., Hefny, A., Srivastava, A., Refaat, K. S., Nayakanti, N., Cornman, A., Chen, K., Douillard, B., Lam, C. P., Anguelov, D., and Sapp, B · 2021
Earlier work this paper cites.
Mpa: Multipath++ based architecture for motion prediction, 2022
Konev, S · 2022
Earlier work this paper cites.
Trajectory prediction with linguistic representations, 2022
Kuo, Y.-L., Huang, X., Barbu, A., McGill, S. G., Katz, B., Leonard, J. J., and Rosman, G · 2022
Earlier work this paper cites.
Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning
Li, Q., Peng, Z., Feng, L., Zhang, Q., Xue, Z., and Zhou, B · 2022
Earlier work this paper cites.
Causalagents: A robustness benchmark for motion forecasting using causal relationships, 2022
Roelofs, R., Sun, L., Caine, B., Refaat, K. S., Sapp, B., Ettinger, S., and Chai, W · 2022
Earlier work this paper cites.
Social interactions for autonomous driving: A review and perspectives
Wang, W., Wang, L., Zhang, C., Liu, C., Sun, L., et al · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Xia, F., Chi, E., Le, Q. V., Zhou, D., et al · 2022
Cited alongside, same era.
Dynamical driving interactions between human and mentalizing-designed autonomous vehicle
Zhang, Y., Zhang, S., Liang, Z., Li, H. L., Wu, H., and Liu, Q · 2022
Cited alongside, same era.
Driving with llms: Fusing object-level vector modality for explainable autonomous driving, 2023
Chen, L., Sinavski, O., Hünermann, J., Karnsund, A., Willmott, A. J., Birch, D., Maund, D., and Shotton, J · 2023
Cited alongside, same era.
A survey of chain of thought reasoning: Advances, frontiers and future
Chu, Z., Chen, J., Chen, Q., Yu, W., He, T., Wang, H., Peng, W., Liu, M., Qin, B., and Liu, T · 2023
Cited alongside, same era.
Incorporating driving knowledge in deep learning based vehicle trajectory prediction: A survey
Ding, Z. and Zhao, H · 2023
Cited alongside, same era.
Drive as you speak: Enabling human-like interaction with large language models in autonomous vehicles
Cui, C., Ma, Y., Cao, X., Ye, W., and Wang, Z · 2024
Closest in time.
Gai, X., Zhou, C., Liu, J., Feng, Y., Wu, J., and Liu, Z · 2024
Closest in time.
Greer, R. and Trivedi, M · 2024
Closest in time.
Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving
Han, W., Guo, D., Xu, C.-Z., and Shen, J · 2024
Closest in time.
Nuscenes-mqa: Integrated evaluation of captions and qa for autonomous driving datasets using markup annotations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Drive like a human: Rethinking autonomous driving with large language models, 2023
Fu, D., Li, X., Wen, L., Dou, M., Cai, P., Shi, B., and Qiao, Y · 2023
Cited alongside, same era.
Llama-adapter v2: Parameter-efficient visual instruction model
Gao, P., Han, J., Zhang, R., Lin, Z., Geng, S., Zhou, A., Zhang, W., Lu, P., He, C., Yue, X., Li, H., and Qiao, Y · 2023
Cited alongside, same era.
Adapt: Action-aware driving caption transformer, 2023
Jin, B., Liu, X., Zheng, Y., Li, P., Zhao, H., Zhang, T., Zheng, Y., Zhou, G., and Liu, J · 2023
Cited alongside, same era.
Scenarionet: Open-source platform for large-scale traffic scenario simulation and modeling
Li, Q., Peng, Z., Feng, L., Liu, Z., Duan, C., Mo, W., and Zhou, B · 2023
Cited alongside, same era.
Drama: Joint risk localization and captioning in driving
Malla, S., Choi, C., Dwivedi, I., Choi, J. H., and Li, J · 2023
Cited alongside, same era.
Reason2drive: Towards interpretable and chain-based reasoning for autonomous driving
Nie, M., Peng, R., Wang, C., Cai, X., Han, J., Xu, H., and Zhang, L · 2023
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer, 2023
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2023
Cited alongside, same era.
Inoue, Y., Yada, Y., Tanahashi, K., and Yamaguchi, Y · 2024
Closest in time.
Tod3cap: Towards 3d dense captioning in outdoor scenes
Jin, B., Zheng, Y., Li, P., Li, W., Zheng, Y., Hu, S., Liu, X., Zhu, J., Yan, Z., Sun, H., et al · 2024
Closest in time.
Visual instruction tuning
Liu, H., Li, C., Wu, Q., and Lee, Y. J · 2024
Closest in time.
Lampilot: An open benchmark dataset for autonomous driving with language model programs
Ma, Y., Cui, C., Cao, X., Ye, W., Liu, P., Lu, J., Abdelraouf, A., Gupta, R., Han, K., Bera, A., Rehg, J. M., and Wang, Z · 2024
Closest in time.
Gpt-4 technical report, 2024
OpenAI, Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., Avila, R., Babuschkin, I., Balaji, S., Balcom, V., Baltescu, P., Bao, H., Bavarian, M., Belgum, J., Bello, I., Berdine, J., Bernadett-Shapiro, G., Berner, C., Bogdonoff, L., Boiko, O., Boyd, M., Brakman, A.-L., Brockman, G., Brooks, T., Brundage, M., Button, K., Cai, T., Campbell, R., Cann, A., Carey, B., Carlson, C., Carmichael, R., Chan, B., Chang, C., Chantzis, F., Chen, D., Chen, S., Chen, R., Chen, J., Chen, M., Chess, B., Cho, C., Chu, C., Chung, H. W., Cummings, D., Currier, J., Dai, Y., Decareaux, C., Degry, T., Deutsch, N., Deville, D., Dhar, A., Dohan, D., Dowling, S., Dunning, S., Ecoffet, A., Eleti, A., Eloundou, T., Farhi, D., Fedus, L., Felix, N., Fishman, S. P., Forte, J., Fulford, I., Gao, L., Georges, E., Gibson, C., Goel, V., Gogineni, T., Goh, G., Gontijo-Lopes, R., Gordon, J., Grafstein, M., Gray, S., Greene, R., Gross, J., Gu, S. S., Guo, Y., Hallacy, C., Han, J., Harris, J., He, Y., Heaton, M., Heidecke, J., Hesse, C., Hickey, A., Hickey, W., Hoeschele, P., Houghton, B., Hsu, K., Hu, S., Hu, X., Huizinga, J., Jain, S., Jain, S., Jang, J., Jiang, A., Jiang, R., Jin, H., Jin, D., Jomoto, S., Jonn, B., Jun, H., Kaftan, T., Łukasz Kaiser, Kamali, A., Kanitscheider, I., Keskar, N. S., Khan, T., Kilpatrick, L., Kim, J. W., Kim, C., Kim, Y., Kirchner, J. H., Kiros, J., Knight, M., Kokotajlo, D., Łukasz Kondraciuk, Kondrich, A., Konstantinidis, A., Kosic, K., Krueger, G., Kuo, V., Lampe, M., Lan, I., Lee, T., Leike, J., Leung, J., Levy, D., Li, C. M., Lim, R., Lin, M., Lin, S., Litwin, M., Lopez, T., Lowe, R., Lue, P., Makanju, A., Malfacini, K., Manning, S., Markov, T., Markovski, Y., Martin, B., Mayer, K., Mayne, A., McGrew, B., McKinney, S. M., McLeavey, C., McMillan, P., McNeil, J., Medina, D., Mehta, A., Menick, J., Metz, L., Mishchenko, A., Mishkin, P., Monaco, V., Morikawa, E., Mossing, D., Mu, T., Murati, M., Murk, O., Mély, D., Nair, A., Nakano, R., Nayak, R., Neelakantan, A., Ngo, R., Noh, H., Ouyang, L., O’Keefe, C., Pachocki, J., Paino, A., Palermo, J., Pantuliano, A., Parascandolo, G., Parish, J., Parparita, E., Passos, A., Pavlov, M., Peng, A., Perelman, A., de Avila Belbute Peres, F., Petrov, M., de Oliveira Pinto, H. P., Michael, Pokorny, Pokrass, M., Pong, V. H., Powell, T., Power, A., Power, B., Proehl, E., Puri, R., Radford, A., Rae, J., Ramesh, A., Raymond, C., Real, F., Rimbach, K., Ross, C., Rotsted, B., Roussez, H., Ryder, N., Saltarelli, M., Sanders, T., Santurkar, S., Sastry, G., Schmidt, H., Schnurr, D., Schulman, J., Selsam, D., Sheppard, K., Sherbakov, T., Shieh, J., Shoker, S., Shyam, P., Sidor, S., Sigler, E., Simens, M., Sitkin, J., Slama, K., Sohl, I., Sokolowsky, B., Song, Y., Staudacher, N., Such, F. P., Summers, N., Sutskever, I., Tang, J., Tezak, N., Thompson, M. B., Tillet, P., Tootoonchian, A., Tseng, E., Tuggle, P., Turley, N., Tworek, J., Uribe, J. F. C., Vallone, A., Vijayvergiya, A., Voss, C., Wainwright, C., Wang, J. J., Wang, A., Wang, B., Ward, J., Wei, J., Weinmann, C., Welihinda, A., Welinder, P., Weng, J., Weng, L., Wiethoff, M., Willner, D., Winter, C., Wolrich, S., Wong, H., Workman, L., Wu, S., Wu, J., Wu, M., Xiao, K., Xu, T., Yoo, S., Yu, K., Yuan, Q., Zaremba, W., Zellers, R., Zhang, C., Zhang, M., Zhao, S., Zheng, T., Zhuang, J., Zhuk, W., and Zoph, B · 2024
Closest in time.
Vlaad: Vision and language assistant for autonomous driving
Park, S., Lee, M., Kang, J., Choi, H., Park, Y., Cho, J., Lee, A., and Kim, D · 2024
Closest in time.
Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario, 2024
Qian, T., Chen, J., Zhuo, L., Jiao, Y., and Jiang, Y.-G · 2024
Closest in time.
A comprehensive survey of hallucination mitigation techniques in large language models, 2024
Tonmoy, S. M. T. I., Zaman, S. M. M., Jain, V., Rani, A., Rawte, V., Chadha, A., and Das, A · 2024
Closest in time.
Drive anywhere: Generalizable end-to-end autonomous driving with multi-modal foundation models
Wang, T.-H., Maalouf, A., Xiao, W., Ban, Y., Amini, A., Rosman, G., Karaman, S., and Rus, D · 2024
Closest in time.
Yuan, J., Sun, S., Omeiza, D., Zhao, B., Newman, P., Kunze, L., and Gadd, M · 2024
Closest in time.
Wisead: Knowledge augmented end-to-end autonomous driving with vision-language model
Zhang, S., Huang, W., Gao, Z., Chen, H., and Lv, C · 2024
Closest in time.
Measuring sociality in driving interaction
Zhao, X., Sun, J., and Wang, M · 2024
Closest in time.
Embodied understanding of driving scenarios
Zhou, Y., Huang, L., Bu, Q., Zeng, J., Li, T., Qiu, H., Zhu, H., Guo, M., Qiao, Y., and Li, H · 2024
Closest in time.
Bai, S., Chen, K., Liu, X., Wang, J., Ge, W., Song, S., Dang, K., Wang, P., Wang, S., Tang, J., et al · 2025
Closest in time.
Vita-1.5: Towards gpt-4o level real-time vision and speech interaction
Fu, C., Lin, H., Wang, X., Zhang, Y.-F., Shen, Y., Liu, X., Cao, H., Long, Z., Gao, H., Li, K., et al · 2025
Closest in time.
Lidar-llm: Exploring the potential of large language models for 3d lidar understanding
Yang, S., Liu, J., Zhang, R., Pan, M., Guo, Z., Li, X., Chen, Z., Gao, P., Li, H., Guo, Y., et al · 2025
Closest in time.