Fetching the paper…
Reading the bibliography…
The advent of foundation models has revolutionized the fields of natural language processing and computer vision, paving the way for their application in autonomous driving (AD).
Kingma, D.P., Welling, M.: Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)
2013
Earlier work this paper cites.
Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., Ganguli, S.: Deep unsupervised learning using nonequilibrium thermodynamics. In: International Conference on Machine Learning, pp. 2256–2265 (2015). PMLR
2015
Earlier work this paper cites.
Ronneberger, O., Fischer, P., Brox, T.: U-net: Convolutional networks for biomedical image segmentation. In: Medical Image Computing and Computer-Assisted Intervention–MICCAI 2015: 18th International Conference, Munich, Germany, October 5-9, 2015, Proceedings, Part III 18, pp. 234–241 (2015). Springer
2015
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I.: Improving language understanding with unsupervised learning (2018)
2018
Earlier work this paper cites.
Kim, J., Rohrbach, A., Darrell, T., Canny, J., Akata, Z.: Textual explanations for self-driving vehicles. In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 563–578 (2018)
2018
Earlier work this paper cites.
Ho, J., Jain, A., Abbeel, P.: Denoising diffusion probabilistic models. Advances in neural information processing systems 33
2020
Earlier work this paper cites.
Caron, M., Touvron, H., Misra, I., Jégou, H., Mairal, J., Bojanowski, P., Joulin, A.: Emerging properties in self-supervised vision transformers. In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 9650–9660 (2021)
2021
Earlier work this paper cites.
Ramesh, A., Pavlov, M., Goh, G., Gray, S., Voss, C., Radford, A., Chen, M., Sutskever, I.: Zero-shot text-to-image generation. In: International Conference on Machine Learning, pp. 8821–8831 (2021). PMLR
2021
Earlier work this paper cites.
Radford, A., Kim, J.W., et al
2021
Earlier work this paper cites.
Wang, S., Sheng, Z., et al
2022
Earlier work this paper cites.
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., et al
2022
Earlier work this paper cites.
Wang, M., Zhang, Z., Yang, G.H.: Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars (2022)
2022
Earlier work this paper cites.
Frantar, E., Ashkboos, S., Hoefler, T., Alistarh, D.: GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers (2022)
2022
Earlier work this paper cites.
Sautier, C., Puy, G., Gidaris, S., Boulch, A., Bursuc, A., Marlet, R.: Image-to-lidar self-supervised distillation for autonomous driving data. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 9891–9901 (2022)
2022
Earlier work this paper cites.
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., Ommer, B.: High-resolution image synthesis with latent diffusion models. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 10684–10695 (2022)
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Ho, J., Salimans, T., Gritsenko, A., Chan, W., Norouzi, M., Fleet, D.J.: Video Diffusion Models (2022)
2022
Earlier work this paper cites.
Yang, Z., Jia, X., Li, H., Yan, J.: LLM4Drive: A Survey of Large Language Models for Autonomous Driving (2023)
2023
Earlier work this paper cites.
Huang, Y., Chen, Y., et al.: Applications of Large Scale Foundation Models for Autonomous Driving (2023)
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
Mao, J., Qian, Y., Ye, J., Zhao, H., Wang, Y.: GPT-Driver: Learning to Drive with GPT (2023)
2023
Cited alongside, same era.
Chen, L., Sinavski, O., Hünermann, J., Karnsund, A., Willmott, A.J., Birch, D., Maund, D., Shotton, J.: Driving with LLMs: Fusing Object-Level Vector Modality for Explainable Autonomous Driving (2023)
2023
Cited alongside, same era.
Cui, C., Ma, Y., Cao, X., Ye, W., Wang, Z.: Receive, Reason, and React: Drive as You Say with Large Language Models in Autonomous Vehicles (2023)
2023
Cited alongside, same era.
Keysan, A., Look, A., Kosman, E., Gürsun, G., Wagner, J., Yao, Y., Rakitsch, B.: Can you text what is happening? Integrating pre-trained language encoders into trajectory prediction models for autonomous driving (2023)
Peng, X., Chen, R., et al.: Learning to Adapt SAM for Segmenting Cross-domain Point Clouds (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Cheng, Y., Li, L., Xu, Y., Li, X., Yang, Z., Wang, W., Yang, Y.: Segment and Track Anything (2023)
2023
Later among the works it cites.
Hu, A., Russell, L., et al.: GAIA-1: A Generative World Model for Autonomous Driving (2023)
2023
Later among the works it cites.
Wang, X., Zhu, Z., Huang, G., Chen, X., Zhu, J., Lu, J.: DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
Yang, Y., Zhang, Q., Li, C., Marta, D.S., Batool, N., Folkesson, J.: Human-Centric Autonomous Systems With LLMs for User Command Reasoning (2023)
2023
Cited alongside, same era.
Deng, Y., Yao, J., Tu, Z., Zheng, X., Zhang, M., Zhang, T.: TARGET: Automated Scenario Generation from Traffic Rules for Testing Autonomous Vehicles (2023)
2023
Cited alongside, same era.
Tan, S., Ivanovic, B., Weng, X., Pavone, M., Kraehenbuehl, P.: Language Conditioned Traffic Generation (2023)
2023
Cited alongside, same era.
Mao, J., Ye, J., Qian, Y., Pavone, M., Wang, Y.: A Language Agent for Autonomous Driving (2023)
2023
Cited alongside, same era.
Sha, H., Mu, Y., et al.: LanguageMPC: Large Language Models as Decision Makers for Autonomous Driving (2023)
2023
Cited alongside, same era.
Wen, L., Fu, D., et al.: DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language Models (2023)
2023
Cited alongside, same era.
Jin, Y., Shen, X., Peng, H., Liu, X., Qin, J., Li, J., Xie, J., Gao, P., Zhou, G., Gong, J.: SurrealDriver: Designing Generative Driver Agent Simulation Framework in Urban Contexts based on Large Language Model (2023)
2023
Cited alongside, same era.
Zhang, L., Xiong, Y., Yang, Z., Casas, S., Hu, R., Urtasun, R.: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion (2023)
2023
Later among the works it cites.
Liu, H., Li, C., Wu, Q., Lee, Y.J.: Visual instruction tuning. In: NeurIPS (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Ding, X., Han, J., Xu, H., Zhang, W., Li, X.: HiLM-D: Towards High-Resolution Understanding in Multimodal Large Language Models for Autonomous Driving (2023)
2023
Later among the works it cites.
Li, J., Li, D., Savarese, S., Hoi, S.: BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Wen, L., Yang, X., Fu, D., Wang, X., Cai, P., Li, X., Ma, T., Li, Y., Xu, L., Shang, D., Zhu, Z., Sun, S., Bai, Y., Cai, X., Dou, M., Hu, S., Shi, B., Qiao, Y.: On the Road with GPT-4V(ision): Early Explorations of Visual-Language Model on Autonomous Driving (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Reis, D., Kupec, J., Hong, J., Daoudi, A.: Real-Time Flying Object Detection with YOLOv8 (2023)
2023
Later among the works it cites.
Rafailov, R., Sharma, A., Mitchell, E., Manning, C.D., Ermon, S., Finn, C.: Direct preference optimization: Your language model is secretly a reward model. Advances in Neural Information Processing Systems 36
2024
Closest in time.