Fetching the paper…
Reading the bibliography…
The applications of Vision-Language Models (VLMs) in the field of Autonomous Driving (AD) have attracted widespread attention due to their outstanding performance and the ability to leverage Large Language Models (LLMs).
P. E. Hart, N. J. Nilsson, and B. Raphael, “A formal basis for the heuristic determination of minimum cost paths,” IEEE Trans. on Systems Science and Cybernetics
1968
Earlier work this paper cites.
J. Barraquand and J.-C. Latombe, “Robot motion planning: A distributed representation approach,” The Int. Journal of Robotics Research
1991
Earlier work this paper cites.
S. LaValle and J. Kuffner, “Randomized kinodynamic planning,” in Proc. IEEE Int. Conf. on Robotics and Aut
1999
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proc. of Association for Computational Linguistics
2002
Earlier work this paper cites.
S. Banerjee and A. Lavie, “METEOR: An automatic metric for MT evaluation with improved correlation with human judgments,” in Proc. of the ACL Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and/or Summarization
2005
Earlier work this paper cites.
P. Dollár, C. Wojek, B. Schiele, and P. Perona, “Pedestrian detection: A benchmark,” in 2009 IEEE Conf. on computer vision and pattern recognition
2009
Earlier work this paper cites.
S. Karaman and E. Frazzoli, “Sampling-based algorithms for optimal motion planning,” 2011
2011
Earlier work this paper cites.
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun, “Vision meets robotics: The kitti dataset,” The Int. Journal of Robotics Research
2013
Earlier work this paper cites.
R. Vedantam, C. L. Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in Proc. of the IEEE Conf. on computer vision and pattern recognition
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” vol. 30, 2017
2017
Earlier work this paper cites.
M. L. Fung, M. Z. Q. Chen, and Y. H. Chen, “Sensor fusion: A review of methods and applications,” in Chinese Control And Decision Conf. (CCDC)
2017
Earlier work this paper cites.
S. Zhang, R. Benenson, and B. Schiele, “Citypersons: A diverse dataset for pedestrian detection,” in Proc. of the IEEE Conf. on computer vision and pattern recognition
2017
Earlier work this paper cites.
K. Ganesan, “Rouge 2.0: Updated and improved measures for evaluation of summarization tasks,” 2018
2018
Earlier work this paper cites.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” 2018
2018
Earlier work this paper cites.
J. Kim, A. Rohrbach, T. Darrell, J. Canny, and Z. Akata, “Textual explanations for self-driving vehicles,” in Proc. of the European Conf. on computer vision (ECCV)
2018
Earlier work this paper cites.
A. B. Vasudevan, D. Dai, and L. Van Gool, “Object referring in videos with language and human gaze,” in Proc. of the IEEE Conf. on Computer Vision and Pattern Recognition
2018
Earlier work this paper cites.
A. H. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang, and O. Beijbom, “Pointpillars: Fast encoders for object detection from point clouds,” in Proc. of the IEEE/CVF Conf. on computer vision and pattern recognition
2019
Earlier work this paper cites.
L. Chen, X. Hu, B. Tang, and D. Cao, “Parallel motion planning: Learning a deep planning model against emergencies,” IEEE Intelligent Transportation Systems Magazine
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” 2019
2019
Earlier work this paper cites.
N. Sriram, T. Maniar, J. Kalyanasundaram, V. Gandhi, B. Bhowmick, and K. M. Krishna, “Talk to the vehicle: Language conditioned autonomous navigation of self driving cars,” in Int. Conf. on Intelligent Robots and Systems (IROS)
2019
Earlier work this paper cites.
T. Unterthiner, S. van Steenkiste, K. Kurach, R. Marinier, M. Michalski, and S. Gelly, “Towards accurate generative models of video: A new metric & challenges,” 2019
2019
Earlier work this paper cites.
N. Sriram, T. Maniar, J. Kalyanasundaram, V. Gandhi, B. Bhowmick, and K. M. Krishna, “Talk to the vehicle: Language conditioned autonomous navigation of self driving cars,” in Int. Conf. on Intelligent Robots and Systems (IROS)
2019
Earlier work this paper cites.
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall, “Semantickitti: A dataset for semantic scene understanding of lidar sequences,” in Proc. of IEEE/CVF int. conf. on computer vision
2019
Earlier work this paper cites.
Z. Tang, M. Naphade, M.-Y. Liu, X. Yang, S. Birchfield, S. Wang, R. Kumar, D. Anastasiu, and J.-N. Hwang, “Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification,” in Proc. of IEEE/CVF Conf. on Computer Vision and Pattern Recognition
2019
Earlier work this paper cites.
H. Chen, A. Suhr, D. Misra, N. Snavely, and Y. Artzi, “Touchdown: Natural language navigation and spatial reasoning in visual street environments,” in Proc. of the IEEE/CVF Conf. on Computer Vision and Pattern Recognition
2019
Earlier work this paper cites.
J. Kim, T. Misu, Y. T. Chen, A. Tawari, and J. Canny, “Grounding human-to-vehicle advice for self-driving vehicles,” in Proc. of the IEEE Computer Society Conf. on Computer Vision and Pattern Recognition
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al
2020
Earlier work this paper cites.
E. Yurtsever, J. Lambert, A. Carballo, and K. Takeda, “A survey of autonomous driving: Common practices and emerging technologies,” IEEE access
2020
Earlier work this paper cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” 2020
2020
Earlier work this paper cites.
A. Sadat, S. Casas, M. Ren, X. Wu, P. Dhawan, and R. Urtasun, “Perceive, predict, and plan: Safe motion planning through interpretable semantic representations,” in Proc. of European Conf. on Computer Vision
2020
Earlier work this paper cites.
J. Kaplan, S. McCandlish, T. Henighan, T. B. Brown, B. Chess, R. Child, S. Gray, A. Radford, J. Wu, and D. Amodei, “Scaling laws for neural language models,” 2020
2020
Earlier work this paper cites.
W. Li, Z. Qu, H. Song, P. Wang, and B. Xue, “The traffic scene understanding and prediction based on image captioning,” IEEE Access
2020
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in Proc. of IEEE/CVF conf. on comp. vision and pattern recognition
2020
Earlier work this paper cites.
T. Zhang*, V. Kishore*, F. Wu*, K. Q. Weinberger, and Y. Artzi, “Bertscore: Evaluating text generation with bert,” in Int. Conf. on Learning Representations
2020
Earlier work this paper cites.
J. Kim, S. Moon, A. Rohrbach, T. Darrell, and J. Canny, “Advisable learning for self-driving vehicles by internalizing observation-to-action rules,” in Proc. of the IEEE/CVF Conf. on Computer Vision and Pattern Recognition
2020
Earlier work this paper cites.
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, et al
2020
Earlier work this paper cites.
F. Yu, H. Chen, X. Wang, W. Xian, Y. Chen, F. Liu, V. Madhavan, and T. Darrell, “Bdd100k: A diverse driving dataset for heterogeneous multitask learning,” in Proc. of the IEEE/CVF Conf. on computer vision and pattern recognition
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” 2020
2020
Earlier work this paper cites.
J. Houston, G. Zuidhof, L. Bergamini, Y. Ye, L. Chen, A. Jain, S. Omari, V. Iglovikov, and P. Ondruska, “One thousand and one hours: Self-driving motion prediction dataset,” 2020
2020
Earlier work this paper cites.
Y. Xu, X. Yang, L. Gong, H.-C. Lin, T.-Y. Wu, Y. Li, and N. Vasconcelos, “Explainable object-induced action decision for autonomous vehicles,” 2020
2020
Earlier work this paper cites.
S. Int., “Taxonomy and definitions for terms related to driving automation systems for on-road motor vehicles,” 2021
2021
Earlier work this paper cites.
W. Luo, J. Xing, A. Milan, X. Zhang, W. Liu, and T.-K. Kim, “Multiple object tracking: A literature review,” Artificial Intelligence
2021
Earlier work this paper cites.
K. Mangalam, Y. An, H. Girase, and J. Malik, “From goals, waypoints & paths to long term human trajectory forecasting,” in Proc. of IEEE/CVF Int. Conf. on Computer Vision
2021
Earlier work this paper cites.
T. Gilles, S. Sabatini, D. Tsishkou, B. Stanciulescu, and F. Moutarde, “Home: Heatmap output for future motion estimation,” in IEEE Int. Intelligent Transportation Systems Conf
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” 2 2021
2021
Earlier work this paper cites.
H. Xu, G. Ghosh, P.-Y. Huang, D. Okhonko, A. Aghajanyan, F. Metze, L. Zettlemoyer, and C. Feichtenhofer, “Videoclip: Contrastive pre-training for zero-shot video-text understanding,” 2021
2021
Earlier work this paper cites.
S. W. Kim, J. Philion, A. Torralba, and S. Fidler, “Drivegan: Towards a controllable high-quality neural simulation,” pp. 5820–5829, 2021
2021
Earlier work this paper cites.
J. Mao, M. Niu, C. Jiang, H. Liang, J. Chen, X. Liang, Y. Li, C. Ye, W. Zhang, Z. Li, J. Yu, H. Xu, and C. Xu, “One million scenes for autonomous driving: Once dataset,” 2021
2021
Earlier work this paper cites.
A. Prakash, K. Chitta, and A. Geiger, “Multi-modal fusion transformer for end-to-end autonomous driving,” 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “Lora: Low-rank adaptation of large language models,” 2021
2021
Earlier work this paper cites.
B. Lester, R. Al-Rfou, and N. Constant, “The power of scale for parameter-efficient prompt tuning,” 2021
2021
Earlier work this paper cites.
C. Zhang, Y. Xie, H. Bai, B. Yu, W. Li, and Y. Gao, “A survey on federated learning,” Know.-Based Systems
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
T. Yuan, W. D. R. Neto, C. E. Rothenberg, K. Obraczka, C. Barakat, and T. Turletti, “Machine learning for next-generation intelligent transportation systems: A survey,” Trans. on Emerging Telecommunications Technologies
2022
Earlier work this paper cites.
W. Jiang and J. Luo, “Graph neural network for traffic forecasting: A survey,” Expert Systems with Applications
2022
Earlier work this paper cites.
E. Erçelik, E. Yurtsever, M. Liu, Z. Yang, H. Zhang, P. Topçam, M. Listl, Y. K. Caylı, and A. Knoll, “3d object detection with a self-supervised lidar scene flow backbone,” in European Conf. on Computer Vision
2022
Cited alongside, same era.
Z. Li, W. Wang, H. Li, E. Xie, C. Sima, T. Lu, Q. Yu, and J. Dai, “Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers,” 2022
2022
Cited alongside, same era.
T. Gilles, S. Sabatini, D. Tsishkou, B. Stanciulescu, and F. Moutarde, “Gohome: Graph-oriented heatmap output for future motion estimation,” in Int. conf. on robotics and autom
2022
Cited alongside, same era.
Z. Zhu and H. Zhao, “Multi-task conditional imitation learning for autonomous navigation at crowded intersections,” 2022
2022
Cited alongside, same era.
2023
Closest in time.
K. Jain, V. Chhangani, A. Tiwari, K. M. Krishna, and V. Gandhi, “Ground then navigate: Language-guided navigation in dynamic scenes,” in 2023 IEEE Int. Conf. on Robotics and Automation (ICRA)
2023
Closest in time.
2023
Closest in time.
S. Yang, J. Liu, R. Zhang, M. Pan, Z. Guo, X. Li, Z. Chen, P. Gao, Y. Guo, and S. Zhang, “Lidar-llm: Exploring the potential of large language models for 3d lidar understanding,” 12 2023
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Gao, Z. Gu, C. Qiu, L. Lei, S. E. Li, S. Zheng, W. Jing, and J. Chen, “Cola-hrl: Continuous-lattice hierarchical reinforcement learning for autonomous driving,” in 2022 IEEE/RSJ Int. Conf. on Intelligent Robots and Systems (IROS)
2022
Cited alongside, same era.
A. Tampuu, T. Matiisen, M. Semikin, D. Fishman, and N. Muhammad, “A survey of end-to-end driving: Architectures and training methods,” IEEE Trans. on Neural Networks and Learning Systems
2022
Cited alongside, same era.
M. Cherti, R. Beaumont, R. Wightman, M. Wortsman, G. Ilharco, C. Gordon, C. Schuhmann, L. Schmidt, and J. Jitsev, “Reproducible scaling laws for contrastive language-image learning,” 2022
2022
Cited alongside, same era.
C. Creß, W. Zimmer, L. Strand, M. Fortkord, S. Dai, V. Lakshminarasimhan, and A. Knoll, “A9-dataset: Multi-sensor infrastructure-based dataset for mobility research,” in 2022 IEEE Intelligent Vehicles Symposium (IV)
2022
Cited alongside, same era.
J. Zhang, X. Lin, M. Jiang, Y. Yu, C. Gong, W. Zhang, X. Tan, Y. Li, E. Ding, and G. Li, “A multi-granularity retrieval system for natural language-based vehicle retrieval,” in Proc. of the IEEE/CVF Conf. on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
S. Wang, D. C. Anastasiu, S. Sclaroff, and A. Li, “The 6th AI City Challenge,” pp. 3347–3356, 2022
2022
Cited alongside, same era.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” Int. Journal of Computer Vision
2022
Cited alongside, same era.
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, et al
2022
Cited alongside, same era.
M. Nie, R. Peng, C. Wang, X. Cai, J. Han, H. Xu, and L. Zhang, “Reason2drive: Towards interpretable and chain-based reasoning for autonomous driving,” 12 2023
2023
Closest in time.
2023
Closest in time.
OpenAI, “Gpt-4v(ision) system card,” 2023
2023
Closest in time.
H. D.-A. Le, Q. Q.-V. Nguyen, D. T. Luu, T. T.-T. Chau, N. M. Chung, and S. V.-U. Ha, “Tracked-vehicle retrieval by natural language descriptions with multi-contextual adaptive knowledge,” in Proc. IEEE/CVF Conf. on Comp. Vision and Pattern Recog
2023
Closest in time.
D. Xie, L. Liu, S. Zhang, and J. Tian, “A unified multi-modal structure for retrieving tracked vehicles through natural language descriptions,” in 2023 IEEE/CVF Conf. on Computer Vision and Pattern Recognition Workshops (CVPRW)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Y. Inoue, Y. Yada, K. Tanahashi, and Y. Yamaguchi, “Nuscenes-mqa: Integrated evaluation of captions and qa for autonomous driving datasets using markup annotations,” 12 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
G. Chen, X. Liu, G. Wang, K. Zhang, P. H. Torr, X.-P. Zhang, and Y. Tang, “Tem-adapter: Adapting image-text pretraining for video question answer,” in IEEE/CVF Int. Conf. on Computer Vision
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
P. Wang, M. Zhu, H. Lu, H. Zhong, X. Chen, S. Shen, X. Wang, and Y. Wang, “Bevgpt: Generative pre-trained large model for autonomous driving prediction, decision-making, and planning,” 10 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
W. Wang, J. Xie, C. Hu, H. Zou, J. Fan, W. Tong, Y. Wen, S. Wu, H. Deng, Z. Li, H. Tian, L. Lu, X. Zhu, X. Wang, Y. Qiao, and J. Dai, “Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,” 12 2023
2023
Closest in time.
2023
Closest in time.
F. Jia, W. Mao, Y. Liu, Y. Zhao, Y. Wen, C. Zhang, X. Zhang, and T. Wang, “Adriver-i: A general world model for autonomous driving,” 11 2023
2023
Closest in time.
S. Wang, D. C. Anastasiu, M. S. Arya, and S. Sclaroff, “The 7th AI City Challenge,” 2023
2023
Closest in time.
A. Wantiez, T. Qiu, S. Matthes, and H. Shen, “Scene understanding for autonomous driving using visual question answering,” in Int. Joint Conf. on Neural Networks (IJCNN)
2023
Closest in time.
P. Gao, J. Han, R. Zhang, Z. Lin, S. Geng, A. Zhou, W. Zhang, P. Lu, C. He, X. Yue, H. Li, and Y. Qiao, “Llama-adapter v2: Parameter-efficient visual instruction model,” 2023
2023
Closest in time.
P. Gao, J. Han, R. Zhang, Z. Lin, S. Geng, A. Zhou, W. Zhang, P. Lu, C. He, X. Yue, H. Li, and Y. Qiao, “Llama-adapter v2: Parameter-efficient visual instruction model,” 2023
2023
Closest in time.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” 2023
2023
Closest in time.
H. Liu, C. Li, Y. Li, and Y. J. Lee, “Improved baselines with visual instruction tuning,” 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
S. Malla, C. Choi, I. Dwivedi, J. H. Choi, and J. Li, “Drama: Joint risk localization and captioning in driving,” in Proc. of IEEE/CVF Winter Conf. on Applications of Comp. Vision
2023
Closest in time.
2023
Closest in time.
X. Wang, X. Zhang, Y. Cao, W. Wang, C. Shen, and T. Huang, “Seggpt: Segmenting everything in context,” 2023
2023
Closest in time.
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, P. Dollár, and R. Girshick, “Segment anything,” 2023
2023
Closest in time.
X. Wang, W. Wang, Y. Cao, C. Shen, and T. Huang, “Images speak in images: A generalist painter for in-context visual learning,” 2023
2023
Closest in time.
J. Li, D. Li, S. Savarese, and S. Hoi, “Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,” 2023
2023
Closest in time.
R. Xu, H. Xiang, X. Han, X. Xia, Z. Meng, C.-J. Chen, C. Correa-Jullian, and J. Ma, “The opencda open-source ecosystem for cooperative driving automation research,” IEEE Trans. on Intelligent Vehicles
2023
Closest in time.
R. Xu, X. Xia, J. Li, H. Li, S. Zhang, Z. Tu, Z. Meng, H. Xiang, X. Dong, R. Song, H. Yu, B. Zhou, and J. Ma, “V2v4real: A real-world large-scale dataset for vehicle-to-vehicle cooperative perception,” in Proc. of IEEE/CVF Conf. on Computer Vision and Pattern Recognition (CVPR)
2023
Closest in time.
J. Li, R. Xu, X. Liu, J. Ma, Z. Chi, J. Ma, and H. Yu, “Learning for vehicle-to-vehicle cooperative perception under lossy communication,” IEEE Trans. on Intelligent Vehicles
2023
Closest in time.
Q. Zhang, M. Chen, A. Bukharin, P. He, Y. Cheng, W. Chen, and T. Zhao, “Adaptive budget allocation for parameter-efficient fine-tuning,” 2023
2023
Closest in time.
Y. Gu, L. Dong, F. Wei, and M. Huang, “Knowledge distillation of large language models,” 2023
2023
Closest in time.
G. Xiao, J. Lin, M. Seznec, H. Wu, J. Demouth, and S. Han, “Smoothquant: Accurate and efficient post-training quantization for large language models,” in Int. Conf. on Machine Learning
2023
Closest in time.
H. Zhang, X. Li, and L. Bing, “Video-llama: An instruction-tuned audio-visual language model for video understanding,” 2023
2023
Closest in time.
M. Maaz, H. Rasheed, S. Khan, and F. S. Khan, “Video-chatgpt: Towards detailed video understanding via large vision and language models,” 2023
2023
Closest in time.
L. Huang, W. Yu, W. Ma, W. Zhong, Z. Feng, H. Wang, Q. Chen, W. Peng, X. Feng, B. Qin, and T. Liu, “A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions,” 2023
2023
Closest in time.
P. Christiano, J. Leike, T. Brown, M. Martic, S. Legg, and D. Amodei, “Deep reinforcement learning from human preferences,” 2023
2023
Closest in time.
R. Rafailov, A. Sharma, E. Mitchell, S. Ermon, C. D. Manning, and C. Finn, “Direct preference optimization: Your language model is secretly a reward model,” 2023
2023
Closest in time.
Q. Li, Z. Wen, Z. Wu, S. Hu, N. Wang, Y. Li, X. Liu, and B. He, “A survey on federated learning systems: Vision, hype and reality for data privacy and protection,” IEEE Trans. on Knowledge and Data Engineering
2023
Closest in time.
H. Wu, S. Yunas, S. Rowlands, W. Ruan, and J. Wahlström, “Adversarial driving: Attacking end-to-end autonomous driving,” in 2023 IEEE Intelligent Vehicles Symposium (IV)
2023
Closest in time.
H. Li, Y. Zhao, Z. Mao, Y. Qin, Z. Xiao, J. Feng, Y. Gu, W. Ju, X. Luo, and M. Zhang, “A survey on graph neural networks in intelligent transportation systems,” 2024
2024
Closest in time.
W. Zimmer, G. A. Wardana, S. Sritharan, X. Zhou, R. Song, and A. C. Knoll, “Tumtraf v2x cooperative perception dataset,” in Proc. of IEEE/CVF Conf. on Comp. Vision and Pattern Recog
2024
Closest in time.
W. Ju, Y. Zhao, Y. Qin, S. Yi, J. Yuan, Z. Xiao, X. Luo, X. Yan, and M. Zhang, “Cool: A conjoint perspective on spatio-temporal graph neural network for traffic forecasting,” 2024
2024
Closest in time.
B. Zhang, P. Zhang, X. Dong, Y. Zang, and J. Wang, “Long-clip: Unlocking the long-text capability of clip,” 2024
2024
Closest in time.
D. Wei, T. Gao, Z. Jia, C. Cai, C. Hou, P. Jia, F. Liu, K. Zhan, J. Fan, Y. Zhao, and Y. Wang, “Bev-clip: Multi-modal bev retrieval methodology for complex scene in autonomous driving,” 1 2024
2024
Closest in time.
X. Tian, J. Gu, B. Li, Y. Liu, C. Hu, Y. Wang, K. Zhan, P. Jia, X. Lang, and H. Zhao, “Drivevlm: The convergence of autonomous driving and large vision-language models,” 2 2024
2024
Closest in time.
C. Pan, B. Yaman, T. Nesti, A. Mallik, A. G. Allievi, S. Velipasalar, and L. Ren, “Vlp: Vision language planning for autonomous driving,” 1 2024
2024
Closest in time.
A. Swerdlow, R. Xu, and B. Zhou, “Street-view image generation from a bird’s-eye view layout,” 2024
2024
Closest in time.