Fetching the paper…
Reading the bibliography…
Traffic accidents present complex challenges for autonomous driving, often featuring unpredictable scenarios that hinder accurate system interpretation and responses.
Xintao Wang, Liangbin Xie, Chao Dong, and Ying Shan, ”Real-ESRGAN: Training real-world blind super-resolution with pure synthetic data,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
1914
Earlier work this paper cites.
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu, ”BLEU: A method for automatic evaluation of machine translation,” in Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics
2002
Earlier work this paper cites.
Chin-Yew Lin and Franz Josef Och, ”Automatic evaluation of machine translation quality using longest common subsequence and skip-bigram statistics,” in Proceedings of the 42nd Annual Meeting of the Association for Computational Linguistics (ACL-04)
2004
Earlier work this paper cites.
Satanjeev Banerjee and Alon Lavie, ”METEOR: An automatic metric for MT evaluation with improved correlation with human judgments,” in Proceedings of the ACL Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and/or Summarization
2005
Earlier work this paper cites.
Dominique Brunet, Edward R Vrscay, and Zhou Wang, ”On the mathematical properties of the structural similarity index,” IEEE Transactions on Image Processing
2011
Earlier work this paper cites.
Jari Korhonen and Junyong You, ”Peak signal-to-noise ratio revisited: Is simple beautiful?” in 2012 Fourth International Workshop on Quality of Multimedia Experience
2012
Earlier work this paper cites.
Subhashini Venugopalan, Marcus Rohrbach, Jeffrey Donahue, Raymond Mooney, Trevor Darrell, and Kate Saenko, ”Sequence to sequence-video to text,” in Proceedings of the IEEE International Conference on Computer Vision
2015
Earlier work this paper cites.
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh, ”CIDER: Consensus-based image description evaluation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2015
Earlier work this paper cites.
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba, ”Generating videos with scene dynamics,” Advances in Neural Information Processing Systems
2016
Earlier work this paper cites.
Alexey Dosovitskiy and Thomas Brox, ”Generating images with perceptual similarity metrics based on deep networks,” Advances in Neural Information Processing Systems
2016
Earlier work this paper cites.
Jinkyu Kim and John Canny, ”Interpretable learning for self-driving cars by visualizing causal attention,” in Proceedings of the IEEE International Conference on Computer Vision
2017
Earlier work this paper cites.
Zaeem Hussain, Mingda Zhang, Xiaozhong Zhang, Keren Ye, Christopher Thomas, Zuha Agha, Nathan Ong, and Adriana Kovashka, ”Automatic understanding of image and video advertisements,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2017
Earlier work this paper cites.
Steven J Rennie, Etienne Marcheret, Youssef Mroueh, Jerret Ross, and Vaibhava Goel, ”Self-critical sequence training for image captioning,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2017
Earlier work this paper cites.
A Vaswani, ”Attention is all you need,” Advances in Neural Information Processing Systems
2017
Earlier work this paper cites.
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter, ”GANs trained by a two time-scale update rule converge to a local Nash equilibrium,” Advances in Neural Information Processing Systems
2017
Earlier work this paper cites.
Jinkyu Kim, Anna Rohrbach, Trevor Darrell, John Canny, and Zeynep Akata, ”Textual explanations for self-driving vehicles,” in Proceedings of the European Conference on Computer Vision (ECCV)
2018
Earlier work this paper cites.
Haoye Cai, Chunyan Bai, Yu-Wing Tai, and Chi-Keung Tang, ”Deep video generation, prediction and completion of human action sequences,” in Proceedings of the European Conference on Computer Vision (ECCV)
2018
Earlier work this paper cites.
Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang, and Jan Kautz, ”Mocogan: Decomposing motion and content for video generation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2018
Earlier work this paper cites.
Emily Denton and Rob Fergus, ”Stochastic video generation with a learned prior,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
Agrim Gupta, Justin Johnson, Li Fei-Fei, Silvio Savarese, and Alexandre Alahi, ”Social GAN: Socially acceptable trajectories with generative adversarial networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2018
Earlier work this paper cites.
Xintao Wang, Ke Yu, Shixiang Wu, Jinjin Gu, Yihao Liu, Chao Dong, Yu Qiao, and Chen Change Loy, ”ESRGAN: Enhanced super-resolution generative adversarial networks,” in Proceedings of the European Conference on Computer Vision (ECCV) Workshops
2018
Earlier work this paper cites.
Ruotian Luo, Brian Price, Scott Cohen, and Gregory Shakhnarovich, ”Discriminability objective for training descriptive captions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Wenyuan Zeng, Wenjie Luo, Simon Suo, Abbas Sadat, Bin Yang, Sergio Casas, and Raquel Urtasun, ”End-to-end interpretable neural motion planner,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2019
Cited alongside, same era.
Ning Xu, Hanwang Zhang, An-An Liu, Weizhi Nie, Yuting Su, Jie Nie, and Yongdong Zhang, ”Multi-level policy and reward-based deep reinforcement learning framework for image captioning,” IEEE Transactions on Multimedia
2019
Cited alongside, same era.
Thao Minh Le, Vuong Le, Svetha Venkatesh, and Truyen Tran, ”Hierarchical conditional relation networks for video question answering,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Chongqi Zhang, Ziwen Zhang, Yao Deng, Yueyi Zhang, Mingzhe Chong, Yunhua Tan, and Pukun Liu, ”Blind super-resolution for SAR images with speckle noise based on deep learning probabilistic degradation model and SAR priors,” Remote Sensing
2023
Later among the works it cites.
2023
Later among the works it cites.
Li Chen, Penghao Wu, Kashyap Chitta, Bernhard Jaeger, Andreas Geiger, and Hongyang Li, ”End-to-end autonomous driving: Challenges and frontiers,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2024
Later among the works it cites.
Ana-Maria Marcu, Long Chen, Jan Hünermann, Alice Karnsund, Benoit Hanotte, Prajwal Chidananda, Saurabh Nair, Vijay Badrinarayanan, Alex Kendall, Jamie Shotton, and others, ”Lingoqa: Visual question answering for autonomous driving,” in European Conference on Computer Vision
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruotian Luo, ”A better variant of self-critical sequence training,” arXiv preprint arXiv:2003.09971
2020
Cited alongside, same era.
Artem Obukhov and Mikhail Krasnyanskiy, ”Quality assessment method for GAN based on modified metrics inception score and Fréchet inception distance,” in Software Engineering Perspectives in Intelligent Systems: Proceedings of 4th Computational Methods in Systems and Software 2020, Vol. 1
2020
Cited alongside, same era.
Li Xu, He Huang, and Jun Liu, ”Sutd-trafficqa: A question answering benchmark and an efficient network for video reasoning over traffic events,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2021
Cited alongside, same era.
Richa Nahata, Daniel Omeiza, Rhys Howard, and Lars Kunze, ”Assessing and explaining collision risk in dynamic environments for autonomous driving safety,” in 2021 IEEE International Intelligent Transportation Systems Conference (ITSC)
2021
Cited alongside, same era.
Teng Wang, Ruimao Zhang, Zhichao Lu, Feng Zheng, Ran Cheng, and Ping Luo, ”End-to-end dense video captioning with parallel decoding,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2021
Cited alongside, same era.
Yingxue Pang, Jianxin Lin, Tao Qin, and Zhibo Chen, ”Image-to-image translation: Methods and applications,” IEEE Transactions on Multimedia
2021
Cited alongside, same era.
Kevin Lin, Linjie Li, Chung-Ching Lin, Faisal Ahmed, Zhe Gan, Zicheng Liu, Yumao Lu, and Lijuan Wang, ”Swinbert: End-to-end transformers with sparse attention for video captioning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
Bu Jin, Xinyu Liu, Yupeng Zheng, Pengfei Li, Hao Zhao, Tong Zhang, Yuhang Zheng, Guyue Zhou, and Jingjing Liu, ”Adapt: Action-aware driving caption transformer,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Cited alongside, same era.
2024
Later among the works it cites.
Tianwen Qian, Jingjing Chen, Linhai Zhuo, Yang Jiao, and Yu-Gang Jiang, ”Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario,” in Proceedings of the AAAI Conference on Artificial Intelligence
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Ziqi Huang, Yinan He, Jiashuo Yu, Fan Zhang, Chenyang Si, Yuming Jiang, Yuanhan Zhang, Tianxing Wu, Qingyang Jin, Nattapol Chanpaisit, and others, ”Vbench: Comprehensive benchmark suite for video generative models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2024
Later among the works it cites.
Yuqing Wen, Yucheng Zhao, Yingfei Liu, Fan Jia, Yanhui Wang, Chong Luo, Chi Zhang, Tiancai Wang, Xiaoyan Sun, and Xiangyu Zhang, ”Panacea: Panoramic and controllable video generation for autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Huan-ang Gao, Mingju Gao, Jiaju Li, Wenyi Li, Rong Zhi, Hao Tang, and Hao Zhao, ”SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis,” in European Conference on Computer Vision
2024
Later among the works it cites.
Wenyi Li, Haoran Xu, Guiyu Zhang, Huan-ang Gao, Mingju Gao, Mengyu Wang, and Hao Zhao, ”Fairdiff: Fair segmentation with point-image diffusion,” in International Conference on Medical Image Computing and Computer-Assisted Intervention
2024
Later among the works it cites.
Cheng Chang, Siqi Wang, Jiawei Zhang, Jingwei Ge, and Li Li, ”LLMScenario: Large Language Model Driven Scenario Generation,” IEEE Transactions on Systems, Man, and Cybernetics: Systems
2024
Later among the works it cites.
Jianwu Fang, Lei-lei Li, Junfei Zhou, Junbin Xiao, Hongkai Yu, Chen Lv, Jianru Xue, and Tat-Seng Chua, ”Abductive Ego-View Accident Video Understanding for Safe Driving Perception,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2024
Later among the works it cites.
2024
Later among the works it cites.
Ads of the World
2024
Later among the works it cites.
Cannes, Cannes Lions
2024
Later among the works it cites.
2024
Later among the works it cites.
Baolin Li, Yankai Jiang, and Devesh Tiwari, ”Carbon in Motion: Characterizing Open-Sora on the Sustainability of Generative AI for Video Generation,” in 3rd Workshop on Sustainable Computer Systems (HotCarbon)
2024
Later among the works it cites.
Songen Gu, Jinming Su, Yiting Duan, Xingyue Chen, Junfeng Luo, and Hao Zhao, ”Text2street: Controllable text-to-image generation for street views,” in International Conference on Pattern Recognition
2025
Closest in time.