Fetching the paper…
Reading the bibliography…
We introduce Traffic-R1, a 3B-parameter foundation model with human-like reasoning for Traffic signal control (TSC), developed via self-exploration and iterative reinforcement of LLM with expert guidance in a simulated traffic environment.
Colight: Learning network-level cooperation for traffic signal control
Hua Wei, Nan Xu, Huichu Zhang, Guanjie Zheng, Xinshi Zang, Chacha Chen, Weinan Zhang, Yanmin Zhu, Kai Xu, and Zhenhui Li. 2019 · 1922
Earlier work this paper cites.
A mathematical model for the fixed-time traffic control problem
Paolo Serafini and Walter Ukovich. 1989 · 1989
Earlier work this paper cites.
Reinforcement learning for true adaptive traffic signal control
Baher Abdulhai, Rob Pringle, and Grigoris J Karakoulas. 2003 · 2003
Earlier work this paper cites.
Neural networks for real-time traffic signal control
Dipti Srinivasan, Min Chee Choy, and Ruey Long Cheu. 2006 · 2006
Earlier work this paper cites.
Traffic signal timing manual
Peter Koonce and 1 others. 2008 · 2008
Earlier work this paper cites.
Computational intelligence in urban traffic signal control: A survey
Dongbin Zhao, Yujie Dai, and Zhen Zhang. 2011 · 2011
Earlier work this paper cites.
Max pressure control of a network of signalized intersections
Pravin Varaiya. 2013 · 2013
Earlier work this paper cites.
A survey on reinforcement learning models and algorithms for traffic signal control
Kok-Lim Alvin Yau, Junaid Qadir, Hooi Ling Khoo, Mee Hong Ling, and Peter Komisarczuk. 2017 · 2017
Earlier work this paper cites.
Literature review on traffic control systems used worldwide
Vaishali Mahavar and Jayesh Juremalani. 2018 · 2018
Earlier work this paper cites.
Intelligent road traffic control system for traffic congestion: a perspective
Pallavi A Mandhare, Vilas Kharat, and CY Patil. 2018 · 2018
Earlier work this paper cites.
Intellilight: A reinforcement learning approach for intelligent traffic light control
Hua Wei, Guanjie Zheng, Huaxiu Yao, and Zhenhui Li. 2018 · 2018
Earlier work this paper cites.
Optimization and simulation of fixed-time traffic signal control in real-world applications
Theresa Thunig, Robert Scheffler, Martin Strehler, and Kai Nagel. 2019 · 2019
Earlier work this paper cites.
A survey of model predictive control methods for traffic signal control
Bao-Lin Ye, Weimin Wu, Keyu Ruan, Lingxi Li, Tehuan Chen, Huimin Gao, and Yaobin Chen. 2019 · 2019
Earlier work this paper cites.
Cityflow: A multi-agent reinforcement learning environment for large scale city traffic scenario
Huichu Zhang, Siyuan Feng, Chang Liu, Yaoyao Ding, Yichen Zhu, Zihan Zhou, Weinan Zhang, Yong Yu, Haiming Jin, and Zhenhui Li. 2019 · 2019
Earlier work this paper cites.
Toward a thousand lights: Decentralized deep reinforcement learning for large-scale traffic signal control
Chacha Chen, Hua Wei, Nan Xu, Guanjie Zheng, Ming Yang, Yuanhao Xiong, Kai Xu, and Zhenhui Li. 2020 · 2020
Earlier work this paper cites.
Max-pressure traffic controller based on travel times: An experimental analysis
Pedro Mercader, Wasim Uwayid, and Jack Haddad. 2020 · 2020
Earlier work this paper cites.
Attendlight: Universal attention-based reinforcement learning model for traffic signal control
Afshin Oroojlooy, Mohammadreza Nazari, Davood Hajinezhad, and Jorge Silva. 2020 · 2020
Earlier work this paper cites.
State-of-art review of traffic signal control methods: challenges and opportunities
Syed Shah Sultan Mohiuddin Qadri, Mahmut Ali Gökçe, and Erdinç Öner. 2020 · 2020
Earlier work this paper cites.
Deep reinforcement learning for traffic signal control: A review
Faizan Rasheed, Kok-Lim Alvin Yau, Rafidah Md Noor, Celimuge Wu, and Yeh-Ching Low. 2020 · 2020
Earlier work this paper cites.
Towards real-world deployment of reinforcement learning for traffic signal control
Arthur Müller, Vishal Rangras, Tobias Ferfers, Florian Hufen, Lukas Schreckenberg, Jürgen Jasperneite, Georg Schnittker, Michael Waldmann, Maxim Friesen, and Marco Wiering. 2021 · 2021
Cited alongside, same era.
Recent advances in reinforcement learning for traffic signal control: A survey of models and evaluation
Hua Wei, Guanjie Zheng, Vikash Gayah, and Zhenhui Li. 2021 · 2021
Cited alongside, same era.
Efficient pressure: Improving efficiency for signalized intersections
Qiang Wu, Liang Zhang, Jun Shen, Linyuan Lü, Bo Du, and Jianqing Wu. 2021 · 2021
Cited alongside, same era.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Yuntao Bai, Andy Jones, Kamal Ndousse, Amanda Askell, Anna Chen, Nova DasSarma, Dawn Drain, Stanislav Fort, Deep Ganguli, Tom Henighan, and 1 others. 2022 · 2022
Cited alongside, same era.
Libsignal: An open library for traffic signal control
Hao Mei, Xiaoliang Lei, Longchao Da, Bin Shi, and Hua Wei. 2024 · 2024
Later among the works it cites.
Large language models: A survey
Shervin Minaee, Tomas Mikolov, Narjes Nikzad, Meysam Chenaghlu, Richard Socher, Xavier Amatriain, and Jianfeng Gao. 2024 · 2024
Later among the works it cites.
Coslight: Co-optimizing collaborator selection and decision-making to enhance traffic signal control
Jingqing Ruan, Ziyue Li, Hua Wei, Haoyuan Jiang, Jiaming Lu, Xuantang Xiong, Hangyu Mao, and Rui Zhao. 2024 · 2024
Later among the works it cites.
Hybridflow: A flexible and efficient rlhf framework
Guangming Sheng, Chi Zhang, Zilingfeng Ye, Xibin Wu, Wang Zhang, Ru Zhang, Yanghua Peng, Haibin Lin, and Chuan Wu. 2024 · 2024
Later among the works it cites.
EvalScope: Evaluation framework for large models
ModelScope Team. 2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rex Chen, Fei Fang, and Norman Sadeh. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, and 1 others. 2022 · 2022
Cited alongside, same era.
Rajkumar Ramamurthy, Prithviraj Ammanabrolu, Kianté Brantley, Jack Hessel, Rafet Sifa, Christian Bauckhage, Hannaneh Hajishirzi, and Yejin Choi. 2022 · 2022
Cited alongside, same era.
Explainable deep reinforcement learning: state of the art and challenges
George A Vouros. 2022 · 2022
Cited alongside, same era.
Expression might be enough: Representing pressure and demand for reinforcement learning based traffic signal control
Liang Zhang, Qiang Wu, Jun Shen, Linyuan Lü, Bo Du, and Jianqing Wu. 2022 · 2022
Cited alongside, same era.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer. 2023 · 2023
Cited alongside, same era.
Scaling laws for reward model overoptimization
Leo Gao, John Schulman, and Jacob Hilton. 2023 · 2023
Cited alongside, same era.
Llmlight: Large language models as traffic signal control agents
Siqi Lai, Zhao Xu, Weijia Zhang, Hao Liu, and Hui Xiong. 2023 · 2023
Cited alongside, same era.
Maonan Wang, Aoyu Pang, Yuheng Kan, Man-On Pun, Chung Shue Chen, and Bo Huang. 2024 · 2024
Later among the works it cites.
Trafficgpt: Viewing, processing and interacting with traffic foundation models
Siyao Zhang, Daocheng Fu, Wenzhe Liang, Zhao Zhang, Bin Yu, Pinlong Cai, and Baozhen Yao. 2024 · 2024
Later among the works it cites.
Tsclip: Robust clip fine-tuning for worldwide cross-regional traffic sign recognition
Guoyang Zhao, Fulong Ma, Weiqing Qi, Chenguang Zhang, Yuxuan Liu, Ming Liu, and Jun Ma. 2024 · 2024
Later among the works it cites.
Understanding world or predicting future? a comprehensive survey of world models
Jingtao Ding, Yunke Zhang, Yu Shang, Yuheng Zhang, Zefang Zong, Jie Feng, Yuan Yuan, Hongyuan Su, Nian Li, Nicholas Sukiennik, and 1 others. 2025 · 2025
Closest in time.
Citybench: Evaluating the capabilities of large language models for urban tasks
Jie Feng, Jun Zhang, Tianhui Liu, Xin Zhang, Tianjian Ouyang, Junbo Yan, Yuwei Du, Siqi Guo, and Yong Li. 2025 · 2025
Closest in time.
A survey of frontiers in llm reasoning: Inference scaling, learning to reason, and agentic systems
Zixuan Ke, Fangkai Jiao, Yifei Ming, Xuan-Phi Nguyen, Austin Xu, Do Xuan Long, Minzhi Li, Chengwei Qin, Peifeng Wang, Silvio Savarese, and 1 others. 2025 · 2025
Closest in time.
Urban computing in the era of large language models
Zhonghang Li, Lianghao Xia, Xubin Ren, Jiabin Tang, Tianyi Chen, Yong Xu, and Chao Huang. 2025 · 2025
Closest in time.
A review of faithfulness metrics for hallucination assessment in large language models
Ben Malin, Tatiana Kalganova, and Nikolaos Boulgouris. 2025 · 2025
Closest in time.
Llm fine-tuning: Concepts, opportunities, and challenges
Xiao-Kun Wu, Min Chen, Wanyi Li, Rui Wang, Limeng Lu, Jia Liu, Kai Hwang, Yixue Hao, Yanru Pan, Qingguo Meng, and 1 others. 2025 · 2025
Closest in time.
Towards large reasoning models: A survey of reinforced reasoning with large language models
Fengli Xu, Qianyue Hao, Zefang Zong, Jingwei Wang, Yunke Zhang, Jingyi Wang, Xiaochong Lan, Jiahui Gong, Tianjian Ouyang, Fanjin Meng, and 1 others. 2025 · 2025
Closest in time.
Dapo: An open-source llm reinforcement learning system at scale
Qiying Yu, Zheng Zhang, Ruofei Zhu, Yufeng Yuan, Xiaochen Zuo, Yu Yue, Weinan Dai, Tiantian Fan, Gaohong Liu, Lingjun Liu, and 1 others. 2025 · 2025
Closest in time.
Collmlight: Cooperative large language model agents for network-wide traffic signal control
Zirui Yuan, Siqi Lai, and Hao Liu. 2025 · 2025
Closest in time.
UrbanVideo-bench: Benchmarking vision-language models on embodied intelligence with video data in urban spaces
Baining Zhao, Jianjie Fang, Zichao Dai, Ziyou Wang, Jirong Zha, Weichen Zhang, Chen Gao, Yue Wang, Jinqiang Cui, Xinlei Chen, and Yong Li. 2025 · 2025
Closest in time.