Fetching the paper…
Reading the bibliography…
Reinforcement Learning (RL) is a potent tool for sequential decision-making and has achieved performance surpassing human capabilities across many challenging real-world tasks.
J. F. Nash Jr, “Equilibrium points in n-person games,” Proceedings of the national academy of sciences
1950
Earlier work this paper cites.
John Wiley, 1960
R. A. Howard, Dynamic programming and markov processes · 1960
Earlier work this paper cites.
P. E. Hart, N. J. Nilsson, and B. Raphael, “A formal basis for the heuristic determination of minimum cost paths,” IEEE Transactions on Systems Science and Cybernetics
1968
Earlier work this paper cites.
C. J. Watkins and P. Dayan, “Q-learning,” Machine learning
1992
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning
1992
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in International Conference on Machine Learning
1993
Earlier work this paper cites.
University of Cambridge, Department of Engineering Cambridge, UK, 1994
G. A. Rummery and M. Niranjan, On-line Q-learning using connectionist systems · 1994
Earlier work this paper cites.
L. Steels, “The biology and technology of intelligent autonomous agents,” Robotics and Autonomous Systems
1995
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation
1997
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” Advances in neural information processing systems
1999
Earlier work this paper cites.
B. Wymann, E. Espié, C. Guionneau, C. Dimitrakakis, R. Coulom, and A. Sumner, “Torcs, the open racing car simulator,” Software available at http://torcs. sourceforge. net
2000
Earlier work this paper cites.
D. Krajzewicz, G. Hertkorn, C. Rössel, and P. Wagner, “Sumo (simulation of urban mobility)-an open-source traffic simulation,” in Proceedings of the 4th middle East Symposium on Simulation and Modelling (MESM20002)
2002
Earlier work this paper cites.
M. Goslin and M. Mine, “The panda3d graphics engine,” Computer
2004
Earlier work this paper cites.
N. Koenig and A. Howard, “Design and use paradigms for gazebo, an open-source multi-robot simulator,” in 2004 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE Cat. No.04CH37566)
2004
Earlier work this paper cites.
V. Alexiadis, J. Colyar, J. Halkias, R. Hranac, and G. McHale, “The next generation simulation program,” Institute of Transportation Engineers. ITE Journal
2004
Earlier work this paper cites.
E. A. Hansen, D. S. Bernstein, and S. Zilberstein, “Dynamic programming for partially observable stochastic games,” in AAAI
2004
Earlier work this paper cites.
L. Matignon, G. J. Laurent, and N. Le Fort-Piat, “Hysteretic q-learning: an algorithm for decentralized reinforcement learning in cooperative multi-agent teams,” in 2007 IEEE/RSJ International Conference on Intelligent Robots and Systems
2007
Earlier work this paper cites.
M. Quigley, K. Conley, B. Gerkey, J. Faust, T. Foote, J. Leibs, R. Wheeler, A. Y. Ng, et al
2009
Earlier work this paper cites.
Springer Science & Business Media, 2009
M. Buehler, K. Iagnemma, and S. Singh, The DARPA urban challenge: autonomous vehicles in city traffic · 2009
Earlier work this paper cites.
D. Loiacono, P. L. Lanzi, J. Togelius, E. Onieva, D. A. Pelta, M. V. Butz, T. D. Lönneker, L. Cardamone, D. Perez, Y. Sáez, M. Preuss, and J. Quadflieg, “The 2009 simulated car racing championship,” IEEE Transactions on Computational Intelligence and AI in Games
2010
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in 2012 IEEE Conference on Computer Vision and Pattern Recognition
2012
Earlier work this paper cites.
J. Cameron, S. Myint, C. Kuo, A. Jain, H. Grip, P. Jayakumar, and J. Overholt, “Real-time and high-fidelity simulation environment for autonomous ground vehicle dynamics,” in Annual Ground Vehicle Systems Engineering And Technology Symposium (GVSETS) Symposium
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
S. Kar, J. M. F. Moura, and H. V. Poor, “ 𝒬𝒟 {{\cal Q}{\cal D}} -learning: A collaborative distributed strategy for multi-agent reinforcement learning through consensus + innovations {\rm consensus}+{\rm innovations} ,” IEEE Transactions on Signal Processing
2013
Earlier work this paper cites.
J. Zhang and S. Singh, “Loam: Lidar odometry and mapping in real-time.,” in Robotics: Science and systems
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
A. D. Ames, J. W. Grizzle, and P. Tabuada, “Control barrier function based quadratic programs with application to adaptive cruise control,” in 53rd IEEE Conference on Decision and Control
2014
Earlier work this paper cites.
B. Wymann, C. Dimitrakakis, A. Sumner, E. Espié, and C. Guionneau, “Torcs: The open racing car simulator,” 2015
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in Medical Image Computing and Computer-Assisted Intervention–MICCAI 2015: 18th International Conference, Munich, Germany, October 5-9, 2015, Proceedings, Part III 18
2015
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,” nature
2015
Earlier work this paper cites.
U. Briefs, “Mcity grand opening,” Research Review
2015
Earlier work this paper cites.
B. Paden, M. Čáp, S. Z. Yong, D. Yershov, and E. Frazzoli, “A survey of motion planning and control techniques for self-driving urban vehicles,” IEEE Transactions on Intelligent Vehicles
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
S. Sukhbaatar, R. Fergus, et al
2016
Earlier work this paper cites.
J. Foerster, I. A. Assael, N. De Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” Advances in neural information processing systems
2016
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International conference on machine learning
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” in International Conference on Learning Representations
2016
Earlier work this paper cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Proceedings of the AAAI conference on artificial intelligence
2016
Earlier work this paper cites.
E. Jang, S. Gu, and B. Poole, “Categorical reparameterization with gumbel-softmax,” in International Conference on Learning Representations
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
M. T. Ribeiro, S. Singh, and C. Guestrin, “” why should i trust you?” explaining the predictions of any classifier,” in Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining
2016
Earlier work this paper cites.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al
2017
Earlier work this paper cites.
J. Perolat, J. Z. Leibo, V. Zambaldi, C. Beattie, K. Tuyls, and T. Graepel, “A multi-agent reinforcement learning model of common-pool resource appropriation,” Advances in neural information processing systems
2017
Earlier work this paper cites.
R. Lowe, Y. WU, A. Tamar, J. Harb, O. Pieter Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Advances in Neural Information Processing Systems
2017
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine
2017
Earlier work this paper cites.
S. D. Pendleton, H. Andersen, X. Du, X. Shen, M. Meghjani, Y. H. Eng, D. Rus, and M. H. Ang, “Perception, planning, control, and coordination for autonomous vehicles,” Machines
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Computing Surveys (CSUR)
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “CARLA: An open urban driving simulator,” in Proceedings of the 1st Annual Conference on Robot Learning
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
L. Paull, J. Tani, H. Ahn, J. Alonso-Mora, L. Carlone, M. Cap, Y. F. Chen, C. Choi, J. Dusek, Y. Fang, et al
2017
Earlier work this paper cites.
J. Foerster, N. Nardelli, G. Farquhar, T. Afouras, P. H. Torr, P. Kohli, and S. Whiteson, “Stabilising experience replay for deep multi-agent reinforcement learning,” in International Conference on Machine Learning
2017
Earlier work this paper cites.
S. Omidshafiei, J. Pazis, C. Amato, J. P. How, and J. Vian, “Deep decentralized multi-task multi-agent reinforcement learning under partial observability,” in International Conference on Machine Learning
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems
2017
Earlier work this paper cites.
R. Lowe, Y. I. Wu, A. Tamar, J. Harb, O. Pieter Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” Advances in neural information processing systems
2017
Earlier work this paper cites.
J. Kim and J. Canny, “Interpretable learning for self-driving cars by visualizing causal attention,” in Proceedings of the IEEE international conference on computer vision
2017
Earlier work this paper cites.
MIT press, 2018
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction · 2018
Earlier work this paper cites.
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel, et al
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
E. Liang, R. Liaw, R. Nishihara, P. Moritz, R. Fox, K. Goldberg, J. Gonzalez, M. Jordan, and I. Stoica, “Rllib: Abstractions for distributed reinforcement learning,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
S. Shah, D. Dey, C. Lovett, and A. Kapoor, “Airsim: High-fidelity visual and physical simulation for autonomous vehicles,” in Field and Service Robotics: Results of the 11th International Conference
2018
Earlier work this paper cites.
J. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in Proceedings of the AAAI conference on artificial intelligence
2018
Earlier work this paper cites.
K. Zhang, Z. Yang, H. Liu, T. Zhang, and T. Basar, “Fully decentralized multi-agent reinforcement learning with networked agents,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
D. Ha and J. Schmidhuber, “Recurrent world models facilitate policy evolution,” in Advances in Neural Information Processing Systems
2018
Earlier work this paper cites.
J. Jiang and Z. Lu, “Learning attentional communication for multi-agent cooperation,” Advances in neural information processing systems
2018
Earlier work this paper cites.
P. Sunehag, G. Lever, A. Gruslys, W. M. Czarnecki, V. Zambaldi, M. Jaderberg, M. Lanctot, N. Sonnerat, J. Z. Leibo, K. Tuyls, et al
2018
Earlier work this paper cites.
P. K. Sharma and J. H. Park, “Blockchain based hybrid network architecture for the smart city,” Future Generation Computer Systems
2018
Earlier work this paper cites.
J. Kim, A. Rohrbach, T. Darrell, J. Canny, and Z. Akata, “Textual explanations for self-driving vehicles,” in Proceedings of the European conference on computer vision (ECCV)
2018
Earlier work this paper cites.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al
2019
Earlier work this paper cites.
M. Brittain and P. Wei, “Autonomous separation assurance in an high-density en route sector: A deep multi-agent reinforcement learning approach,” in 2019 IEEE Intelligent Transportation Systems Conference (ITSC)
2019
Earlier work this paper cites.
H. Zhang, S. Feng, C. Liu, Y. Ding, Y. Zhu, Z. Zhou, W. Zhang, Y. Yu, H. Jin, and Z. Li, “Cityflow: A multi-agent reinforcement learning environment for large scale city traffic scenario,” in The world wide web conference
2019
Earlier work this paper cites.
M. Samvelyan, T. Rashid, C. Schroeder de Witt, G. Farquhar, N. Nardelli, T. G. Rudner, C.-M. Hung, P. H. Torr, J. Foerster, and S. Whiteson, “The starcraft multi-agent challenge,” in Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems
2019
Earlier work this paper cites.
J. Chen, B. Yuan, and M. Tomizuka, “Model-free deep reinforcement learning for urban autonomous driving,” in 2019 IEEE Intelligent Transportation Systems Conference (ITSC)
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Jaderberg, W. M. Czarnecki, I. Dunning, L. Marris, G. Lever, A. G. Castaneda, C. Beattie, N. C. Rabinowitz, A. S. Morcos, A. Ruderman, et al
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
W. Kim, M. Cho, and Y. Sung, “Message-dropout: An efficient training method for multi-agent deep reinforcement learning,” in Proceedings of the AAAI conference on artificial intelligence
2019
Earlier work this paper cites.
G. Sartoretti, J. Kerr, Y. Shi, G. Wagner, T. K. S. Kumar, S. Koenig, and H. Choset, “Primal: Pathfinding via reinforcement and imitation multi-agent learning,” IEEE Robotics and Automation Letters
2019
Earlier work this paper cites.
W. Schwarting, A. Pierson, J. Alonso-Mora, S. Karaman, and D. Rus, “Social behavior for autonomous vehicles,” Proceedings of the National Academy of Sciences
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. Gambi, M. Mueller, and G. Fraser, “Automatically testing self-driving cars with search-based procedural content generation,” in Proceedings of the 28th ACM SIGSOFT International Symposium on Software Testing and Analysis
2019
Earlier work this paper cites.
F. Hauer, T. Schmidt, B. Holzmüller, and A. Pretschner, “Did we test all scenarios for automated and autonomous driving systems?,” in 2019 IEEE Intelligent Transportation Systems Conference (ITSC)
2019
Earlier work this paper cites.
A. Raffin, A. Hill, R. Traoré, T. Lesort, N. Díaz-Rodríguez, and D. Filliat, “Decoupling feature extraction from policy learning: assessing benefits of state representation learning in goal based robotics,” in ICLR Workshop on Structure and Priors in Reinforcement Learning
2019
Cited alongside, same era.
R. Traoré, H. Caselles-Dupré, T. Lesort, T. Sun, G. Cai, D. Filliat, and N. Díaz-Rodríguez, “Discorl: Continual reinforcement learning via policy distillation,” in NeurIPS Workshop on Deep Reinforcement Learning
2019
Cited alongside, same era.
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” in International Conference on Learning Representations
2019
Cited alongside, same era.
J. D. M.-W. C. Kenton and L. K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of NAACL-HLT
2019
Cited alongside, same era.
“Isaac sim.” [Online]. Available: https://developer.nvidia.com/isaac-sim , 2022
2022
Later among the works it cites.
S. Macenski, T. Foote, B. Gerkey, C. Lalancette, and W. Woodall, “Robot operating system 2: Design, architecture, and uses in the wild,” Science Robotics
2022
Later among the works it cites.
Q. Sun, X. Huang, B. C. Williams, and H. Zhao, “Intersim: Interactive traffic simulation via explicit relation modeling,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2022
Later among the works it cites.
E. Vinitsky, N. Lichtlé, X. Yang, B. Amos, and J. Foerster, “Nocturne: a scalable driving benchmark for bringing multi-agent learning one step closer to the real world,” Advances in Neural Information Processing Systems
2022
Later among the works it cites.
W. Mao, L. Yang, K. Zhang, and T. Basar, “On improving model-free algorithms for decentralized multi-agent reinforcement learning,” in International Conference on Machine Learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. T. Nguyen, N. D. Nguyen, and S. Nahavandi, “Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications,” IEEE Transactions on Cybernetics
2020
Cited alongside, same era.
G. Chen, H. Cao, J. Conradt, H. Tang, F. Rohrbein, and A. Knoll, “Event-based neuromorphic vision for autonomous driving: A paradigm shift for bio-inspired visual sensing and perception,” IEEE Signal Processing Magazine
2020
Cited alongside, same era.
S. Aradi, “Survey of deep reinforcement learning for motion planning of autonomous vehicles,” IEEE Transactions on Intelligent Transportation Systems
2020
Cited alongside, same era.
T. Chu, J. Wang, L. Codecà, and Z. Li, “Multi-agent deep reinforcement learning for large-scale traffic signal control,” IEEE Transactions on Intelligent Transportation Systems
2020
Cited alongside, same era.
2020
Cited alongside, same era.
PhD thesis, Université de Lille, 2020
E. Leurent, Safe and efficient reinforcement learning for behavioural planning in autonomous driving · 2020
Cited alongside, same era.
2020
Cited alongside, same era.
P. Palanisamy, “Multi-agent connected autonomous driving using deep reinforcement learning,” in 2020 International Joint Conference on Neural Networks (IJCNN)
2020
Cited alongside, same era.
2022
Later among the works it cites.
V. Egorov and A. Shpilman, “Scalable multi-agent model-based reinforcement learning,” in Autonomous Agents and Multiagent Systems
2022
Later among the works it cites.
C. Yu, A. Velu, E. Vinitsky, J. Gao, Y. Wang, A. Bayen, and Y. Wu, “The surprising effectiveness of ppo in cooperative multi-agent games,” Advances in Neural Information Processing Systems
2022
Later among the works it cites.
L. Brunke, M. Greeff, A. W. Hall, Z. Yuan, S. Zhou, J. Panerati, and A. P. Schoellig, “Safe learning in robotics: From learning-based control to safe reinforcement learning,” Annual Review of Control, Robotics, and Autonomous Systems
2022
Later among the works it cites.
B. Toghi, R. Valiente, D. Sadigh, R. Pedarsani, and Y. P. Fallah, “Social coordination and altruism in autonomous driving,” IEEE Transactions on Intelligent Transportation Systems
2022
Later among the works it cites.
E. Candela, L. Parada, L. Marques, T.-A. Georgescu, Y. Demiris, and P. Angeloudis, “Transferring multi-agent reinforcement learning policies for autonomous driving using sim-to-real,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2022
Later among the works it cites.
X. Bai, Z. Hu, X. Zhu, Q. Huang, Y. Chen, H. Fu, and C.-L. Tai, “Transfusion: Robust lidar-camera fusion for 3d object detection with transformers,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
2022
Later among the works it cites.
A. Brunnbauer, L. Berducci, A. Brandstátter, M. Lechner, R. Hasani, D. Rus, and R. Grosu, “Latent imagination facilitates zero-shot transfer in autonomous racing,” in 2022 International Conference on Robotics and Automation (ICRA)
2022
Later among the works it cites.
T. Liang, H. Xie, K. Yu, Z. Xia, Z. Lin, Y. Wang, T. Tang, B. Wang, and Z. Tang, “Bevfusion: A simple and robust lidar-camera fusion framework,” Advances in Neural Information Processing Systems
2022
Later among the works it cites.
2022
Later among the works it cites.
X. Fang, Q. Zhang, Y. Gao, and D. Zhao, “Offline reinforcement learning for autonomous driving with real world driving data,” in 2022 IEEE 25th International Conference on Intelligent Transportation Systems (ITSC)
2022
Later among the works it cites.
X. Zhang, N. Tseng, A. Syed, R. Bhasin, and N. Jaipuria, “Simbar: Single image-based scene relighting for effective data augmentation for automated driving vision tasks,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2022
Later among the works it cites.
X. Wu, L. Xiao, Y. Sun, J. Zhang, T. Ma, and L. He, “A survey of human-in-the-loop for machine learning,” Future Generation Computer Systems
2022
Later among the works it cites.
Z. Peng, Q. Li, C. Liu, and B. Zhou, “Safe driving via expert guided policy optimization,” in Conference on Robot Learning
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Snell, D. Klein, and R. Zhong, “Learning by distilling context,” arXiv preprint arXiv:2209.15189
2022
Later among the works it cites.
S. Teng, X. Hu, P. Deng, B. Li, Y. Li, Y. Ai, D. Yang, L. Li, Z. Xuanyuan, F. Zhu, and L. Chen, “Motion planning for autonomous driving: The state of the art and future perspectives,” IEEE Transactions on Intelligent Vehicles
2023
Later among the works it cites.
Y. Song, A. Romero, M. Müller, V. Koltun, and D. Scaramuzza, “Reaching the limit in autonomous racing: Optimal control versus reinforcement learning,” Science Robotics
2023
Later among the works it cites.
E. Kaufmann, L. Bauersfeld, A. Loquercio, M. Müller, V. Koltun, and D. Scaramuzza, “Champion-level drone racing using deep reinforcement learning,” Nature
2023
Later among the works it cites.
L. Ding, Z. Lin, X. Shi, and G. Yan, “Target-value-competition-based multi-agent deep reinforcement learning algorithm for distributed nonconvex economic dispatch,” IEEE Transactions on Power Systems
2023
Later among the works it cites.
L. Chen, Y. Li, C. Huang, B. Li, Y. Xing, D. Tian, L. Li, Z. Hu, X. Na, Z. Li, S. Teng, C. Lv, J. Wang, D. Cao, N. Zheng, and F.-Y. Wang, “Milestones in autonomous driving and intelligent vehicles: Survey of surveys,” IEEE Transactions on Intelligent Vehicles
2023
Later among the works it cites.
L. Chen, Y. Li, C. Huang, Y. Xing, D. Tian, L. Li, Z. Hu, S. Teng, C. Lv, J. Wang, D. Cao, N. Zheng, and F.-Y. Wang, “Milestones in autonomous driving and intelligent vehicles—part i: Control, computing system design, communication, hd map, testing, and human behaviors,” IEEE Transactions on Systems, Man, and Cybernetics: Systems
2023
Later among the works it cites.
L. Chen, S. Teng, B. Li, X. Na, Y. Li, Z. Li, J. Wang, D. Cao, N. Zheng, and F.-Y. Wang, “Milestones in autonomous driving and intelligent vehicles—part ii: Perception and planning,” IEEE Transactions on Systems, Man, and Cybernetics: Systems
2023
Later among the works it cites.
L. Chen, P. Wu, K. Chitta, B. Jaeger, A. Geiger, and H. Li, “End-to-end autonomous driving: Challenges and frontiers,” arXiv
2023
Later among the works it cites.
P. Yadav, A. Mishra, and S. Kim, “A comprehensive survey on multi-agent reinforcement learning for connected and automated vehicles,” Sensors
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Zhou, Z. Wan, H. Wang, M. Wen, R. Wu, Y. Wen, Y. Yang, Y. Yu, J. Wang, and W. Zhang, “Malib: A parallel framework for population-based multi-agent reinforcement learning,” Journal of Machine Learning Research
2023
Later among the works it cites.
J. Guan, G. Chen, J. Huang, Z. Li, L. Xiong, J. Hou, and A. Knoll, “A discrete soft actor-critic decision-making strategy with sample filter for freeway autonomous driving,” IEEE Transactions on Vehicular Technology
2023
Later among the works it cites.
Z. Dai, T. Zhou, K. Shao, D. H. Mguni, B. Wang, and H. Jianye, “Socially-attentive policy optimization in multi-agent self-driving system,” in Conference on Robot Learning
2023
Later among the works it cites.
Q. Li, Z. Peng, L. Feng, Q. Zhang, Z. Xue, and B. Zhou, “Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2023
Later among the works it cites.
L. Feng, Q. Li, Z. Peng, S. Tan, and B. Zhou, “Trafficgen: Learning to generate diverse and realistic traffic scenarios,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
S. Lan, Z. Wang, E. Wei, A. K. Roy-Chowdhury, and Q. Zhu, “Collaborative multi-agent video fast-forwarding,” IEEE Transactions on Multimedia
2023
Later among the works it cites.
D. Ye, T. Zhu, C. Zhu, W. Zhou, and P. S. Yu, “Model-based self-advising for multi-agent learning,” IEEE Transactions on Neural Networks and Learning Systems
2023
Later among the works it cites.
D. Xu, Y. Chen, B. Ivanovic, and M. Pavone, “Bits: Bi-level imitation for traffic simulation,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
Q. Li, Z. M. Peng, L. Feng, Z. Liu, C. Duan, W. Mo, and B. Zhou, “Scenarionet: Open-source platform for large-scale traffic scenario simulation and modeling,” in Advances in Neural Information Processing Systems
2023
Later among the works it cites.
K. Jiang, W. Liu, Y. Wang, L. Dong, and C. Sun, “Credit assignment in heterogeneous multi-agent reinforcement learning for fully cooperative tasks,” Applied Intelligence
2023
Later among the works it cites.
W. Chen, W. Li, X. Liu, S. Yang, and Y. Gao, “Learning explicit credit assignment for cooperative multi-agent reinforcement learning via polarization policy gradient,” in Proceedings of the AAAI Conference on Artificial Intelligence
2023
Later among the works it cites.
N. Gupta, G. Srinivasaraghavan, S. Mohalik, N. Kumar, and M. E. Taylor, “Hammer: Multi-level coordination of reinforcement learning agents via learned messaging,” Neural Computing and Applications
2023
Later among the works it cites.
C. Huang, J. Zhao, H. Zhou, H. Zhang, X. Zhang, and C. Ye, “Multi-agent decision-making at unsignalized intersections with reinforcement learning from demonstrations,” in 2023 IEEE Intelligent Vehicles Symposium (IV)
2023
Later among the works it cites.
S. Gu, J. G. Kuba, Y. Chen, Y. Du, L. Yang, A. Knoll, and Y. Yang, “Safe multi-agent reinforcement learning for multi-robot control,” Artificial Intelligence
2023
Later among the works it cites.
2023
Later among the works it cites.
L. Chen, Y. Wang, Z. Miao, Y. Mo, M. Feng, Z. Zhou, and H. Wang, “Transformer-based imitative reinforcement learning for multirobot path planning,” IEEE Transactions on Industrial Informatics
2023
Later among the works it cites.
C. Dawson, S. Gao, and C. Fan, “Safe control with learned certificates: A survey of neural lyapunov, barrier, and contraction methods for robotics and control,” IEEE Transactions on Robotics
2023
Later among the works it cites.
A. Singh, “Transformer-based sensor fusion for autonomous driving: A survey,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2023
Later among the works it cites.
X. Chen, T. Zhang, Y. Wang, Y. Wang, and H. Zhao, “Futr3d: A unified sensor fusion framework for 3d detection,” in proceedings of the IEEE/CVF conference on computer vision and pattern recognition
2023
Later among the works it cites.
Z. Liu, H. Tang, A. Amini, X. Yang, H. Mao, D. L. Rus, and S. Han, “Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
Y. Zhang, B. Kang, B. Hooi, S. Yan, and J. Feng, “Deep long-tailed learning: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2023
Later among the works it cites.
L. Montaut, Q. L. Lidec, A. Bambade, V. Petrik, J. Sivic, and J. Carpentier, “Differentiable collision detection: a randomized smoothing approach,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
Z. Liu, G. Chen, Z. Li, Y. Kang, S. Qu, and C. Jiang, “Psdc: A prototype-based shared-dummy classifier model for open-set domain adaptation,” IEEE Transactions on Cybernetics
2023
Later among the works it cites.
K. Chitta, A. Prakash, B. Jaeger, Z. Yu, K. Renz, and A. Geiger, “Transfuser: Imitation with transformer-based sensor fusion for autonomous driving,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2023
Later among the works it cites.
Z. Zheng, Y. Cheng, Z. Xin, Z. Yu, and B. Zheng, “Robust perception under adverse conditions for autonomous driving based on data augmentation,” IEEE Transactions on Intelligent Transportation Systems
2023
Later among the works it cites.
L. Meng, M. Wen, C. Le, X. Li, D. Xing, W. Zhang, Y. Wen, H. Zhang, J. Wang, Y. Yang, et al
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
C. Gulino, J. Fu, W. Luo, G. Tucker, E. Bronstein, Y. Lu, J. Harb, X. Pan, Y. Wang, X. Chen, et al
2024
Closest in time.
D. Lee, C. Eom, and M. Kwon, “Ad4rl: Autonomous driving benchmarks for offline reinforcement learning with value-based dataset,” in 2024 IEEE International Conference on Robotics and Automation (ICRA)
2024
Closest in time.
C. Zhu, M. Dastani, and S. Wang, “A survey of multi-agent deep reinforcement learning with communication,” Autonomous Agents and Multi-Agent Systems
2024
Closest in time.
S. Bhattacharya, S. Kailas, S. Badyal, S. Gil, and D. Bertsekas, “Multiagent reinforcement learning: Rollout and policy iteration for pomdp with application to multirobot problems,” IEEE Transactions on Robotics
2024
Closest in time.
D. Ying, Y. Zhang, Y. Ding, A. Koppel, and J. Lavaei, “Scalable primal-dual actor-critic method for safe multi-agent rl with general utilities,” Advances in Neural Information Processing Systems
2024
Closest in time.
2024
Closest in time.
S. Han, S. Zhou, J. Wang, L. Pepin, C. Ding, J. Fu, and F. Miao, “A multi-agent reinforcement learning approach for safe and efficient behavior planning of connected autonomous vehicles,” IEEE Transactions on Intelligent Transportation Systems
2024
Closest in time.
B. Cai, C. Wei, and Z. Ji, “Deep reinforcement learning with multiple unrelated rewards for agv mapless navigation,” IEEE Transactions on Automation Science and Engineering
2024
Closest in time.
R. Zhao, Y. Li, F. Gao, Z. Gao, and T. Zhang, “Multi-agent constrained policy optimization for conflict-free management of connected autonomous vehicles at unsignalized intersections,” IEEE Transactions on Intelligent Transportation Systems
2024
Closest in time.
2024
Closest in time.
L. Wang, X. Zhang, H. Su, and J. Zhu, “A comprehensive survey of continual learning: Theory, method and application,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2024
Closest in time.
X. Hu, S. Li, T. Huang, B. Tang, R. Huai, and L. Chen, “How simulation helps autonomous driving: A survey of sim2real, digital twins, and parallel intelligence,” IEEE Transactions on Intelligent Vehicles
2024
Closest in time.
J. Li, Z. Yu, Z. Du, L. Zhu, and H. T. Shen, “A comprehensive survey on source-free domain adaptation,” IEEE Transactions on Pattern Analysis and Machine Intelligence
2024
Closest in time.
H. Lin, W. Ding, Z. Liu, Y. Niu, J. Zhu, Y. Niu, and D. Zhao, “Safety-aware causal representation for trustworthy offline reinforcement learning in autonomous driving,” IEEE Robotics and Automation Letters
2024
Closest in time.
Y. Yuan, J. Hao, Y. Ma, Z. Dong, H. Liang, J. Liu, Z. Feng, K. Zhao, and Y. Zheng, “Uni-rlhf: Universal platform and benchmark suite for reinforcement learning with diverse human feedback,” in International Conference on Learning Representations
2024
Closest in time.
2024
Closest in time.
L. Wen, X. Yang, D. Fu, X. Wang, P. Cai, X. Li, M. Tao, Y. Li, X. Linran, D. Shang, et al
2024
Closest in time.
C. Cui, Y. Ma, X. Cao, W. Ye, and Z. Wang, “Drive as you speak: Enabling human-like interaction with large language models in autonomous vehicles,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision
2024
Closest in time.