Fetching the paper…
Reading the bibliography…
Multi-agent reinforcement learning (MARL) has been increasingly used in a wide range of safety-critical applications, which require guaranteed safety (e.g., no unsafe states are ever visited) during the learning process.Unfortunately, current MARL methods do not have safety guarantees.
The temporal logic of programs. In 18th Annual Symposium on Foundations of Computer Science (sfcs 1977) . IEEE, 46–57
Amir Pnueli. 1977 · 1977
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents. In Proceedings of the tenth international conference on machine learning . 330–337
Ming Tan. 1993 · 1993
Earlier work this paper cites.
Model checking of safety properties
Orna Kupferman and Moshe Y Vardi. 2001 · 2001
Earlier work this paper cites.
Principles of model checking
Christel Baier and Joost-Pieter Katoen. 2008 · 2008
Earlier work this paper cites.
Temporal-logic-based reactive mission and motion planning
Hadas Kress-Gazit, Georgios E Fainekos, and George J Pappas. 2009 · 2009
Earlier work this paper cites.
Learning of coordination: Exploiting sparse interactions in multiagent systems. In Proceedings of The 8th International Conference on Autonomous Agents and Multiagent Systems-Volume 2 . 773–780
Francisco S Melo and Manuela Veloso. 2009 · 2009
Earlier work this paper cites.
Learning multi-agent state space representations. In Proceedings of the 9th International Conference on Autonomous Agents and Multiagent Systems: volume 1-Volume 1 . 715–722
Yann-Michaël De Hauwere, Peter Vrancx, and Ann Nowé. 2010 · 2010
Earlier work this paper cites.
Optimality and robustness in multi-robot path planning with temporal logic constraints
Alphan Ulusoy, Stephen L Smith, Xu Chu Ding, Calin Belta, and Daniela Rus. 2013 · 2013
Earlier work this paper cites.
Principles of cyber-physical systems
Rajeev Alur. 2015 · 2015
Earlier work this paper cites.
Shield synthesis. In International Conference on Tools and Algorithms for the Construction and Analysis of Systems . Springer, 533–548
Roderick Bloem, Bettina Könighofer, Robert Könighofer, and Chao Wang. 2015 · 2015
Cited alongside, same era.
A comprehensive survey on safe reinforcement learning
Javier Garcıa and Fernando Fernández. 2015 · 2015
Cited alongside, same era.
Slugs: Extensible gr (1) synthesis. In International Conference on Computer Aided Verification . Springer, 333–339
Rüdiger Ehlers and Vasumathi Raman. 2016 · 2016
Cited alongside, same era.
Safe, multi-agent, reinforcement learning for autonomous driving
Shai Shalev-Shwartz, Shaked Shammah, and Amnon Shashua. 2016 · 2016
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments. In Advances in neural information processing systems . 6379–6390
Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch. 2017 · 2017
A survey and critique of multiagent deep reinforcement learning
Pablo Hernandez-Leal, Bilal Kartal, and Matthew E Taylor. 2019 · 2019
Later among the works it cites.
Decentralized runtime synthesis of shields for multi-agent systems
Dhananjay Raju, Suda Bharadwaj, and Ufuk Topcu. 2019 · 2019
Later among the works it cites.
CM3: Cooperative Multi-goal Multi-stage Multi-agent Reinforcement Learning. In International Conference on Learning Representations
Jiachen Yang, Alireza Nakhaei, David Isele, Kikuo Fujimura, and Hongyuan Zha. 2019 · 2019
Later among the works it cites.
Coordinated Multiagent Reinforcement Learning for Teams of Mobile Sensing Robots. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . 2297–2299
Chao Yu, Xin Wang, and Zhanbo Feng. 2019 · 2019
Later among the works it cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Safe Reinforcement Learning via Shielding. In AAAI-18: 32nd AAAI Conference on Artificial Intelligence . 2669–2678
Mohammed Alshiekh, Roderick Bloem, Rüdiger Ehlers, Bettina Könighofer, Scott Niekum, and Ufuk Topcu. 2018 · 2018
Cited alongside, same era.
Synthesis of Minimum-Cost Shields for Multi-agent Systems. In 2019 American Control Conference (ACC) . IEEE, 1048–1055
Suda Bharadwaj, Roderik Bloem, Rayna Dimitrova, Bettina Konighofer, and Ufuk Topcu. 2019 · 2019
Cited alongside, same era.
Omega-regular objectives in model-free reinforcement learning. In International Conference on Tools and Algorithms for the Construction and Analysis of Systems . Springer, 395–412
Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi, Ashutosh Trivedi, and Dominik Wojtczak. 2019 · 2019
Cited alongside, same era.
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar. 2019 · 2019
Later among the works it cites.
Control synthesis from linear temporal logic specifications using model-free reinforcement learning. In 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 10349–10355
Alper Kamil Bozkurt, Yu Wang, Michael M Zavlanos, and Miroslav Pajic. 2020 · 2020
Later among the works it cites.
Cautious Reinforcement Learning with Logical Constraints. In Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems . 483–491
Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening. 2020 · 2020
Later among the works it cites.
Hierarchical Multiagent Reinforcement Learning for Maritime Traffic Management. In Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems . 1278–1286
Arambam James Singh, Akshat Kumar, and Hoong Chuin Lau. 2020 · 2020
Later among the works it cites.