Fetching the paper…
Reading the bibliography…
The combination of Formal Methods with Reinforcement Learning (RL) has recently attracted interest as a way for single-agent RL to learn multiple-task specifications.
Michael O Rabin and Dana Scott, ‘Finite automata and their decision problems’, IBM journal of research and development
1959
Earlier work this paper cites.
Amir Pnueli, ‘The temporal logic of programs’, in 18th Annual Symposium on Foundations of Computer Science (sfcs 1977)
1977
Earlier work this paper cites.
JCH Christopher, ‘Watkins and peter dayan’, Q-Learning. Machine Learning
1992
Earlier work this paper cites.
Ming Tan, ‘Multi-agent reinforcement learning: Independent vs. cooperative agents’, in Proceedings of the tenth international conference on machine learning
1993
Earlier work this paper cites.
Michael L Littman, ‘Markov games as a framework for multi-agent reinforcement learning’, in Machine learning proceedings 1994
1994
Earlier work this paper cites.
Fahiem Bacchus, Craig Boutilier, and Adam Grove, ‘Rewarding behaviors’, in Proceedings of the National Conference on Artificial Intelligence
1996
Earlier work this paper cites.
Junling Hu, Michael P Wellman, et al., ‘Multiagent reinforcement learning: theoretical framework and an algorithm.’, in ICML
1998
Earlier work this paper cites.
Andrew Y Ng, Daishi Harada, and Stuart Russell, ‘Policy invariance under reward transformations: Theory and application to reward shaping’, in ICML
1999
Earlier work this paper cites.
Fahiem Bacchus and Froduald Kabanza, ‘Using temporal logics to express search control knowledge for planning’, Artificial intelligence
2000
Earlier work this paper cites.
Orna Kupferman and Moshe Y Vardi, ‘Model checking of safety properties’, Formal Methods in System Design
2001
Earlier work this paper cites.
2002
Earlier work this paper cites.
Xavier Glorot, Antoine Bordes, and Yoshua Bengio, ‘Deep sparse rectifier neural networks’, in Proceedings of the fourteenth international conference on artificial intelligence and statistics
2011
Earlier work this paper cites.
Yongcan Cao, Wenwu Yu, Wei Ren, and Guanrong Chen, ‘An overview of recent progress in the study of distributed multi-agent coordination’, IEEE Transactions on Industrial informatics
2012
Earlier work this paper cites.
Giuseppe De Giacomo and Moshe Y Vardi, ‘Linear temporal logic and linear dynamic logic on finite traces’, in Twenty-Third International Joint Conference on Artificial Intelligence
2013
Cited alongside, same era.
2014
Cited alongside, same era.
Bruno Lacerda, David Parker, and Nick Hawes, ‘Optimal policy generation for partially satisfiable co-safe ltl specifications’, in Twenty-Fourth International Joint Conference on Artificial Intelligence
2015
Cited alongside, same era.
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al., ‘Human-level control through deep reinforcement learning’, Nature
2015
Cited alongside, same era.
2017
Later among the works it cites.
George Rupert Mason, Radu Constantin Calinescu, Daniel Kudenko, and Alec Banks, ‘Assured reinforcement learning with formally verified abstract policies’, in 9th International Conference on Agents and Artificial Intelligence (ICAART)
2017
Later among the works it cites.
2017
Later among the works it cites.
Ardi Tampuu, Tambet Matiisen, Dorian Kodelja, Ilya Kuzovkin, Kristjan Korjus, Juhan Aru, Jaan Aru, and Raul Vicente, ‘Multiagent cooperation and competition with deep reinforcement learning’, PloS one
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
Dayong Ye, Minjie Zhang, and Yun Yang, ‘A multi-agent framework for packet routing in wireless sensor networks’, sensors
2015
Cited alongside, same era.
Sebastian Junges, Nils Jansen, Christian Dehnert, Ufuk Topcu, and Joost-Pieter Katoen, ‘Safety-constrained reinforcement learning for mdps’, in International Conference on Tools and Algorithms for the Construction and Analysis of Systems
2016
Cited alongside, same era.
2016
Cited alongside, same era.
Jacob Andreas, Dan Klein, and Sergey Levine, ‘Modular multitask reinforcement learning with policy sketches’, in Proceedings of the 34th International Conference on Machine Learning-Volume 70
2017
Cited alongside, same era.
Openai baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov · 2017
Cited alongside, same era.
Jakob Foerster, Nantas Nardelli, Gregory Farquhar, Triantafyllos Afouras, Philip HS Torr, Pushmeet Kohli, and Shimon Whiteson, ‘Stabilising experience replay for deep multi-agent reinforcement learning’, in Proceedings of the 34th International Conference on Machine Learning-Volume 70
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Min Wen, Ivan Papusha, and Ufuk Topcu, ‘Learning from demonstrations with high-level side information’, in Proceedings of the Twenty-Sixth International Joint Conference on Artificial Intelligence
2017
Later among the works it cites.
Mohammed Alshiekh, Roderick Bloem, Rüdiger Ehlers, Bettina Könighofer, Scott Niekum, and Ufuk Topcu, ‘Safe reinforcement learning via shielding’, in Thirty-Second AAAI Conference on Artificial Intelligence
2018
Later among the works it cites.
Ronen I Brafman, Giuseppe De Giacomo, and Fabio Patrizi, ‘Ltlf/ldlf non-markovian rewards’, in Thirty-Second AAAI Conference on Artificial Intelligence
2018
Later among the works it cites.
2018
Later among the works it cites.
Devaprakash Muniraj, Kyriakos G Vamvoudakis, and Mazen Farhood, ‘Enforcing signal temporal logic specifications in multi-agent adversarial environments: A deep q-learning approach’, in 2018 IEEE Conference on Decision and Control (CDC)
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Rodrigo Toro Icarte, Toryn Q Klassen, Richard Valenzano, and Sheila A McIlraith, ‘Teaching multiple tasks to an rl agent using ltl’, in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems
2018
Later among the works it cites.