Fetching the paper…
Reading the bibliography…
In this paper, we study the cooperative Multi-Agent Reinforcement Learning (MARL) problems using Reward Machines (RMs) to specify the reward functions such that the prior knowledge of high-level events in a task can be leveraged to facilitate the learning efficiency.
The temporal logic of programs
Amir Pnueli · 1977
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Hierarchical multi-agent reinforcement learning
Mohammad Ghavamzadeh, Sridhar Mahadevan, and Rajbala Makar · 2006
Earlier work this paper cites.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Rob Fergus, et al · 2016
Earlier work this paper cites.
Modular multitask reinforcement learning with policy sketches
Jacob Andreas, Dan Klein, and Sergey Levine · 2017
Earlier work this paper cites.
The option-critic architecture
Pierre-Luc Bacon, Jean Harb, and Doina Precup · 2017
Earlier work this paper cites.
A survey of learning in multiagent environments: Dealing with non-stationarity
Pablo Hernandez-Leal, Michael Kaisers, Tim Baarslag, and Enrique Munoz de Cote · 2017
Earlier work this paper cites.
Reinforcement learning with temporal logic rewards
Xiao Li, Cristian-Ioan Vasile, and Calin Belta · 2017
Earlier work this paper cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Earlier work this paper cites.
Reinforcement learning for ltlf/ldlf goals
Giuseppe De Giacomo, Luca Iocchi, Marco Favorito, and Fabio Patrizi · 2018
Earlier work this paper cites.
Diversity is all you need: Learning skills without a reward function
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2018
Earlier work this paper cites.
Using reward machines for high-level task specification and decomposition in reinforcement learning
Rodrigo Toro Icarte, Toryn Klassen, Richard Valenzano, and Sheila McIlraith · 2018
Earlier work this paper cites.
Enforcing signal temporal logic specifications in multi-agent adversarial environments: A deep q-learning approach
Devaprakash Muniraj, Kyriakos G Vamvoudakis, and Mazen Farhood · 2018
Earlier work this paper cites.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson · 2018
Earlier work this paper cites.
Hierarchical deep multiagent reinforcement learning with temporal abstraction
Hongyao Tang, Jianye Hao, Tangjie Lv, Yingfeng Chen, Zongzhang Zhang, Hangtian Jia, Chunxu Ren, Yan Zheng, Zhaopeng Meng, Changjie Fan, et al · 2018
Cited alongside, same era.
Teaching multiple tasks to an rl agent using ltl
Rodrigo Toro Icarte, Toryn Q Klassen, Richard Valenzano, and Sheila A McIlraith · 2018
Cited alongside, same era.
Feudal multi-agent hierarchies for cooperative reinforcement learning
Sanjeevan Ahilan and Peter Dayan · 2019
Cited alongside, same era.
Ltl and beyond: Formal languages for reward function specification in reinforcement learning
Alberto Camacho, Rodrigo Toro Icarte, Toryn Q Klassen, Richard Anthony Valenzano, and Sheila A McIlraith · 2019
Cited alongside, same era.
Induction of subgoal automata for reinforcement learning
Daniel Furelos-Blanco, Mark Law, Alessandra Russo, Krysia Broda, and Anders Jonsson · 2020
Later among the works it cites.
Reinforcement learning with non-markovian rewards
Maor Gaon and Ronen Brafman · 2020
Later among the works it cites.
Extended markov games to learn multiple tasks in multi-agent reinforcement learning
Borja G Leon and Francesco Belardinelli · 2020
Later among the works it cites.
Systematic generalisation through task temporal logic and deep reinforcement learning
Borja G León, Murray Shanahan, and Francesco Belardinelli · 2020
Later among the works it cites.
Learning non-markovian reward models in mdps
Gavin Rens and Jean-François Raskin · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jhelum Chakravorty, Nadeem Ward, Julien Roy, Maxime Chevalier-Boisvert, Sumana Basu, Andrei Lupu, and Doina Precup · 2019
Cited alongside, same era.
A composable specification language for reinforcement learning tasks
Kishor Jothimurugan, Rajeev Alur, and Osbert Bastani · 2019
Cited alongside, same era.
Learning to coordinate manipulation skills via skill behavior diversification
Youngwoon Lee, Jingyun Yang, and Joseph J Lim · 2019
Cited alongside, same era.
Dealing with non-stationarity in multi-agent deep reinforcement learning
Georgios Papoudakis, Filippos Christianos, Arrasy Rahman, and Stefano V Albrecht · 2019
Cited alongside, same era.
Learning reward machines for partially observable reinforcement learning
Rodrigo Toro Icarte, Ethan Waldie, Toryn Klassen, Rick Valenzano, Margarita Castro, and Sheila McIlraith · 2019
Cited alongside, same era.
Influence-based multi-agent exploration
Tonghan Wang, Jianhao Wang, Yi Wu, and Chongjie Zhang · 2019
Cited alongside, same era.
Hierarchical cooperative multi-agent reinforcement learning with skill discovery
Jiachen Yang, Igor Borovikov, and Hongyuan Zha · 2019
Cited alongside, same era.
Modular deep reinforcement learning with temporal logic specifications
Lim Zun Yuan, Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2019
Cited alongside, same era.
Reward machines for vision-based robotic manipulation
Alberto Camacho, Jacob Varley, Andy Zeng, Deepali Jain, Atil Iscen, and Dmitry Kalashnikov · 2021
Later among the works it cites.
Learning quadruped locomotion policies with reward machines
David DeFazio and Shiqi Zhang · 2021
Later among the works it cites.
Decentralized graph-based multi-agent reinforcement learning using reward machines
Jueming Hu, Zhe Xu, Weichang Wang, Guannan Qu, Yutian Pang, and Yongming Liu · 2021
Later among the works it cites.
Reward machines for cooperative multi-agent reinforcement learning
Cyrus Neary, Zhe Xu, Bo Wu, and Ufuk Topcu · 2021
Later among the works it cites.
Hierarchical reinforcement learning: A comprehensive survey
Shubham Pateria, Budhitama Subagdja, Ah-hwee Tan, and Chai Quek · 2021
Later among the works it cites.
Learning probabilistic reward machines from non-markovian stochastic reward processes
Alvaro Velasquez, Andre Beckus, Taylor Dohmen, Ashutosh Trivedi, Noah Topper, and George Atia · 2021
Later among the works it cites.
Active finite reward automaton inference and reinforcement learning using queries and counterexamples
Zhe Xu, Bo Wu, Aditya Ojha, Daniel Neider, and Ufuk Topcu · 2021
Later among the works it cites.
Multi-agent deep reinforcement learning: a survey
Sven Gronauer and Klaus Diepold · 2022
Later among the works it cites.