Fetching the paper…
Reading the bibliography…
In this study, we address the challenge of learning generalizable policies for compositional tasks defined by logical specifications.
Benchmarking safe exploration in deep reinforcement learning
Ray, A.; Achiam, J.; and Amodei, D. 2019 · 1910
Earlier work this paper cites.
The temporal logic of programs
Pnueli, A. 1977 · 1977
Earlier work this paper cites.
TALplanner: A temporal logic based forward chaining planner
Kvarnström, J.; and Doherty, P. 2000 · 2000
Earlier work this paper cites.
Systematic generalisation through task temporal logic and deep reinforcement learning
León, B. G.; Shanahan, M.; and Belardinelli, F. 2020 · 2006
Earlier work this paper cites.
The graph neural network model
Scarselli, F.; Gori, M.; Tsoi, A. C.; Hagenbuchner, M.; and Monfardini, G. 2008 · 2008
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Taylor, M. E.; and Stone, P. 2009 · 2009
Earlier work this paper cites.
Translating embeddings for modeling multi-relational data
Bordes, A.; Usunier, N.; Garcia-Duran, A.; Weston, J.; and Yakhnenko, O. 2013 · 2013
Earlier work this paper cites.
Linear temporal logic and linear dynamic logic on finite traces
De Giacomo, G.; and Vardi, M. Y. 2013 · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Mnih, V. 2013 · 2013
Earlier work this paper cites.
Towards manipulation planning with temporal logic specifications
He, K.; Lahijanian, M.; Kavraki, L. E.; and Vardi, M. Y. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Earlier work this paper cites.
Modular multitask reinforcement learning with policy sketches
Andreas, J.; Klein, D.; and Levine, S. 2017 · 2017
Earlier work this paper cites.
Environment-independent task specifications via GLTL
Littman, M. L.; Topcu, U.; Fu, J.; Isbell, C.; Wen, M.; and MacGlashan, J. 2017 · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J.; Wolski, F.; Dhariwal, P.; Radford, A.; and Klimov, O. 2017 · 2017
Cited alongside, same era.
Using reward machines for high-level task specification and decomposition in reinforcement learning
Icarte, R. T.; Klassen, T.; Valenzano, R.; and McIlraith, S. 2018 · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Cited alongside, same era.
Teaching multiple tasks to an RL agent using LTL
Toro Icarte, R.; Klassen, T. Q.; Valenzano, R.; and McIlraith, S. A. 2018 · 2018
Cited alongside, same era.
Learning to plan with logical automata
Araki, B.; Vodrahalli, K.; Leech, T.; Vasile, C.-I.; Donahue, M. D.; and Rus, D. L. 2019 · 2019
Cited alongside, same era.
LTL and Beyond: Formal Languages for Reward Function Specification in Reinforcement Learning
Goal-conditioned reinforcement learning with imagined subgoals
Chane-Sane, E.; Schmid, C.; and Laptev, I. 2021 · 2021
Later among the works it cites.
Compositional reinforcement learning from logical specifications
Jothimurugan, K.; Bansal, S.; Bastani, O.; and Alur, R. 2021 · 2021
Later among the works it cites.
In a Nutshell, the Human Asked for This: Latent Goals for Following Temporal Specifications
León, B. G.; Shanahan, M.; and Belardinelli, F. 2021 · 2021
Later among the works it cites.
Active hierarchical exploration with stable subgoal representation learning
Li, S.; Zhang, J.; Wang, J.; Yu, Y.; and Zhang, C. 2021 · 2021
Later among the works it cites.
Ltl2action: Generalizing ltl instructions for multi-task rl
Vaezipoor, P.; Li, A. C.; Icarte, R. A. T.; and Mcilraith, S. A. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Camacho, A.; Icarte, R. T.; Klassen, T. Q.; Valenzano, R. A.; and McIlraith, S. A. 2019 · 2019
Cited alongside, same era.
A composable specification language for reinforcement learning tasks
Jothimurugan, K.; Alur, R.; and Bastani, O. 2019 · 2019
Cited alongside, same era.
T: A heuristic search based path planning algorithm for temporal logic specifications
Khalidi, D.; Gujarathi, D.; and Saha, I. 2020 · 2020
Cited alongside, same era.
Contrastive Learning of Structured World Models
Kipf, T.; van der Pol, E.; and Welling, M. 2020 · 2020
Cited alongside, same era.
Encoding formulas as deep networks: Reinforcement learning for zero-shot execution of LTL formulas
Kuo, Y.-L.; Katz, B.; and Barbu, A. 2020 · 2020
Cited alongside, same era.
Plannable Approximations to MDP Homomorphisms: Equivariance under Actions
van der Pol, E.; Kipf, T.; Oliehoek, F.; and Welling, M. 2020 · 2020
Cited alongside, same era.
Graph neural networks: A review of methods and applications
Zhou, J.; Cui, G.; Hu, S.; Zhang, Z.; Yang, C.; Liu, Z.; Wang, L.; Li, C.; and Sun, M. 2020 · 2020
Cited alongside, same era.
den Hengst, F.; François-Lavet, V.; Hoogendoorn, M.; and van Harmelen, F. 2022 · 2022
Later among the works it cites.
MT*: Multi-robot path planning for temporal logic specifications
Gujarathi, D.; and Saha, I. 2022 · 2022
Later among the works it cites.
Reward machines: Exploiting reward function structure in reinforcement learning
Icarte, R. T.; Klassen, T. Q.; Valenzano, R.; and McIlraith, S. A. 2022 · 2022
Later among the works it cites.
Skill Transfer for Temporally-Extended Task Specifications
Liu, J. X.; Shah, A.; Rosen, E.; Konidaris, G.; and Tellex, S. 2022 · 2022
Later among the works it cites.
Goal-conditioned reinforcement learning: Problems and solutions
Liu, M.; Zhu, M.; and Zhang, W. 2022 · 2022
Later among the works it cites.
Deep reinforcement learning: A survey
Wang, X.; Wang, S.; Liang, X.; Zhao, D.; Huang, J.; Xu, X.; Dai, B.; and Miao, Q. 2022 · 2022
Later among the works it cites.
Eventual discounting temporal logic counterfactual experience replay
Voloshin, C.; Verma, A.; and Yue, Y. 2023 · 2023
Later among the works it cites.
Generalization of temporal logic tasks via future dependent options
Xu, D.; and Fekri, F. 2024 · 2024
Closest in time.