Fetching the paper…
Reading the bibliography…
This work introduces a neuro-symbolic agent that combines deep reinforcement learning (DRL) with temporal logic (TL) to achieve systematic zero-shot, i.e., never-seen-before, generalisation of formally specified instructions.
Modular Deep Reinforcement Learning with Temporal Logic Specifications
Yuan, L. Z.; Hasanbeig, M.; Abate, A.; and Kroening, D. 2019 · 1909
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V.; Badia, A. P.; Mirza, M.; Graves, A.; Lillicrap, T.; Harley, T.; Silver, D.; and Kavukcuoglu, K. 2016 · 1937
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Fodor, J. A.; and Pylyshyn, Z. W. 1988 · 1988
Earlier work this paper cites.
On the proper treatment of connectionism
Smolensky, P. 1988 · 1988
Earlier work this paper cites.
Rewarding behaviors
Bacchus, F.; Boutilier, C.; and Grove, A. 1996 · 1996
Earlier work this paper cites.
Learning to forget: Continual prediction with LSTM
Gers, F. A.; Schmidhuber, J.; and Cummins, F. 2000 · 2000
Earlier work this paper cites.
Syntactic structures
Chomsky, N.; and Lightfoot, D. W. 2002 · 2002
Earlier work this paper cites.
Logic in Computer Science: Modelling and reasoning about systems
Huth, M.; and Ryan, M. 2004 · 2004
Earlier work this paper cites.
Encoding formulas as deep networks: Reinforcement learning for zero-shot execution of LTL formulas
Kuo, Y.; Katz, B.; and Barbu, A. 2020 · 2006
Earlier work this paper cites.
Principles of model checking
Baier, C.; and Katoen, J.-P. 2008 · 2008
Earlier work this paper cites.
Letting structure emerge: connectionist and dynamical systems approaches to cognition
McClelland, J. L.; Botvinick, M. M.; Noelle, D. C.; Plaut, D. C.; Rogers, T. T.; Seidenberg, M. S.; and Smith, L. B. 2010 · 2010
Earlier work this paper cites.
Linear temporal logic and linear dynamic logic on finite traces
De Giacomo, G.; and Vardi, M. Y. 2013 · 2013
Earlier work this paper cites.
Deep learning
LeCun, Y.; Bengio, Y.; and Hinton, G. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Earlier work this paper cites.
Deep learning , volume 1
Goodfellow, I.; Bengio, Y.; Courville, A.; and Bengio, Y. 2016 · 2016
Earlier work this paper cites.
Modular multitask reinforcement learning with policy sketches
Andreas, J.; Klein, D.; and Levine, S. 2017 · 2017
Cited alongside, same era.
Environment-independent task specifications via gltl
Littman, M. L.; Topcu, U.; Fu, J.; Isbell, C.; Wen, M.; and MacGlashan, J. 2017 · 2017
Cited alongside, same era.
Zero-shot task generalization with multi-task deep reinforcement learning
Oh, J.; Singh, S.; Lee, H.; and Kohli, P. 2017 · 2017
Cited alongside, same era.
Safe reinforcement learning via shielding
Alshiekh, M.; Bloem, R.; Ehlers, R.; Könighofer, B.; Niekum, S.; and Topcu, U. 2018 · 2018
Cited alongside, same era.
LTLf/LDLf non-markovian rewards
Brafman, R. I.; De Giacomo, G.; and Patrizi, F. 2018 · 2018
Cited alongside, same era.
Minimalistic Gridworld Environment for OpenAI Gym
Chevalier-Boisvert, M.; Willems, L.; and Pal, S. 2018 · 2018
Foundations for restraining bolts: Reinforcement learning with LTLf/LDLf restraining specifications
De Giacomo, G.; Iocchi, L.; Favorito, M.; and Patrizi, F. 2019 · 2019
Later among the works it cites.
Neural architecture search: A survey
Elsken, T.; Metzen, J. H.; Hutter, F.; et al. 2019 · 2019
Later among the works it cites.
Learning Reward Machines for Partially Observable Reinforcement Learning
Icarte, R. T.; Waldie, E.; Klassen, T.; Valenzano, R.; Castro, M.; and McIlraith, S. 2019 · 2019
Later among the works it cites.
A Composable Specification Language for Reinforcement Learning Tasks
Jothimurugan, K.; Alur, R.; and Bastani, O. 2019 · 2019
Later among the works it cites.
Compositional generalization through meta sequence-to-sequence learning
Lake, B. M. 2019 · 2019
Later among the works it cites.
A Survey of Reinforcement Learning Informed by Natural Language
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures
Espeholt, L.; Soyer, H.; Munos, R.; Simonyan, K.; Mnih, V.; Ward, T.; Doron, Y.; Firoiu, V.; Harley, T.; Dunning, I.; Legg, S.; and Kavukcuoglu, K. 2018 · 2018
Cited alongside, same era.
Using reward machines for high-level task specification and decomposition in reinforcement learning
Icarte, R. T.; Klassen, T.; Valenzano, R.; and McIlraith, S. 2018 · 2018
Cited alongside, same era.
Generalization without Systematicity: On the Compositional Skills of Sequence-to-Sequence Recurrent Networks
Lake, B.; and Baroni, M. 2018 · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Cited alongside, same era.
Teaching multiple tasks to an RL agent using LTL
Toro Icarte, R.; Klassen, T. Q.; Valenzano, R.; and McIlraith, S. A. 2018 · 2018
Cited alongside, same era.
Interactive Grounded Language Acquisition and Generalization in a 2D World
Yu, H.; Zhang, H.; and Xu, W. 2018 · 2018
Cited alongside, same era.
Luketina, J.; Nardelli, N.; Farquhar, G.; Foerster, J. N.; Andreas, J.; Grefenstette, E.; Whiteson, S.; and Rocktäschel, T. 2019 · 2019
Later among the works it cites.
The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Mao, J.; Gan, C.; Kohli, P.; Tenenbaum, J. B.; and Wu, J. 2019 · 2019
Later among the works it cites.
Environmental drivers of systematicity and generalization in a situated agent
Hill, F.; Lampinen, A.; Schneider, R.; Clark, S.; Botvinick, M.; McClelland, J. L.; and Santoro, A. 2020 · 2020
Closest in time.
Extended Markov Games to Learn Multiple Tasks in Multi-Agent Reinforcement Learning
León, B. G.; and Belardinelli, F. 2020 · 2020
Closest in time.
Multi-Agent Reinforcement Learning with Temporal Logic Specifications
Hammond, L.; Abate, A.; Gutierrez, J.; and Wooldridge, M. 2021 · 2021
Closest in time.
DeepSynth: Automata Synthesis for Automatic Task Segmentation in Deep Reinforcement Learning
Hasanbeig, M.; Jeppu, N. Y.; Abate, A.; Melham, T.; and Kroening, D. 2021 · 2021
Closest in time.
Grounded Language Learning Fast and Slow
Hill, F.; Tieleman, O.; von Glehn, T.; Wong, N.; Merzic, H.; and Clark, S. 2021 · 2021
Closest in time.
AlwaysSafe: Reinforcement learning without safety constraint violations during training
Simão, T. D.; Jansen, N.; and Spaan, M. T. 2021 · 2021
Closest in time.
Rapidly Adaptable Legged Robots via Evolutionary Meta-Learning
Song, X.; Yang, Y.; Choromanski, K.; Caluwaerts, K.; Gao, W.; Finn, C.; and Tan, J. 2020 · 2021
Closest in time.