Fetching the paper…
Reading the bibliography…
Natural and formal languages provide an effective mechanism for humans to specify instructions and reward functions.
A Markovian Decision Process
Richard Bellman · 1957
Earlier work this paper cites.
Introduction to Automata Theory, Languages, and Computation, 3rd Edition
John E. Hopcroft, Rajeev Motwani, and Jeffrey D. Ullman · 2007
Earlier work this paper cites.
Q-learning with linear function approximation
Francisco S Melo and M Isabel Ribeiro · 2007
Earlier work this paper cites.
LTL Control in Uncertain Environments with Probabilistic Satisfaction Guarantees
Xu Chu Dennis Ding, Stephen L Smith, Calin Belta, and Daniela Rus · 2011
Earlier work this paper cites.
Distribution Temporal Logic: Combining Correctness with Quality of Estimation
Austin Jones, Mac Schwager, and Calin Belta · 2013
Earlier work this paper cites.
Q-learning for Robust Satisfaction of Signal Temporal Logic Specifications
Derya Aksaray, Austin Jones, Zhaodan Kong, Mac Schwager, and Calin Belta · 2016
Earlier work this paper cites.
Safe Control under Uncertainty with Probabilistic Signal Temporal Logic
Dorsa Sadigh and Ashish Kapoor · 2016
Earlier work this paper cites.
Modular Multitask Reinforcement Learning with Policy Sketches
Jacob Andreas, Dan Klein, and Sergey Levine · 2017
Earlier work this paper cites.
Reinforcement Learning with Temporal Logic Rewards
Xiao Li, Cristian Ioan Vasile, and Calin Belta · 2017
Earlier work this paper cites.
Environment-Independent Task Specifications via GLTL
Michael L Littman, Ufuk Topcu, Jie Fu, Charles Isbell, Min Wen, and James MacGlashan · 2017
Earlier work this paper cites.
Zero-shot Task Generalization with Multi-Task Deep Reinforcement Learning
Junhyuk Oh, Satinder Singh, Honglak Lee, and Pushmeet Kohli · 2017
Earlier work this paper cites.
Proximal Policy Optimization Algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Logically-Constrained Reinforcement Learning
Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2018
Earlier work this paper cites.
A Policy Search Method for Temporal Logic Specified Reinforcement Learning Tasks
Xiao Li, Yao Ma, and Calin Belta · 2018
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto · 2018
Earlier work this paper cites.
Using Reward Machines for High-Level Task Specification and Decomposition in Reinforcement Learning
Rodrigo Toro Icarte, Toryn Q. Klassen, Richard Valenzano, and Sheila A. McIlraith · 2018
Earlier work this paper cites.
LTL and Beyond: Formal Languages for Reward Function Specification in Reinforcement Learning
Alberto Camacho, Rodrigo Toro Icarte, Toryn Q. Klassen, Richard Valenzano, and Sheila A. McIlraith · 2019
Earlier work this paper cites.
Active Perception and Control from Temporal Logic Specifications
Rafael Rodrigues da Silva, Vince Kurtz, and Hai Lin · 2019
Cited alongside, same era.
Language as an Abstraction for Hierarchical Deep Reinforcement Learning
Yiding Jiang, Shixiang Shane Gu, Kevin P Murphy, and Chelsea Finn · 2019
Cited alongside, same era.
A Composable Specification Language for Reinforcement Learning Tasks
Kishor Jothimurugan, Rajeev Alur, and Osbert Bastani · 2019
Cited alongside, same era.
A Survey of Reinforcement Learning Informed by Natural Language
Jelena Luketina, Nantas Nardelli, Gregory Farquhar, Jakob Foerster, Jacob Andreas, Edward Grefenstette, Shimon Whiteson, and Tim Rocktäschel · 2019
Cited alongside, same era.
Learning Reward Machines for Partially Observable Reinforcement Learning
Rodrigo Toro Icarte, Ethan Waldie, Toryn Q. Klassen, Rick Valenzano, Margarita P. Castro, and Sheila A. McIlraith · 2019
Cited alongside, same era.
Reinforcement Learning Based Temporal Logic Control with Maximum Probabilistic Satisfaction
Mingyu Cai, Shaoping Xiao, Baoluo Li, Zhiliang Li, and Zhen Kan · 2021
Later among the works it cites.
Reward Machines for Vision-Based Robotic Manipulation
Alberto Camacho, Jacob Varley, Andy Zeng, Deepali Jain, Atil Iscen, and Dmitry Kalashnikov · 2021
Later among the works it cites.
Learning Quadruped Locomotion Policies with Reward Machines
David DeFazio and Shiqi Zhang · 2021
Later among the works it cites.
Temporal-Logic-Based Reward Shaping for Continuing Reinforcement Learning Tasks
Yuqian Jiang, Suda Bharadwaj, Bo Wu, Rishi Shah, Ufuk Topcu, and Peter Stone · 2021
Later among the works it cites.
Compositional Reinforcement Learning from Logical Specifications
Kishor Jothimurugan, Suguman Bansal, Osbert Bastani, and Rajeev Alur · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhe Xu and Ufuk Topcu · 2019
Cited alongside, same era.
Modular Deep Reinforcement Learning with Temporal Logic Specifications
Lim Zun Yuan, Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2019
Cited alongside, same era.
Temporal Logic Monitoring Rewards via Transducers
Giuseppe de Giacomo, Marco Favorito, Luca Iocchi, Fabio Patrizi, and Alessandro Ronca · 2020
Cited alongside, same era.
Induction of Subgoal Automata for Reinforcement Learning
Daniel Furelos-Blanco, Mark Law, Alessandra Russo, Krysia Broda, and Anders Jonsson · 2020
Cited alongside, same era.
Reinforcement Learning with non-Markovian Rewards
Maor Gaon and Ronen Brafman · 2020
Cited alongside, same era.
Task-Oriented Active Perception and Planning in Environments with Partially Known Semantics
Mahsa Ghasemi, Erdem Arinc Bulgur, and Ufuk Topcu · 2020
Cited alongside, same era.
Deep Reinforcement Learning with Temporal Logics
Mohammadhosein Hasanbeig, Daniel Kroening, and Alessandro Abate · 2020
Cited alongside, same era.
Reward Machines for Cooperative Multi-Agent Reinforcement Learning
Cyrus Neary, Zhe Xu, Bo Wu, and Ufuk Topcu · 2021
Later among the works it cites.
LTL2Action: Generalizing LTL Instructions for Multi-Task RL
Pashootan Vaezipoor, Andrew Li, Rodrigo Toro Icarte, and Sheila McIlraith · 2021
Later among the works it cites.
Learning Probabilistic Reward Machines from Non-Markovian Stochastic Reward Processes
Alvaro Velasquez, Andre Beckus, Taylor Dohmen, Ashutosh Trivedi, Noah Topper, and George Atia · 2021
Later among the works it cites.
Active Finite Reward Automaton Inference and Reinforcement Learning using Queries and Counterexamples
Zhe Xu, Bo Wu, Aditya Ojha, Daniel Neider, and Ufuk Topcu · 2021
Later among the works it cites.
Reinforcement Learning with Stochastic Reward Machines
Jan Corazza, Ivan Gavran, and Daniel Neider · 2022
Closest in time.
Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Wenlong Huang, Pieter Abbeel, Deepak Pathak, and Igor Mordatch · 2022
Closest in time.
In a Nutshell, the Human Asked for This: Latent Goals for Following Temporal Specifications
Borja G León, Murray Shanahan, and Francesco Belardinelli · 2022
Closest in time.
Skill Transfer for Temporally-Extended Task Specifications
Jason Xinyu Liu, Ankit Shah, Eric Rosen, George Konidaris, and Stefanie Tellex · 2022
Closest in time.
Reward Machines: Exploiting Reward Function Structure in Reinforcement Learning
Rodrigo Toro Icarte, Toryn Q. Klassen, Richard Valenzano, and Sheila A. McIlraith · 2022
Closest in time.
Learning to Follow Instructions in Text-Based Games
Mathieu Tuli, Andrew C. Li, Pashootan Vaezipoor, Toryn Q. Klassen, Scott Sanner, and Sheila A. McIlraith · 2022
Closest in time.
Joint Learning of Reward Machines and Policies in Environments with Partially Known Semantics
Christos Verginis, Cevahir Koprulu, Sandeep Chinchali, and Ufuk Topcu · 2022
Closest in time.
Lifelong Reinforcement Learning with Temporal Logic Formulas and Reward Machines
Xuejing Zheng, Chao Yu, and Minjie Zhang · 2022
Closest in time.