Fetching the paper…
Reading the bibliography…
This paper addresses the problem of designing control policies for agents with unknown stochastic dynamics and control objectives specified using Linear Temporal Logic (LTL).
Principles of model checking , volume 26202649
Christel Baier and Joost-Pieter Katoen · 2008
Earlier work this paper cites.
Playing Atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Verification of Markov decision processes using learning algorithms
Tomáš Brázdil, Krishnendu Chatterjee, Martin Chmelik, Vojtěch Forejt, Jan Křetínskỳ, Marta Kwiatkowska, David Parker, and Mateusz Ujma · 2014
Earlier work this paper cites.
Reinforcement learning and the reward engineering principle
Daniel Dewey · 2014
Earlier work this paper cites.
Probably approximately correct MDP learning and control with temporal logic constraints
Jie Fu and Ufuk Topcu · 2014
Earlier work this paper cites.
Distributed intermittent connectivity control of mobile robot networks
Yiannis Kantaros and Michael M Zavlanos · 2016
Earlier work this paper cites.
Persistent surveillance for unmanned aerial vehicles subject to charging and temporal logic constraints
Kevin Leahy, Dingjiang Zhou, Cristian-Ioan Vasile, Konstantinos Oikonomopoulos, Mac Schwager, and Calin Belta · 2016
Earlier work this paper cites.
Socially aware motion planning with deep reinforcement learning
Yu Fan Chen, Michael Everett, Miao Liu, and Jonathan P How · 2017
Earlier work this paper cites.
Distributed data gathering with buffer constraints and intermittent communication
Meng Guo and Michael M. Zavlanos · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
PRM-RL: Long-range robotic navigation tasks by combining reinforcement learning and sampling-based planning
Aleksandra Faust, Kenneth Oslund, Oscar Ramirez, Anthony Francis, Lydia Tapia, Marek Fiser, and James Davidson · 2018
Earlier work this paper cites.
Probabilistic motion planning under temporal tasks and soft constraints
Meng Guo and Michael M Zavlanos · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Earlier work this paper cites.
Anytime planning for decentralized multirobot active information gathering
Brent Schlotfeldt, Dinesh Thakur, Nikolay Atanasov, Vijay Kumar, and George J Pappas · 2018
Earlier work this paper cites.
Reduced variance deep reinforcement learning with temporal logic specifications
Qitong Gao, Davood Hajinezhad, Yan Zhang, Yiannis Kantaros, and Michael M Zavlanos · 2019
Cited alongside, same era.
Reinforcement learning for temporal logic control synthesis with probabilistic satisfaction guarantees
Hosein Hasanbeig, Yiannis Kantaros, Alessandro Abate, Daniel Kroening, George J. Pappas, and Insup Lee · 2019
Cited alongside, same era.
Transfer of temporal logic formulas in reinforcement learning
Zhe Xu and Ufuk Topcu · 2019
Cited alongside, same era.
Control synthesis from linear temporal logic specifications using model-free reinforcement learning
Alper Kamil Bozkurt, Yu Wang, Michael M Zavlanos, and Miroslav Pajic · 2020
Cited alongside, same era.
Deep reinforcement learning with temporal logics
Hosein Hasanbeig, Daniel Kroening, and Alessandro Abate · 2020
Cited alongside, same era.
Model-free reinforcement learning for symbolic automata-encoded objectives
Anand Balakrishnan, Stefan Jaksic, Edgar Aguilar, Dejan Nickovic, and Jyotirmoy Deshmukh · 2022
Later among the works it cites.
Specification-guided reinforcement learning
Suguman Bansal · 2022
Later among the works it cites.
Learning minimally-violating continuous control for infeasible linear temporal logic specifications
Mingyu Cai, Makai Mann, Zachary Serlin, Kevin Leahy, and Cristian-Ioan Vasile · 2022
Later among the works it cites.
LCRL: Certified policy synthesis via logically-constrained reinforcement learning
Hosein Hasanbeig, Daniel Kroening, and Alessandro Abate · 2022
Later among the works it cites.
Reward machines: Exploiting reward function structure in reinforcement learning
Rodrigo Toro Icarte, Toryn Q Klassen, Richard Valenzano, and Sheila A McIlraith · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yiannis Kantaros and Michael M Zavlanos · 2020
Cited alongside, same era.
Formal controller synthesis for continuous-space MDPs via model-free reinforcement learning
Abolfazl Lavaei, Fabio Somenzi, Sadegh Soudjani, Ashutosh Trivedi, and Majid Zamani · 2020
Cited alongside, same era.
Continuous motion planning with temporal logic specifications using deep neural networks
Chuanzheng Wang, Yinan Li, Stephen L Smith, and Jun Liu · 2020
Cited alongside, same era.
Reinforcement learning based temporal logic control with maximum probabilistic satisfaction
Mingyu Cai, Shaoping Xiao, Baoluo Li, Zhiliang Li, and Zhen Kan · 2021
Cited alongside, same era.
Model-based reinforcement learning for approximate optimal control with temporal logic specifications
Max H Cohen and Calin Belta · 2021
Cited alongside, same era.
DeepSynth: Program synthesis for automatic task segmentation in deep reinforcement learning
Hosein Hasanbeig, Natasha Yogananda Jeppu, Alessandro Abate, Tom Melham, and Daniel Kroening · 2021
Cited alongside, same era.
Temporal-logic-based reward shaping for continuing reinforcement learning tasks
Yuqian Jiang, Suda Bharadwaj, Bo Wu, Rishi Shah, Ufuk Topcu, and Peter Stone · 2021
Cited alongside, same era.
Accelerated reinforcement learning for temporal logic control objectives
Yiannis Kantaros · 2022
Later among the works it cites.
Perception-based temporal logic planning in uncertain semantic maps
Yiannis Kantaros, Samarth Kalluraya, Qi Jin, and George J Pappas · 2022
Later among the works it cites.
Computational benefits of intermediate rewards for goal-reaching policy learning
Yuexiang Zhai, Christina Baek, Zhengyuan Zhou, Jiantao Jiao, and Yi Ma · 2022
Later among the works it cites.
Programmatic reward design by example
Weichao Zhou and Wenchao Li · 2022
Later among the works it cites.
Overcoming exploration: Deep reinforcement learning for continuous control in cluttered environments from temporal logic specifications
Mingyu Cai, Erfan Aasi, Calin Belta, and Cristian-Ioan Vasile · 2023
Closest in time.
Sample efficient model-free reinforcement learning from LTL specifications with optimality guarantees
Daqian Shao and Marta Kwiatkowska · 2023
Closest in time.
Exploiting transformer in sparse reward reinforcement learning for interpretable temporal logic motion planning
Hao Zhang, Hao Wang, and Zhen Kan · 2023
Closest in time.
Identify, estimate and bound the uncertainty of reinforcement learning for autonomous driving
Weitao Zhou, Zhong Cao, Nanshan Deng, Kun Jiang, and Diange Yang · 2023
Closest in time.
Sample-efficient reinforcement learning with temporal logic objectives: Leveraging the task specification to guide exploration
Yiannis Kantaros and Jun Wang · 2024
Closest in time.
Uncertainty and noise aware decision making for autonomous vehicles-a bayesian approach
Rewat Sachdeva, Raghav Gakhar, Sharad Awasthi, Kavinder Singh, Ashutosh Pandey, and Anil Singh Parihar · 2024
Closest in time.