Fetching the paper…
Reading the bibliography…
In this paper, we propose a model-free reinforcement learning method to synthesize control policies for motion planning problems with continuous states and actions.
Learning from Delayed Rewards
Christopher John Cornish Hellaby Watkins · 1989
Earlier work this paper cites.
The complexity of probabilistic verification
Costas Courcoubetis and Mihalis Yannakakis · 1995
Earlier work this paper cites.
Introduction to Reinforcement Learning
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David A McAllester, Satinder P Singh, and Yishay Mansour · 2000
Earlier work this paper cites.
Probabilistic Robotics
Wolfram Burgard Sebastian Thrun and Dieter Fox · 2002
Earlier work this paper cites.
Automatic synthesis of multi-agent motion tasks based on ltl specifications
Savvas G Loizou and Kostas J Kyriakopoulos · 2004
Earlier work this paper cites.
Modular deep reinforcement learning with temporal logic specifications
Lim Zun Yuan, Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2004
Earlier work this paper cites.
Natural actor-critic
Jan Peters, Sethu Vijayakumar, and Stefan Schaal · 2005
Earlier work this paper cites.
Principles of Model Checking
Christel Baier and Joost-Pieter Katoen · 2008
Earlier work this paper cites.
Temporal logic motion planning for dynamic robots
Georgios E Fainekos, Antoine Girard, Hadas Kress-Gazit, and George J Pappas · 2009
Earlier work this paper cites.
Receding horizon temporal logic planning for dynamical systems
Tichakorn Wongpiromsarn, Ufuk Topcu, and Richard M Murray · 2009
Earlier work this paper cites.
Sampling-based motion planning with temporal goals
Amit Bhatia, Lydia E Kavraki, and Moshe Y Vardi · 2010
Cited alongside, same era.
Sampling-based algorithms for optimal motion planning
Sertac Karaman and Emilio Frazzoli · 2011
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, and et al · 2013
Cited alongside, same era.
A learning based approach to control synthesis of markov decision processes for linear temporal logic specifications
Dorsa Sadigh, Eric S Kim, Samuel Coogan, S Shankar Sastry, and Sanjit A Seshia · 2014
Cited alongside, same era.
Deterministic policy gradient algorithms
David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller · 2014
Cited alongside, same era.
Logically-constrained neural fitted q-iteration
Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2018
Later among the works it cites.
A policy search method for temporal logic specified reinforcement learning tasks
Xiao Li, Yao Ma, and Calin Belta · 2018
Later among the works it cites.
Robustly complete synthesis of memoryless controllers for nonlinear systems with reach-and-stay specifications
Yinan Li and Jun Liu · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Reduced variance deep reinforcement learning with temporal logic specifications
Qitong Gao, Davood Hajinezhad, Yan Zhang, Yiannis Kantaros, and Michael M Zavlanos · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, et al · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, et al · 2015
Cited alongside, same era.
Benchmarking deep reinforcement learning for continuous control
Yan Duan, Xi Chen, Rein Houthooft, John Schulman, and Pieter Abbeel · 2016
Cited alongside, same era.
Limit-Deterministic Büchi Automata for Linear Temporal Logic
Salomon Sickert, Javier Esparza, Stefan Jaax, and Jan Křetínský · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, et al · 2016
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi, Ashutosh Trivedi, and Dominik Wojtczak · 2019
Later among the works it cites.
Certified reinforcement learning with logic guidance
Mohammadhosein Hasanbeig, Alessandro Abate, and Daniel Kroening · 2019
Later among the works it cites.
Mohammadhosein Hasanbeig, Yiannis Kantaros, Alessandro Abate, Daniel Kroening, George J Pappas, and Insup Lee · 2019
Later among the works it cites.
Formal controller synthesis for continuous-space mdps via model-free reinforcement learning
Abolfazl Lavaei, Fabio Somenzi, Sadegh Soudjani, Ashutosh Trivedi, and Majid Zamani · 2020
Closest in time.
Ryohei Oura, Ami Sakakibara, and Toshimitsu Ushio · 2020
Closest in time.