Fetching the paper…
Reading the bibliography…
Reinforcement Learning (RL) has emerged as an efficient method of choice for solving complex sequential decision making problems in automatic control, computer science, economics, and biology.
A. Pnueli, “The temporal logic of programs,” in Foundations of Computer Science . IEEE, 1977, pp. 46–57
1977
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, 1998, vol. 1
1998
Earlier work this paper cites.
R. Durrett, Essentials of stochastic processes . Springer, 1999, vol. 1
1999
Earlier work this paper cites.
R. Alur and S. La Torre, “Deterministic generators and games for LTL fragments,” TOCL , vol. 5, no. 1, pp. 1–25, 2004
2004
Earlier work this paper cites.
G. E. Fainekos, H. Kress-Gazit, and G. J. Pappas, “Hybrid controllers for path planning: A temporal logic approach,” in CDC and ECC , December 2005, pp. 4885–4890
2005
Earlier work this paper cites.
C. Baier and J.-P. Katoen, Principles of model checking . MIT Press, 2008
2008
Earlier work this paper cites.
A. Platzer, “Differential dynamic logic for hybrid systems,” Journal of Automated Reasoning , vol. 41, no. 2, pp. 143–189, 2008
2008
Earlier work this paper cites.
A. Abate, M. Prandini, J. Lygeros, and S. Sastry, “Probabilistic reachability and safety for controlled discrete time stochastic hybrid systems,” Automatica , vol. 44, no. 11, pp. 2724–2734, 2008
2008
Earlier work this paper cites.
H. Kress-Gazit, G. E. Fainekos, and G. J. Pappas, “Temporal-logic-based reactive mission and motion planning,” IEEE Transactions on Robotics , vol. 25, no. 6, pp. 1370–1381, 2009
2009
Earlier work this paper cites.
2009
Earlier work this paper cites.
A. Bhatia, L. E. Kavraki, and M. Y. Vardi, “Sampling-based motion planning with temporal goals,” in ICRA , 2010, pp. 2689–2696
2010
Earlier work this paper cites.
A. Abate, J.-P. Katoen, J. Lygeros, and M. Prandini, “Approximate model checking of stochastic hybrid systems,” European Journal of Control , vol. 16, no. 6, pp. 624–641, 2010
2010
Earlier work this paper cites.
X. C. Ding, S. L. Smith, C. Belta, and D. Rus, “MDP optimal control under temporal logic constraints,” in CDC and ECC , 2011, pp. 532–538
2011
Earlier work this paper cites.
M. Kwiatkowska, G. Norman, and D. Parker, “PRISM 4.0: Verification of probabilistic real-time systems,” in CAV . Springer, 2011, pp. 585–591
2011
Cited alongside, same era.
E. M. Wolff, U. Topcu, and R. M. Murray, “Robust control of uncertain Markov decision processes with temporal logic specifications,” in CDC , 2012, pp. 3372–3379
2012
Cited alongside, same era.
V. Forejt, M. Kwiatkowska, and D. Parker, “Pareto curves for probabilistic model checking,” in ATVA . Springer, 2012, pp. 317–332
2012
Cited alongside, same era.
X. Ding, S. L. Smith, C. Belta, and D. Rus, “Optimal control of Markov decision processes with linear temporal logic constraints,” IEEE Transactions on Automatic Control , vol. 59, no. 5, pp. 1244–1257, 2014
2014
Cited alongside, same era.
J. Fu and U. Topcu, “Probably approximately correct MDP learning and control with temporal logic constraints,” in Robotics: Science and Systems X , 2014
S. Sickert and J. Křetínskỳ, “MoChiBA: Probabilistic LTL model checking using limit-deterministic Büchi automata,” in ATVA . Springer, 2016, pp. 130–137
2016
Later among the works it cites.
——, “Distributed intermittent connectivity control of mobile robot networks,” IEEE Transactions on Automatic Control , vol. 62, no. 7, pp. 3109–3121, 2017
2017
Later among the works it cites.
I. Tkachev, A. Mereacre, J.-P. Katoen, and A. Abate, “Quantitative model-checking of controlled discrete-time Markov processes,” Information and Computation , vol. 253, pp. 1–35, 2017
2017
Later among the works it cites.
Y. Kantaros and M. M. Zavlanos, “Sampling-based optimal control synthesis for multi-robot systems under global temporal tasks,” IEEE Transactions on Automatic Control , 2018. [Online]. Available: DOI:10.1109/TAC.2018.2853558
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
T. Brázdil, K. Chatterjee, M. Chmelík, V. Forejt, J. Křetínskỳ, M. Kwiatkowska, D. Parker, and M. Ujma, “Verification of Markov decision processes using learning algorithms,” in ATVA . Springer, 2014, pp. 98–114
2014
Cited alongside, same era.
D. Sadigh, E. S. Kim, S. Coogan, S. S. Sastry, and S. A. Seshia, “A learning based approach to control synthesis of Markov decision processes for linear temporal logic specifications,” in CDC . IEEE, 2014, pp. 1091–1096
2014
Cited alongside, same era.
——, “Logically-constrained neural fitted Q-iteration,” in AAMAS , 2019, pp. 2012–2014
2014
Cited alongside, same era.
M. L. Puterman, Markov decision processes: Discrete stochastic dynamic programming . John Wiley & Sons, 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
J. Wang, X. Ding, M. Lahijanian, I. C. Paschalidis, and C. A. Belta, “Temporal logic motion control using actor–critic methods,” The International Journal of Robotics Research , vol. 34, no. 10, pp. 1329–1344, 2015
2015
Cited alongside, same era.
S. Sickert, J. Esparza, S. Jaax, and J. Křetínskỳ, “Limit-deterministic Büchi automata for linear temporal logic,” in CAV . Springer, 2016, pp. 312–332
2016
Cited alongside, same era.
M. Guo and M. M. Zavlanos, “Probabilistic motion planning under temporal tasks and soft constraints,” IEEE Transactions on Automatic Control , 2018
2018
Later among the works it cites.
E. M. Clarke, O. Grumberg, D. Kroening, D. Peled, and H. Veith, Model Checking , 2nd ed. MIT Press, 2018
2018
Later among the works it cites.
N. Fulton, “Verifiably safe autonomy for cyber-physical systems,” Ph.D. dissertation, Carnegie Mellon University Pittsburgh, PA, 2018
2018
Later among the works it cites.
N. Fulton and A. Platzer, “Safe reinforcement learning via formal methods: Toward safe control through proof and learning,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Q. Gao, D. Hajinezhad, Y. Zhang, Y. Kantaros, and M. M. Zavlanos, “Reduced variance deep reinforcement learning with temporal logic specifications,” 2019 (to appear)
2019
Closest in time.