Fetching the paper…
Reading the bibliography…
Designing reliable decision strategies for autonomous urban driving is challenging.
“Dynamic Programming”
Richard Bellman · 1957
Earlier work this paper cites.
“Learning from delayed rewards”, 1989
Christopher John Cornish Watkins · 1989
Earlier work this paper cites.
“Reinforcement Learning: A Survey”
Leslie Kaelbling, Michael. Littman and Andrew. Moore · 1996
Earlier work this paper cites.
“Congested traffic states in empirical observations and microscopic simulations”
Martin Treiber, Ansgar Hennecke and Dirk Helbing · 2000
Earlier work this paper cites.
“Principles of model checking”
Christel Baier and Joost-Pieter Katoen · 2008
Earlier work this paper cites.
“Motion planning and control from temporal logic specifications with probabilistic satisfaction guarantees”
Morteza Lahijanian, Joseph Wasniewski, Sean Andersson and Calin Belta · 2010
Earlier work this paper cites.
“Control of Markov decision processes from PCTL specifications”
M. Lahijanian, S.. Andersson and C. Belta · 2011
Earlier work this paper cites.
“Intention-Aware Motion Planning”
Tirthankar Bandyopadhyay, Kok Won, Emilio Frazzoli, David Hsu, Wee Lee and Daniela Rus · 2012
Cited alongside, same era.
“A comprehensive survey on safe reinforcement learning”
Javier Garc“’a and Fernando Fern“’andez · 2015
Cited alongside, same era.
“Decision Making Under Uncertainty: Theory and Application”
Mykel Kochenderfer · 2015
Cited alongside, same era.
“Human-level control through deep reinforcement learning”
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei. Rusu, Joel Veness, Marc. Bellemare, Alex Graves, Martin. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg and Demis Hassabis · 2015
Cited alongside, same era.
“Prioritized experience replay”
Tom Schaul, John Quan, Ioannis Antonoglou and David Silver · 2016
Cited alongside, same era.
“Belief state planning for autonomously navigating urban intersections”
“Tactical Decision Making for Lane Changing with Deep Reinforcement Learning”
Mustafa Mukadam, Akansel Cosgun, Alireza Nakhaei and Kikuo Fujimura · 2017
Later among the works it cites.
“The value of inferring the internal state of traffic participants for autonomous freeway driving”
Zachary. Sunberg, Christopher. Ho and Mykel. Kochenderfer · 2017
Later among the works it cites.
“Safe Reinforcement Learning via Shielding”
Mohammed Alshiekh, Roderick Bloem, R“”udiger Ehlers, Bettina K“”onighofer, Scott Niekum and Ufuk Topcu · 2018
Later among the works it cites.
“Utility Decomposition with Deep Corrections for Scalable Planning under Uncertainty”
Maxime Bouton, Kyle Julian, Alireza Nakhaei, Kikuo Fujimura and Mykel. Kochenderfer · 2018
Later among the works it cites.
“Safe Reinforcement Learning via Formal Methods”
Nathan Fulton and André Platzer · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Maxime Bouton, Akansel Cosgun and Mykel. Kochenderfer · 2017
Cited alongside, same era.
“Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks”
Guy Katz, Clark. Barrett, David. Dill, Kyle Julian and Mykel. Kochenderfer · 2017
Cited alongside, same era.
Nils Jansen, Bettina K“”onighofer, Sebastian Junges and Roderick Bloem · 2018
Later among the works it cites.
“Bounded Policy Synthesis for POMDPs with Safe-Reachability Objectives”
Yue Wang, Swarat Chaudhuri and Lydia. Kavraki · 2018
Later among the works it cites.