Fetching the paper…
Reading the bibliography…
Reinforcement learning using a novel predictive representation is applied to autonomous driving to accomplish the task of driving between lane markings where substantial benefits in performance and generalization are observed on unseen test roads in both simulation and on a real Jackal robot.
Importance resampling off-policy prediction
Matthew Schlegel, Wesley Chung, Daniel Graves Jian Qian, and Martha White · 1906
Earlier work this paper cites.
On the theory of the brownian motion
George Uhlenbeck and Leonard Ornstein · 1930
Earlier work this paper cites.
Predictive representations of state
Michael L. Littman and Richard S Sutton · 1983
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Richard S. Sutton · 1988
Earlier work this paper cites.
Probabilistic robotics
Sebastian Thrun · 2002
Earlier work this paper cites.
A two-point visual control model of steering
Dario D Salvucci and Rob Gray · 2004
Earlier work this paper cites.
Offline reinforcement learning: tutorial, review and perspectives on open problems
Sergey Levine, Aviral Kumar, George Tucker, and Justin Fu · 2005
Earlier work this paper cites.
Using predictive representations to improve generalization in reinforcement learning
Eddie J. Rafols, Mark B. Ring, Richard S. Sutton, and Brian Tanner · 2005
Earlier work this paper cites.
Gps navigation based autonomous driving system design for intelligent vehicles
Bing-Fei Wu, Tsu-Tian Lee, Hsin-Han Chang, Jhong-Jie Jiang, Cheng-Nan Lien, Tien-Yu Liao, and Jau-Woei Perng · 2007
Earlier work this paper cites.
A flexible and scalable slam system with full 3d motion estimation
S. Kohlbrecher, J. Meyer, O. von Stryk, and U. Klingauf · 2011
Earlier work this paper cites.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
Richard Sutton, Joseph Modayil, Michael Delp, Thomas Degris, Patrick Pilarski, Adam White, and Doina Precup · 2011
Earlier work this paper cites.
Multi-timescale nexting in a reinforcement learning robot
Joseph Modayil, Adam White, and Richard S. Sutton · 2012
Earlier work this paper cites.
Whatever next? predictive brains, situated agents, and the future of cognitive science
A. Clark · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Better generalization with forecasts
Tom Schaul and Mark Ring · 2013
Cited alongside, same era.
Torcs, the open racing car simulator, v1.3.5, 2013
Bernhard Wymann, Eric Espie, Christophe Guionneau, Christos Dimitrakakis, Remi Coulom, and Andrew Sumner · 2013
Cited alongside, same era.
Deterministic policy gradient algorithms
David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller · 2014
Cited alongside, same era.
A survey of motion planning and control techniques for self-driving urban vehicles
Brian Paden, Michal Cáp, Sze Zheng Yong, Dmitry S. Yershov, and Emilio Frazzoli · 2016
Later among the works it cites.
Prioritized experience replay
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2016
Later among the works it cites.
Successor features for transfer in reinforcement learning
Andre Barreto, Will Dabney, Remi Munos, Jonathan J Hunt, Tom Schaul, Hado P van Hasselt, and David Silver · 2017
Later among the works it cites.
End-to-end learning for lane keeping of self-driving cars
Z. Chen and X. Huang · 2017
Later among the works it cites.
Deep steering: Learning end-to-end driving model from spatial and temporal visual cues
Lu Chi and Yadong Mu · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deepdriving: Learning affordance for direct perception in autonomous driving
C. Chen, A. Seff, A. Kornhauser, and J. Xiao · 2015
Cited alongside, same era.
Intelligent laser welding through representation, prediction, and control learning: An architecture with deep neural networks and reinforcement learning
Johannes Günther, Patrick M. Pilarski, Gerhard Helfrich, Hao Shen, and Klaus Diepold · 2015
Cited alongside, same era.
Developing a predictive approach to knowledge
Adam White · 2015
Cited alongside, same era.
End to end learning for self-driving cars
Mariusz Bojarski, Davide Del Testa, Daniel Dworakowski, Bernhard Firner, Beat Flepp, Prasoon Goyal, Lawrence D. Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, Xin Zhang, Jake Zhao, and Karol Zieba · 2016
Cited alongside, same era.
Application of real-time machine learning to myoelectric prosthesis control: A case series in adaptive switching
Ann L. Edwards, Michael R. Dawson, Jacqueline S. Hebert, Craig Sherstan, Richard S. Sutton, K. Ming Chan, and Patrick M. Pilarski · 2016
Cited alongside, same era.
Reinforcement learning with unsupervised auxiliary tasks
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z. Leibo, David Silver, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Continuous control with deep reinforcement learning
Timothy Lillicrap, Jonathan J. Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2016
Cited alongside, same era.
Neural network modeling for steering control of an autonomous vehicle
G. Garimella, J. Funke, C. Wang, and M. Kobilarov · 2017
Later among the works it cites.
Predictive representations can link model-based reinforcement learning to model-free mechanisms
Evan M. Russek, Ida Momennejad, Matthew M. Botvinick, Samuel J. Gershman, and Nathaniel D. Daw · 2017
Later among the works it cites.
Deep reinforcement learning framework for autonomous driving
Ahmad Sallab, Mohammed Abdou, Etienne Perot, and Senthil Yogamani · 2017
Later among the works it cites.
Hybrid reward architecture for reinforcement learning
Harm Van Seijen, Mehdi Fatemi, Joshua Romoff, Romain Laroche, Tavian Barnes, and Jeffrey Tsang · 2017
Later among the works it cites.
Off-policy deep reinforcement learning without exploration
Scott Fujimoto, David Meger, and Doina Precup · 2018
Later among the works it cites.
Sina Ghiassian, Andrew Patterson, Martha White, Richard S. Sutton, and Adam White · 2018
Later among the works it cites.
Perception as prediction using general value functions in autonomous driving applications
Sean Scheideman Daniel Graves, Kasra Rezaee · 2019
Later among the works it cites.