Fetching the paper…
Reading the bibliography…
In many environments, only a relatively small subset of the complete state space is necessary in order to accomplish a given task.
Pattern-recognizing control systems
Bernard Widrow and Fred W Smith · 1964
Earlier work this paper cites.
Practical reinforcement learning in continuous spaces
William D Smart and Leslie Pack Kaelbling · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Exploration and apprenticeship learning in reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2005
Earlier work this paper cites.
Apprenticeship learning for initial value functions in reinforcement learning
Frederic Maire and Vadim Bulitko · 2005
Earlier work this paper cites.
Empirical bernstein bounds and sample variance penalization
Andreas Maurer and Massimiliano Pontil · 2009
Earlier work this paper cites.
UAV cooperative control with stochastic risk models
Alborz Geramifard, Joshua Redding, Nicholas Roy, and Jonathan P How · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stephane Ross, Geoffrey J Gordon, and J Andrew Bagnell · 2011
Earlier work this paper cites.
Safe exploration of state and action spaces in reinforcement learning
Javier Garcia and Fernando Fernández · 2012
Earlier work this paper cites.
Learning monocular reactive UAV control in cluttered natural environments
Stéphane Ross, Narek Melik-Barkhudarov, Kumar Shaurya Shankar, Andreas Wendel, Debadeepta Dey, J Andrew Bagnell, and Martial Hebert · 2013
Cited alongside, same era.
An invitation to imitation
J Andrew Bagnell · 2015
Cited alongside, same era.
A comprehensive survey on safe reinforcement learning
Javier Garcıa and Fernando Fernández · 2015
Cited alongside, same era.
OpenAI Gym, 2016
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
SHIV: Reducing supervisor burden in DAgger using support vectors for efficient learning from demonstrations in high dimensional state spaces
Michael Laskey, Sam Staszak, Wesley Yu-Shu Hsieh, Jeffrey Mahler, Florian T Pokorny, Anca D Dragan, and Ken Goldberg · 2016
Safe reinforcement learning via shielding
Mohammed Alshiekh, Roderick Bloem, Rüdiger Ehlers, Bettina Könighofer, Scott Niekum, and Ufuk Topcu · 2018
Later among the works it cites.
JAX: composable transformations of Python+NumPy programs, 2018
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, and Skye Wanderman-Milne · 2018
Later among the works it cites.
Leave no trace: learning to reset for safe and autonomous reinforcement learning, 2018
Benjamin Eysenbach, Shixiang Gu, Julian Ibarz, and Sergey Levine · 2018
Later among the works it cites.
Deep Q-learning from demonstrations
Todd Hester, Matej Vecerik, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, and Ian Osband · 2018
Later among the works it cites.
Variance reduction methods for sublinear reinforcement learning
Sham Kakade, Mengdi Wang, and Lin F Yang · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Minimax regret bounds for reinforcement learning
Mohammad Gheshlaghi Azar, Ian Osband, and Rémi Munos · 2017
Cited alongside, same era.
Uncertainty-aware reinforcement learning for collision avoidance
Gregory Kahn, Adam Villaflor, Vitchyr Pong, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Safe visual navigation via deep learning and novelty detection
Charles Richter and Nicholas Roy · 2017
Cited alongside, same era.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Anusha Nagabandi, Gregory Kahn, Ronald S. Fearing, and Sergey Levine · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto · 2018
Later among the works it cites.
Generative adversarial imitation from observation
Faraz Torabi, Garrett Warnell, and Peter Stone · 2018
Later among the works it cites.