Fetching the paper…
Reading the bibliography…
Reinforcement learning can acquire complex behaviors from high-level specifications.
Policy invariance under reward transformations: Theory and application to reward shaping
Ng, A., Harada, D., and Russell, S · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Ng, A., Russell, S., et al · 2000
Earlier work this paper cites.
Covariant policy search
Bagnell, J. A. and Schneider, J · 2003
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A · 2004
Earlier work this paper cites.
Maximum margin planning
Ratliff, N., Bagnell, J. A., and Zinkevich, M. A · 2006
Earlier work this paper cites.
Linearly-solvable markov decision problems
Todorov, E · 2006
Earlier work this paper cites.
Bayesian inverse reinforcement learning
Ramachandran, D. and Amir, E · 2007
Earlier work this paper cites.
Boosting structured prediction for imitation learning
Ratliff, N., Bradley, D., Bagnell, J. A., and Chestnutt, J · 2007
Earlier work this paper cites.
Learning to search: Functional gradient techniques for imitation learning
Ratliff, N., Silver, D., and Bagnell, J. A · 2009
Earlier work this paper cites.
Relative entropy policy search
Peters, J., Mülling, K., and Altün, Y · 2010
Earlier work this paper cites.
Modeling purposeful adaptive behavior with the principle of maximum causal entropy
Ziebart, B · 2010
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning in continuous state spaces with path integrals
Aghasadeghi, N. and Bretl, T · 2011
Cited alongside, same era.
Relative entropy inverse reinforcement learning
Boularias, A., Kober, J., and Peters, J · 2011
Cited alongside, same era.
Nonlinear inverse reinforcement learning with gaussian processes
Levine, S., Popovic, Z., and Koltun, V · 2011
Cited alongside, same era.
Formalizing assistive teleoperation
Dragan, Anca and Srinivasa, Siddhartha · 2012
Cited alongside, same era.
Continuous inverse optimal control with locally optimal examples
Levine, S. and Koltun, V · 2012
Cited alongside, same era.
MuJoCo: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Cited alongside, same era.
Learning strategies in table tennis using inverse reinforcement learning
Muelling, K., Boularias, A., Mohler, B., Schölkopf, B., and Peters, J · 2014
Later among the works it cites.
Maximum entropy inverse reinforcement learning
Ziebart, B., Maas, A., Bagnell, J. A., and Dey, A. K · 2014
Later among the works it cites.
Maximum Entropy Semi-Supervised Inverse Reinforcement Learning
Audiffren, J., Valko, M., Lazaric, A., and Ghavamzadeh, M · 2015
Later among the works it cites.
Graph-based inverse optimal control for robot manipulation
Byravan, A., Monfort, M., Ziebart, B., Boots, B., and Fox, D · 2015
Later among the works it cites.
Direct loss minimization inverse optimal control
Doerr, A., Ratliff, N., Bohg, J., Toussaint, M., and Schaal, S · 2015
Later among the works it cites.
Learning contact-rich manipulation skills with guided policy search
Levine, S., Wagener, N., and Abbeel, P · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning objective functions for manipulation
Kalakrishnan, M., Pastor, P., Righetti, L., and Schaal, S · 2013
Cited alongside, same era.
On stochastic optimal control and reinforcement learning by approximate inference
Rawlik, K. and Vijayakumar, S · 2013
Cited alongside, same era.
Action-reaction: Forecasting the dynamics of human interaction
Huang, D. and Kitani, K · 2014
Cited alongside, same era.
Learning neural network policies with guided policy search under unknown dynamics
Levine, S. and Abbeel, P · 2014
Cited alongside, same era.
Later among the works it cites.
Softstar: Heuristic-guided probabilistic inference
Monfort, M., Lake, B. M., Ziebart, B., Lucey, P., and Tenenbaum, J · 2015
Later among the works it cites.
Simultaneous deep transfer across domains and tasks
Tzeng, E., Hoffman, J., Darrell, T., and Saenko, K · 2015
Later among the works it cites.
Maximum entropy deep inverse reinforcement learning
Wulfmeier, M., Ondruska, P., and Posner, I · 2015
Later among the works it cites.
Deep spatial autoencoders for visuomotor learning
Finn, Chelsea, Tan, Xin Yu, Duan, Yan, Darrell, Trevor, Levine, Sergey, and Abbeel, Pieter · 2016
Closest in time.