Fetching the paper…
Reading the bibliography…
Inverse reinforcement learning attempts to reconstruct the reward function in a Markov decision problem, using observations of agent actions.
When Is a Linear Control System Optimal?
R. E. Kalman · 1964
Earlier work this paper cites.
Decisions with Multiple Objectives: Preferences and Value Trade-Offs
Ralph L. Keeney and Howard Raiffa · 1976
Earlier work this paper cites.
Econometric policy evaluation: A critique
Robert E. Lucas · 1976
Earlier work this paper cites.
Estimation of dynamic labor demand schedules under rational expectations
Thomas J Sargent · 1978
Earlier work this paper cites.
Linear matrix inequalities in system and control theory
Stephen Boyd, Laurent El Ghaoui, Eric Feron, and Venkataramanan Balakrishnan · 1994
Earlier work this paper cites.
Learning agents for uncertain environments
Stuart Russell · 1998
Earlier work this paper cites.
Policy invariance under reward transformations: Theory and application to reward shaping
Andrew Y. Ng, Daishi Harada, and Stuart Russell · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng and Stuart Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Stochastic optimal control: the discrete-time case
Dimitir P Bertsekas and Steven Shreve · 2004
Cited alongside, same era.
Maximum margin planning
Nathan D. Ratliff, J. Andrew Bagnell, and Martin A. Zinkevich · 2006
Cited alongside, same era.
Non-negative Matrices and Markov chains
E. Seneta · 2006
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Cited alongside, same era.
Inverse optimal control with linearly-solvable mdps
Krishnamurthy Dvijotham and Emanuel Todorov · 2010
Cited alongside, same era.
Modeling Purposeful Adaptive Behavior with the Principle of Maximum Causal Entropy
Brian D Ziebart · 2010
Cited alongside, same era.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Later among the works it cites.
Towards resolving unidentifiability in inverse reinforcement learning
Kareem Amin and Satinder Singh · 2016
Later among the works it cites.
Repeated inverse reinforcement learning
Kareem Amin, Nan Jiang, and Satinder Singh · 2017
Later among the works it cites.
Reinforcement learning with deep energy-based policies
Tuomas Haarnoja, Haoran Tang, Pieter Abbeel, and Sergey Levine · 2017
Later among the works it cites.
Learning robust rewards with adverserial inverse reinforcement learning
Justin Fu, Katie Luo, and Sergey Levine · 2018
Later among the works it cites.
Reinforcement learning and control as probabilistic inference: Tutorial and review
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Relative entropy inverse reinforcement learning
Abdeslam Boularias, Jens Kober, and Jan Peters · 2011
Cited alongside, same era.
A weak convergence approach to the theory of large deviations , volume 902
Paul Dupuis and Richard S Ellis · 2011
Cited alongside, same era.
Nonlinear inverse reinforcement learning with gaussian processes
Sergey Levine, Zoran Popovic, and Vladlen Koltun · 2011
Cited alongside, same era.
Chelsea Finn, Paul Christiano, Pieter Abbeel, and Sergey Levine
Cited in the paper.
Guided cost learning: Deep inverse optimal control via policy optimization
Chelsea Finn, Sergey Levine, and Pieter Abbeel
Cited in the paper.
Sergey Levine · 2018
Later among the works it cites.
Efficient exploration of reward functions in inverse reinforcement learning via bayesian optimization
Sreejith Balakrishnan, Quoc Phong Nguyen, Bryan Kian Hsiang Low, and Harold Soh · 2020
Later among the works it cites.
Reward identification in inverse reinforcement learning
Kuno Kim, Shivam Garg, Kirankumar Shiragur, and Stefano Ermon · 2021
Closest in time.