Fetching the paper…
Reading the bibliography…
Continuous control and planning remains a major challenge in robotics and machine learning.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 1908
Earlier work this paper cites.
Dynamic programming
R Bellman · 1957
Earlier work this paper cites.
Information theory and statistical mechanics
Edwin T Jaynes · 1957
Earlier work this paper cites.
Maximum likelihood from incomplete data via the em algorithm
Arthur P Dempster, Nan M Laird, and Donald B Rubin · 1977
Earlier work this paper cites.
Learning agents for uncertain environments
Stuart Russell · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton, Andrew G Barto, et al · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Bayesian inverse reinforcement learning
Deepak Ramachandran and Eyal Amir · 2007
Cited alongside, same era.
Inverse optimal control with linearly-solvable mdps
Krishnamurthy Dvijotham and Emanuel Todorov · 2010
Cited alongside, same era.
Apprenticeship learning about multiple intentions
Monica Babes, Vukosi Marivate, Kaushik Subramanian, and Michael L Littman · 2011
Cited alongside, same era.
Bayesian theory of mind: Modeling joint belief-desire attribution
Chris Baker, Rebecca Saxe, and Joshua Tenenbaum · 2011
Cited alongside, same era.
Relative entropy inverse reinforcement learning
Abdeslam Boularias, Jens Kober, and Jan Peters · 2011
Cited alongside, same era.
Inverse reinforcement learning in partially observable environments
Jaedeug Choi and Kee-Eung Kim · 2011
Cited alongside, same era.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Later among the works it cites.
Guided cost learning: Deep inverse optimal control via policy optimization
Chelsea Finn, Sergey Levine, and Pieter Abbeel · 2016
Later among the works it cites.
Inverse reinforcement learning with simultaneous estimation of rewards and dynamics
Michael Herman, Tobias Gindele, Jörg Wagner, Felix Schmitt, and Wolfram Burgard · 2016
Later among the works it cites.
I see what you see: Inferring sensor and policy models of human real-world motor behavior
Felix Schmitt, Hans-Joachim Bieg, Michael Herman, and Constantin A Rothkopf · 2017
Later among the works it cites.
A dynamic bayesian observer model reveals origins of bias in visual path integration
Kaushik J Lakshminarasimhan, Marina Petsalis, Hyeshin Park, Gregory C DeAngelis, Xaq Pitkow, and Dora E Angelaki · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning objective functions for manipulation
Mrinal Kalakrishnan, Peter Pastor, Ludovic Righetti, and Stefan Schaal · 2013
Cited alongside, same era.
Maximum likelihood inverse reinforcement learning
Monica C Vroman · 2014
Cited alongside, same era.
Later among the works it cites.
Where do you think you’re going?: Inferring beliefs about dynamics from behavior
Siddharth Reddy, Anca D. Dragan, and Sergey Levine · 2018
Later among the works it cites.
Inverse POMDP: Inferring what you think from what you do
Zhengwei Wu, Paul Schrater, and Xaq Pitkow · 2018
Later among the works it cites.