Fetching the paper…
Reading the bibliography…
Inverse reinforcement learning (IRL) has become a useful tool for learning behavioral models from demonstration data.
Beitrag zur Theorie des Ferromagnetismus
E. Ising · 1925
Earlier work this paper cites.
On the statistical analysis of dirty pictures
J. Besag · 1986
Earlier work this paper cites.
Cognitive models from subcognitive skills
D. Michie, M. Bain, and J. Hayes-Miches · 1990
Earlier work this paper cites.
Learning to fly
C. Sammut, S. Hurst, D. Kedzier, D. Michie, et al · 1992
Earlier work this paper cites.
Q-learning
C. Watkins and P. Dayan · 1992
Earlier work this paper cites.
Novel type of phase transition in a system of self-driven particles
T. Vicsek, A. Czirók, E. Ben-Jacob, I. Cohen, and O. Shochet · 1995
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra · 1998
Earlier work this paper cites.
Reinforcement learning: an introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Learning finite-state controllers for partially observable environments
N. Meuleau, L. Peshkin, K.-E. Kim, and L. P. Kaelbling · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng and S. Russell · 2000
Earlier work this paper cites.
Self-assembly at all scales
G. M. Whitesides and B. Grzybowski · 2002
Earlier work this paper cites.
Least-squares policy iteration
M. G. Lagoudakis and R. Parr · 2003
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Cited alongside, same era.
Current status of nanomedicine and medical nanorobotics
R. A. Freitas · 2005
Cited alongside, same era.
Programmable matter
S. C. Goldstein, J. D. Campbell, and T. C. Mowry · 2005
Cited alongside, same era.
A survey of consensus problems in multi-agent coordination
W. Ren, R. W. Beard, and E. M. Atkins · 2005
Cited alongside, same era.
From disorder to order in marching locusts
J. Buhl, D. Sumpter, I. D. Couzin, J. J. Hale, E. Despland, E. Miller, and S. J. Simpson · 2006
Cited alongside, same era.
Apprenticeship learning using inverse reinforcement learning and gradient methods
G. Neu and C. Szepesvari · 2007
Cited alongside, same era.
Collective cognition in animal groups
I. D. Couzin · 2009
Later among the works it cites.
Multiagent policy teaching
L. Dufton and K. Larson · 2009
Later among the works it cites.
Autonomous helicopter aerobatics through apprenticeship learning
P. Abbeel, A. Coates, and A. Y. Ng · 2010
Later among the works it cites.
Multi-agent inverse reinforcement learning
S. Natarajan, G. Kunapuli, K. Judah, P. Tadepalli, K. Kersting, and J. Shavlik · 2010
Later among the works it cites.
Distributed sensor networks: a multiagent perspective
V. Lesser, C. L. Ortiz J., and M. Tambe · 2012
Later among the works it cites.
Bayesian nonparametric inverse reinforcement learning
B. Michini and J. P. How · 2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Numerical recipes 3rd edition: The art of scientific computing
W. H. Press · 2007
Cited alongside, same era.
Bayesian inverse reinforcement learning
D. Ramachandran and E. Amir · 2007
Cited alongside, same era.
A game-theoretic approach to apprenticeship learning
U. Syed and R. E. Schapire · 2007
Cited alongside, same era.
Exploiting locality of interactions using a policy-gradient approach in multiagent learning
F. S. Melo · 2008
Cited alongside, same era.
Chimera states: the natural link between coherence and incoherence
E. Omel’chenko, Y. L. Maistrenko, and P. A. Tass · 2008
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Cited alongside, same era.
F. A. Oliehoek · 2012
Later among the works it cites.
Inverse reinforcement learning for decentralized non-cooperative multiagent systems
T. S. Reddy, V. Gopikrishna, G. Zaruba, and M. Huber · 2012
Later among the works it cites.
Collective motion
T. Vicsek and A. Zafeiris · 2012
Later among the works it cites.
A survey of inverse reinforcement learning techniques
S. Zhifei and E. M. Joo · 2012
Later among the works it cites.
Convergence of probability measures
P. Billingsley · 2013
Later among the works it cites.