Fetching the paper…
Reading the bibliography…
This paper considers the problem of inverse reinforcement learning in zero-sum stochastic games when expert demonstrations are known to be not optimal.
Learning agents for uncertain environments
Stuart Russell · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Bayesian inverse reinforcement learning
Deepak Ramachandran and Eyal Amir · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
Multiagent reinforcement learning: algorithm converging to nash equilibrium in general-sum discounted stochastic games
Natalia Akchurina · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey J Gordon, and Drew Bagnell · 2011
Cited alongside, same era.
Inverse reinforcement learning for decentralized non-cooperative multiagent systems
Tummalapalli Reddy, Vamsikrishna Gopikrishna, Gergely Zaruba, and Manfred Huber · 2012
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Two-timescale algorithms for learning nash equilibria in general-sum stochastic games
Prasad H.L., Prashanth L.A., and Shalabh Bhatnagar · 2015
Cited alongside, same era.
Guided cost learning: Deep inverse optimal control via policy optimization
Chelsea Finn, Sergey Levine, and Pieter Abbeel · 2016
Later among the works it cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Later among the works it cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Multi-agent inverse reinforcement learning for two-person zero-sum games
Xiaomin Lin, Peter A Beling, and Randy Cogill · 2017
Later among the works it cites.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Prasad H.L. and Shalabh Bhatnagar · 2015
Cited alongside, same era.
Density matching reward learning
Sungjoon Choi, Kyungjae Lee, Andy Park, and Songhwai Oh · 2016
Cited alongside, same era.
Jakob N Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2017
Later among the works it cites.