Fetching the paper…
Reading the bibliography…
We present an algorithm for Inverse Reinforcement Learning (IRL) from expert state observations only.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Generative adversarial networks, 2014
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Earlier work this paper cites.
Improved techniques for training gans
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Earlier work this paper cites.
Towards principled methods for training generative adversarial networks
Martín Arjovsky and Léon Bottou · 2017
Earlier work this paper cites.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Do deep generative models know what they don’t know?
Eric T. Nalisnick, Akihiro Matsukawa, Yee Whye Teh, Dilan Görür, and Balaji Lakshminarayanan · 2019
Cited alongside, same era.
Neural spline flows
Conor Durkan, Artur Bekasov, Iain Murray, and George Papamakarios · 2019
Cited alongside, same era.
A divergence minimization perspective on imitation learning methods
Seyed Kamyar Seyed Ghasemipour, Richard Zemel, and Shixiang Gu · 2020
Cited alongside, same era.
f-irl: Inverse reinforcement learning via state marginal matching
Tianwei Ni, Harshit Sikchi, Yufei Wang, Tejus Gupta, Lisa Lee, and Ben Eysenbach · 2020
Imitation with neural density models
Kuno Kim, Akshat Jindal, Yang Song, Jiaming Song, Yanan Sui, and Stefano Ermon · 2020
Later among the works it cites.
Primal wasserstein imitation learning
Robert Dadashi, Léonard Hussenot, Matthieu Geist, and Olivier Pietquin · 2020
Later among the works it cites.
Why normalizing flows fail to detect out-of-distribution data
Polina Kirichenko, Pavel Izmailov, and Andrew G Wilson · 2020
Later among the works it cites.
Probability distillation: A caveat and alternatives
Chin-Wei Huang, Faruk Ahmed, Kundan Kumar, Alexandre Lacoste, and Aaron Courville · 2020
Later among the works it cites.
Noise regularization for conditional density estimation, 2020
Jonas Rothfuss, Fabio Ferreira, Simon Boehm, Simon Walther, Maxim Ulrich, Tamim Asfour, and Andreas Krause · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Softflow: Probabilistic framework for normalizing flow on manifolds
Hyeongju Kim, Hyeonseung Lee, Woo Hyun Kang, Joun Yeop Lee, and Nam Soo Kim · 2020
Later among the works it cites.