Fetching the paper…
Reading the bibliography…
In this paper, we describe a novel approach to imitation learning that infers latent policies directly from state observations.
Alvinn: An autonomous land vehicle in a neural network
Pomerleau, D. A · 1989
Earlier work this paper cites.
Robot learning from demonstration
Schaal, S · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
Sutton, R. S. and Barto, A. G · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Ng, A. A. Y. and Russell, S. J. S · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y · 2004
Earlier work this paper cites.
The functional role of the parieto-frontal mirror circuit: interpretations and misinterpretations
Rizzolatti, G. and Sinigaglia, C · 2010
Earlier work this paper cites.
Apprenticeship learning about multiple intentions
Babes, M., Marivate, V., Subramanian, K., and Littman, M. L · 2011
Earlier work this paper cites.
Robot learning from human teachers
Chernova, S. and Thomaz, A. L · 2014
Earlier work this paper cites.
Action-conditional video prediction using deep networks in atari games
Oh, J., Guo, X., Lee, H., Lewis, R. L., and Singh, S · 2015
Earlier work this paper cites.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Chen, X., Duan, Y., Houthooft, R., Schulman, J., Sutskever, I., and Abbeel, P · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
Ho, J. and Ermon, S · 2016
Cited alongside, same era.
Unsupervised perceptual rewards for imitation learning
Sermanet, P., Xu, K., and Levine, S · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., van den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., Sander, D., Grewe, D., Nham, J., Kalchbrenner, N., Sutskever, I., Lillicrap, T., Leach, M., Kavukcuoglu, K., Graepel, T., and Hassabis, D · 2016
Cited alongside, same era.
Recurrent environment simulators
Chiappa, S., Racaniere, S., Wierstra, D., and Mohamed, S · 2017
Cited alongside, same era.
Openai baselines
Dhariwal, P., Hesse, C., Klimov, O., Nichol, A., Plappert, M., Radford, A., Schulman, J., Sidor, S., and Wu, Y · 2017
Cited alongside, same era.
State aware imitation learning
Schroecker, Y. and Isbell, C. L · 2017
Later among the works it cites.
Time-contrastive networks: Self-supervised learning from multi-view observation
Sermanet, P., Lynch, C., Hsu, J., and Levine, S · 2017
Later among the works it cites.
Third-person imitation learning
Stadie, B. C., Abbeel, P., and Sutskever, I · 2017
Later among the works it cites.
Toward multimodal image-to-image translation
Zhu, J.-Y., Zhang, R., Pathak, D., Darrell, T., Efros, A. A., Wang, O., and Shechtman, E · 2017
Later among the works it cites.
Playing hard exploration games by watching youtube
Aytar, Y., Pfaff, T., Budden, D., Paine, T. L., Wang, Z., and de Freitas, N · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets
Hausman, K., Chebotar, Y., Schaal, S., Sukhatme, G., and Lim, J. J · 2017
Cited alongside, same era.
Infogail: Interpretable imitation learning from visual demonstrations
Li, Y., Song, J., and Ermon, S · 2017
Cited alongside, same era.
Imitation from observation: Learning to imitate behaviors from raw video via context translation
Liu, Y., Gupta, A., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Overcoming exploration in reinforcement learning with demonstrations
Nair, A., McGrew, B., Andrychowicz, M., Zaremba, W., and Abbeel, P · 2017
Cited alongside, same era.
Behavioral cloning from observation
Torabi, F., Warnell, G., and Stone, P
Cited in the paper.
Generative adversarial imitation from observation
Torabi, F., Warnell, G., and Stone, P
Cited in the paper.
Closest in time.
Quantifying generalization in reinforcement learning
Cobbe, K., Klimov, O., Hesse, C., Kim, T., and Schulman, J · 2018
Closest in time.
Forward-backward reinforcement learning
Edwards, A. D., Downs, L., and Davidson, J. C · 2018
Closest in time.
Recall traces: Backtracking models for efficient reinforcement learning
Goyal, A., Brakel, P., Fedus, W., Lillicrap, T., Levine, S., Larochelle, H., and Bengio, Y · 2018
Closest in time.
Pathak, D., Mahmoudieh, P., Luo, G., Agrawal, P., Chen, D., Shentu, Y., Shelhamer, E., Malik, J., Efros, A. A., and Darrell, T · 2018
Closest in time.