Fetching the paper…
Reading the bibliography…
We consider the Imitation Learning (IL) setup where expert data are not collected on the actual deployment environment but on a different version.
Dream to control: Learning behaviors by latent imagination
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi · 1912
Earlier work this paper cites.
Model predictive heuristic control: application to industrial processes
J. Rault, A. Richalet, J. Testud, and J. Papon · 1978
Earlier work this paper cites.
Model predictive heuristic control
J. Richalet, A. Rault, J. Testud, and J. Papon · 1978
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. A. Pomerleau · 1989
Earlier work this paper cites.
Efficient training of artificial neural networks for autonomous navigation
D. A. Pomerleau · 1991
Earlier work this paper cites.
Approximate nearest neighbors: towards removing the curse of dimensionality
P. Indyk and R. Motwani · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng, S. J. Russell, et al · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Efficient reductions for imitation learning
S. Ross and D. Bagnell · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. Gordon, and D. Bagnell · 2011
Earlier work this paper cites.
Synthesis and stabilization of complex behaviors through online trajectory optimization
Y. Tassa, T. Erez, and E. Todorov · 2012
Earlier work this paper cites.
Probabilistic model-based imitation learning
P. Englert, A. Paraschos, M. P. Deisenroth, and J. Peters · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Model predictive path integral control using covariance variable importance sampling
G. Williams, A. Aldrich, and E. Theodorou · 2015
Earlier work this paper cites.
End to end learning for self-driving cars
M. Bojarski, D. Del Testa, D. Dworakowski, B. Firner, B. Flepp, P. Goyal, L. D. Jackel, M. Monfort, U. Muller, J. Zhang, et al · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
Simple and scalable predictive uncertainty estimation using deep ensembles
B. Lakshminarayanan, A. Pritzel, and C. Blundell · 2017
Earlier work this paper cites.
Information theoretic mpc for model-based reinforcement learning
G. Williams, N. Wagener, B. Goldfain, P. Drews, J. M. Rehg, B. Boots, and E. A. Theodorou · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
End-to-end differentiable adversarial imitation learning
N. Baram, O. Anschel, I. Caspi, and S. Mannor · 2017
Earlier work this paper cites.
One-shot visual imitation learning via meta-learning
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine · 2017
Earlier work this paper cites.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Earlier work this paper cites.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Earlier work this paper cites.
Evaluating reinforcement learning algorithms in observational health settings
O. Gottesman, F. Johansson, J. Meier, J. Dent, D. Lee, S. Srinivasan, L. Zhang, Y. Ding, D. Wihl, X. Peng, et al · 2018
Earlier work this paper cites.
Using simulation and domain adaptation to improve efficiency of deep robotic grasping
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige, et al · 2018
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, et al · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
K. Chua, R. Calandra, R. McAllister, and S. Levine · 2018
Cited alongside, same era.
Roboturk: A crowdsourcing platform for robotic skill learning through imitation
A. Mandlekar, Y. Zhu, A. Garg, J. Booher, M. Spero, A. Tung, J. Gao, J. Emmons, A. Gupta, E. Orbay, et al · 2018
Cited alongside, same era.
Safe end-to-end imitation learning for model predictive control
K. Lee, K. Saigol, and E. A. Theodorou · 2018
Cited alongside, same era.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, and R. Martín-Martín · 2020
Later among the works it cites.
A divergence minimization perspective on imitation learning methods
S. K. S. Ghasemipour, R. Zemel, and S. Gu · 2020
Later among the works it cites.
Deep dynamics models for learning dexterous manipulation
A. Nagabandi, K. Konolige, S. Levine, and V. Kumar · 2020
Later among the works it cites.
Behavioral cloning from noisy demonstrations
F. Sasaki and R. Yamashina · 2020
Later among the works it cites.
Accelerating large-scale inference with anisotropic vector quantization
R. Guo, P. Sun, E. Lindgren, Q. Geng, D. Simcha, F. Chern, and S. Kumar · 2020
Later among the works it cites.
Safari: Safe and active robot imitation learning with imagination
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Imitation learning via kernel mean embedding
K.-E. Kim and H. S. Park · 2018
Cited alongside, same era.
An algorithmic perspective on imitation learning
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, J. Peters, et al · 2018
Cited alongside, same era.
Deep imitative models for flexible inference, planning, and control
N. Rhinehart, R. McAllister, and S. Levine · 2018
Cited alongside, same era.
Task-embedded control networks for few-shot imitation learning
S. James, M. Bloesch, and A. J. Davison · 2018
Cited alongside, same era.
One-shot imitation from observing humans via domain-adaptive meta-learning
T. Yu, C. Finn, A. Xie, S. Dasari, T. Zhang, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Preventing undesirable behavior of intelligent machines
P. S. Thomas, B. C. da Silva, A. G. Barto, S. Giguere, Y. Brun, and E. Brunskill · 2019
Cited alongside, same era.
Sqil: Imitation learning via reinforcement learning with sparse rewards
S. Reddy, A. D. Dragan, and S. Levine · 2019
Cited alongside, same era.
N. Di Palo and E. Johns · 2020
Later among the works it cites.
Can autonomous vehicles identify, recover from, and adapt to distribution shifts?
A. Filos, P. Tigkas, R. McAllister, N. Rhinehart, S. Levine, and Y. Gal · 2020
Later among the works it cites.
Model-based behavioral cloning with future image similarity learning
A. Wu, A. Piergiovanni, and M. S. Ryoo · 2020
Later among the works it cites.
Learning latent plans from play
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet · 2020
Later among the works it cites.
Active domain randomization
B. Mehta, M. Diaz, F. Golemo, C. J. Pal, and L. Paull · 2020
Later among the works it cites.
Acme: A research framework for distributed reinforcement learning
M. Hoffman, B. Shahriari, J. Aslanides, G. Barth-Maron, F. Behbahani, T. Norman, A. Abdolmaleki, A. Cassirer, F. Yang, K. Baumli, et al · 2020
Later among the works it cites.
Learning dynamics models for model predictive agents
M. Lutter, L. Hasenclever, A. Byravan, G. Dulac-Arnold, P. Trochim, N. Heess, J. Merel, and Y. Tassa · 2021
Later among the works it cites.
Primal wasserstein imitation learning
R. Dadashi, L. Hussenot, M. Geist, and O. Pietquin · 2021
Later among the works it cites.
What matters in learning from offline human demonstrations for robot manipulation
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Martín-Martín · 2021
Later among the works it cites.
What matters for adversarial imitation learning?
M. Orsini, A. Raichuk, L. Hussenot, D. Vincent, R. Dadashi, S. Girgin, M. Geist, O. Bachem, O. Pietquin, and M. Andrychowicz · 2021
Later among the works it cites.
Combo: Conservative offline model-based policy optimization
T. Yu, A. Kumar, R. Rafailov, A. Rajeswaran, S. Levine, and C. Finn · 2021
Later among the works it cites.
Offline reinforcement learning from images with latent space models
R. Rafailov, T. Yu, A. Rajeswaran, and C. Finn · 2021
Later among the works it cites.
Model-based offline planning with trajectory pruning
X. Zhan, X. Zhu, and H. Xu · 2021
Later among the works it cites.
No need for interactions: Robust model-based imitation learning using neural ode
H. Lin, B. Li, X. Zhou, J. Wang, and M. Q.-H. Meng · 2021
Later among the works it cites.
A survey of generalisation in deep reinforcement learning
R. Kirk, A. Zhang, E. Grefenstette, and T. Rocktäschel · 2021
Later among the works it cites.
Bc-z: Zero-shot task generalization with robotic imitation learning
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn · 2021
Later among the works it cites.
Kitchenshift: Evaluating zero-shot generalization of imitation-based policy learning under domain shifts
E. Xing, A. Gupta, S. Powers, and V. Dean · 2021
Later among the works it cites.
Actionable models: Unsupervised offline reinforcement learning of robotic skills
Y. Chebotar, K. Hausman, Y. Lu, T. Xiao, D. Kalashnikov, J. Varley, A. Irpan, B. Eysenbach, R. Julian, C. Finn, et al · 2021
Later among the works it cites.
Implicit behavioral cloning
P. Florence, C. Lynch, A. Zeng, O. A. Ramirez, A. Wahid, L. Downs, A. Wong, J. Lee, I. Mordatch, and J. Tompson · 2022
Later among the works it cites.
Imitating, fast and slow: Robust learning from demonstrations via decision-time planning
C. Qi, P. Abbeel, and A. Grover · 2022
Later among the works it cites.