Fetching the paper…
Reading the bibliography…
In recent years, the development of robotics and artificial intelligence (AI) systems has been nothing short of remarkable.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Advances in neural information processing systems , vol. 1, 1988
1988
Earlier work this paper cites.
D. A. Pomerleau, “Efficient training of artificial neural networks for autonomous navigation,” Neural computation , vol. 3, no. 1, pp. 88–97, 1991
1991
Earlier work this paper cites.
S. Russell, “Learning agents for uncertain environments,” in Proceedings of the eleventh annual conference on Computational learning theory , 1998, pp. 101–103
1998
Earlier work this paper cites.
S. Schaal, “Is imitation learning the route to humanoid robots?” Trends in cognitive sciences , vol. 3, no. 6, pp. 233–242, 1999
1999
Earlier work this paper cites.
A. Y. Ng, S. Russell et al. , “Algorithms for inverse reinforcement learning.” in Icml , vol. 1, 2000, p. 2
2000
Earlier work this paper cites.
A. J. Ijspeert, J. Nakanishi, and S. Schaal, “Movement imitation with nonlinear dynamical systems in humanoid robots,” in Proceedings 2002 IEEE International Conference on Robotics and Automation (Cat. No. 02CH37292) , vol. 2. IEEE, 2002, pp. 1398–1403
2002
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in Proceedings of the twenty-first international conference on Machine learning , 2004, p. 1
2004
Earlier work this paper cites.
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich, “Maximum margin planning,” in Proceedings of the 23rd international conference on Machine learning , 2006, pp. 729–736
2006
Earlier work this paper cites.
J. Bagnell, J. Chestnutt, D. Bradley, and N. Ratliff, “Boosting structured prediction for imitation learning,” Advances in Neural Information Processing Systems , vol. 19, 2006
2006
Earlier work this paper cites.
U. Syed and R. E. Schapire, “A game-theoretic approach to apprenticeship learning,” Advances in neural information processing systems , vol. 20, 2007
2007
Earlier work this paper cites.
D. Ramachandran and E. Amir, “Bayesian inverse reinforcement learning.” in IJCAI , vol. 7, 2007, pp. 2586–2591
2007
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, A. K. Dey et al. , “Maximum entropy inverse reinforcement learning.” in Aaai , vol. 8. Chicago, IL, USA, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
L. Van der Maaten and G. Hinton, “Visualizing data using t-sne.” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
N. D. Ratliff, D. Silver, and J. A. Bagnell, “Learning to search: Functional gradient techniques for imitation learning,” Autonomous Robots , vol. 27, no. 1, pp. 25–53, 2009
2009
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
N. Aghasadeghi and T. Bretl, “Maximum entropy inverse reinforcement learning in continuous state spaces with path integrals,” in 2011 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2011, pp. 1561–1566
2011
Earlier work this paper cites.
A. Boularias, J. Kober, and J. Peters, “Relative entropy inverse reinforcement learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 182–189
2011
Earlier work this paper cites.
J. Choi and K.-E. Kim, “Map inference for bayesian inverse reinforcement learning,” Advances in Neural Information Processing Systems , vol. 24, 2011
2011
Earlier work this paper cites.
S. Levine, Z. Popovic, and V. Koltun, “Nonlinear inverse reinforcement learning with gaussian processes,” Advances in neural information processing systems , vol. 24, 2011
2011
Earlier work this paper cites.
2012
Earlier work this paper cites.
M. Kalakrishnan, P. Pastor, L. Righetti, and S. Schaal, “Learning objective functions for manipulation,” in 2013 IEEE International Conference on Robotics and Automation . IEEE, 2013, pp. 1331–1336
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” Advances in neural information processing systems , vol. 27, 2014
2014
Earlier work this paper cites.
M. Kuderer, S. Gulati, and W. Burgard, “Learning driving styles for autonomous vehicles from demonstration,” in 2015 IEEE international conference on robotics and automation (ICRA) . IEEE, 2015, pp. 2641–2646
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
B. Piot, M. Geist, and O. Pietquin, “Bridging the gap between imitation learning and inverse reinforcement learning,” IEEE transactions on neural networks and learning systems , vol. 28, no. 8, pp. 1814–1826, 2016
2016
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
D. Hadfield-Menell, S. J. Russell, P. Abbeel, and A. Dragan, “Cooperative inverse reinforcement learning,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
C. Finn, S. Levine, and P. Abbeel, “Guided cost learning: Deep inverse optimal control via policy optimization,” in International conference on machine learning . PMLR, 2016, pp. 49–58
2016
Earlier work this paper cites.
J. Ho, J. Gupta, and S. Ermon, “Model-free imitation learning with policy optimization,” in International Conference on Machine Learning . PMLR, 2016, pp. 2760–2769
2016
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Computing Surveys (CSUR) , vol. 50, no. 2, pp. 1–35, 2017
2017
Earlier work this paper cites.
J. Zhang and K. Cho, “Query-efficient imitation learning for end-to-end simulated driving,” in Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence , ser. AAAI’17. AAAI Press, 2017, p. 2891–2897
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Arjovsky, S. Chintala, and L. Bottou, “Wasserstein generative adversarial networks,” in International conference on machine learning . PMLR, 2017, pp. 214–223
2017
Earlier work this paper cites.
Y. Li, J. Song, and S. Ermon, “Infogail: Interpretable imitation learning from visual demonstrations,” Advances in Neural Information Processing Systems , vol. 30, 2017
2017
Earlier work this paper cites.
W. Sun, A. Venkatraman, G. J. Gordon, B. Boots, and J. A. Bagnell, “Deeply aggrevated: Differentiable imitation learning for sequential prediction,” in International conference on machine learning . PMLR, 2017, pp. 3309–3318
2017
Earlier work this paper cites.
B. C. Stadie, P. Abbeel, and I. Sutskever, “Third-person imitation learning,” in International conference on learning representations , 2017
2017
Earlier work this paper cites.
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, J. Peters et al. , “An algorithmic perspective on imitation learning,” Foundations and Trends® in Robotics , vol. 7, no. 1-2, pp. 1–179, 2018
2018
Earlier work this paper cites.
M. Deng, Z. Li, Y. Kang, C. P. Chen, and X. Chu, “A learning-based hierarchical control scheme for an exoskeleton robot in human–robot cooperative manipulation,” IEEE transactions on cybernetics , vol. 50, no. 1, pp. 112–125, 2018
2018
Cited alongside, same era.
Z. Zhu and H. Hu, “Robot learning from demonstration in robotic assembly: A survey,” Robotics , vol. 7, no. 2, p. 17, 2018
2018
Cited alongside, same era.
Y. Aytar, T. Pfaff, D. Budden, T. Paine, Z. Wang, and N. De Freitas, “Playing hard exploration games by watching youtube,” Advances in neural information processing systems , vol. 31, 2018
2018
Cited alongside, same era.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
J. Song, H. Ren, D. Sadigh, and S. Ermon, “Multi-agent generative adversarial imitation learning,” Advances in neural information processing systems , vol. 31, 2018
R. Hoque, A. Balakrishna, C. Putterman, M. Luo, D. S. Brown, D. Seita, B. Thananjeyan, E. Novoseller, and K. Goldberg, “Lazydagger: Reducing context switching in interactive imitation learning,” in 2021 IEEE 17th International Conference on Automation Science and Engineering (CASE) . IEEE, 2021, pp. 502–509
2021
Later among the works it cites.
R. Dadashi, L. Hussenot, M. Geist, and O. Pietquin, “Primal wasserstein imitation learning,” in International conference on learning representations , 2021
2021
Later among the works it cites.
K. Brantley, “Expert-in-the-loop for sequential decisions and predictions,” Ph.D. dissertation, 2021
2021
Later among the works it cites.
J. Wong, A. Tung, A. Kurenkov, A. Mandlekar, L. Fei-Fei, S. Savarese, and R. Martín-Martín, “Error-aware imitation learning from teleoperation data for mobile manipulation,” in Conference on Robot Learning . PMLR, 2021, pp. 1367–1378
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
J. Fu, K. Luo, and S. Levine, “Learning robust rewards with adversarial inverse reinforcement learning,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
Y. Liu, A. Gupta, P. Abbeel, and S. Levine, “Imitation from observation: Learning to imitate behaviors from raw video via context translation,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 1118–1125
2018
Cited alongside, same era.
P. Sermanet, C. Lynch, Y. Chebotar, J. Hsu, E. Jang, S. Schaal, S. Levine, and G. Brain, “Time-contrastive networks: Self-supervised learning from video,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 1134–1141
2018
Cited alongside, same era.
F. Torabi, G. Warnell, and P. Stone, “Behavioral cloning from observation,” in Proceedings of the 27th International Joint Conference on Artificial Intelligence , ser. IJCAI’18. AAAI Press, 2018, p. 4950–4957
2018
Cited alongside, same era.
W. Jeon, S. Seo, and K.-E. Kim, “A bayesian approach to generative adversarial imitation learning,” Advances in Neural Information Processing Systems , vol. 31, 2018
2018
Cited alongside, same era.
F. Sasaki, T. Yohira, and A. Kawaguchi, “Sample efficient imitation learning for continuous control,” in International conference on learning representations , 2018
2018
Cited alongside, same era.
U. E. Ogenyi, J. Liu, C. Yang, Z. Ju, and H. Liu, “Physical human–robot collaboration: Robotic systems, learning methods, collaborative strategies, sensors, and actuators,” IEEE transactions on cybernetics , vol. 51, no. 4, pp. 1888–1901, 2019
2019
Cited alongside, same era.
J. Chang, M. Uehara, D. Sreenivas, R. Kidambi, and W. Sun, “Mitigating covariate shift in imitation learning via offline data with partial coverage,” Advances in Neural Information Processing Systems , vol. 34, pp. 965–979, 2021
2021
Later among the works it cites.
P. Florence, C. Lynch, A. Zeng, O. A. Ramirez, A. Wahid, L. Downs, A. Wong, J. Lee, I. Mordatch, and J. Tompson, “Implicit behavioral cloning,” in Conference on Robot Learning . PMLR, 2021, pp. 158–168
2021
Later among the works it cites.
2021
Later among the works it cites.
B. Lian, W. Xue, F. L. Lewis, and T. Chai, “Robust inverse q-learning for continuous-time linear systems in adversarial environments,” IEEE Transactions on Cybernetics , 2021
2021
Later among the works it cites.
K. Kim, S. Garg, K. Shiragur, and S. Ermon, “Reward identification in inverse reinforcement learning,” in International Conference on Machine Learning . PMLR, 2021, pp. 5496–5505
2021
Later among the works it cites.
D. Jarrett, A. Hüyük, and M. Van Der Schaar, “Inverse decision modeling: Learning interpretable representations of behavior,” in International Conference on Machine Learning . PMLR, 2021, pp. 4755–4771
2021
Later among the works it cites.
A. J. Chan and M. van der Schaar, “Scalable bayesian inverse reinforcement learning,” in International Conference on Learning Representations , 2021
2021
Later among the works it cites.
A. M. Metelli, G. Ramponi, A. Concetti, and M. Restelli, “Provably efficient learning of transferable rewards,” in International Conference on Machine Learning . PMLR, 2021, pp. 7665–7676
2021
Later among the works it cites.
M. Orsini, A. Raichuk, L. Hussenot, D. Vincent, R. Dadashi, S. Girgin, M. Geist, O. Bachem, O. Pietquin, and M. Andrychowicz, “What matters for adversarial imitation learning?” Advances in Neural Information Processing Systems , vol. 34, pp. 14 656–14 668, 2021
2021
Later among the works it cites.
S. Choe, H. Seong, and E. Kim, “Indoor place category recognition for a cleaning robot by fusing a probabilistic approach and deep learning,” IEEE Transactions on Cybernetics , 2021
2021
Later among the works it cites.
D. S. Raychaudhuri, S. Paul, J. Vanbaar, and A. K. Roy-Chowdhury, “Cross-domain imitation from observations,” in International Conference on Machine Learning . PMLR, 2021, pp. 8902–8912
2021
Later among the works it cites.
A. Jaegle, Y. Sulsky, A. Ahuja, J. Bruce, R. Fergus, and G. Wayne, “Imitation by predicting observations,” in International Conference on Machine Learning . PMLR, 2021, pp. 4665–4676
2021
Later among the works it cites.
Y. Wang, C. Xu, B. Du, and H. Lee, “Learning to weight imperfect demonstrations,” in International Conference on Machine Learning . PMLR, 2021, pp. 10 961–10 970
2021
Later among the works it cites.
J. Lee, W. Jeon, B. Lee, J. Pineau, and K.-E. Kim, “Optidice: Offline policy optimization via stationary distribution correction estimation,” in International Conference on Machine Learning . PMLR, 2021, pp. 6120–6130
2021
Later among the works it cites.
L. Le Mero, D. Yi, M. Dianati, and A. Mouzakitis, “A survey on imitation learning techniques for end-to-end autonomous vehicles,” IEEE Transactions on Intelligent Transportation Systems , 2022
2022
Later among the works it cites.
K. Zhou, Z. Liu, Y. Qiao, T. Xiang, and C. C. Loy, “Domain generalization: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2022
2022
Later among the works it cites.
Q. Li, Z. Peng, and B. Zhou, “Efficient learning of safe driving policy via human-ai copilot optimization,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
P. Florence and C. Lynch, “Decisiveness in Imitation Learning for Robots,” https://ai.googleblog.com/2021/11/decisiveness-in-imitation-learning-for.html , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Hu, G. Chen, Z. Li, and A. Knoll, “Robot policy improvement with natural evolution strategies for stable nonlinear dynamical system,” IEEE Transactions on Cybernetics , 2022
2022
Later among the works it cites.
T. Gangwani, Y. Zhou, and J. Peng, “Imitation learning from observations under transition model disparity,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
J. Zhang, R. Zhu, and E. Ohn-Bar, “Selfd: Self-learning large-scale driving policies from the web,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 17 316–17 326
2022
Later among the works it cites.
G.-H. Kim, S. Seo, J. Lee, W. Jeon, H. Hwang, H. Yang, and K.-E. Kim, “Demodice: Offline imitation learning with supplementary imperfect demonstrations,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
M. Beliaev, A. Shih, S. Ermon, D. Sadigh, and R. Pedarsani, “Imitation learning by estimating expertise of demonstrators,” in Proceedings of the 39th International Conference on Machine Learning . PMLR, 2022, pp. 1732–1748
2022
Later among the works it cites.
J. Chae, S. Han, W. Jung, M. Cho, S. Choi, and Y. Sung, “Robust imitation learning against variations in environment dynamics,” in Proceedings of the 39th International Conference on Machine Learning , vol. 162. PMLR, 2022, pp. 2828–2852
2022
Later among the works it cites.
K. Zakka, A. Zeng, P. Florence, J. Tompson, J. Bohg, and D. Dwibedi, “Xirl: Cross-embodiment inverse reinforcement learning,” in Conference on Robot Learning . PMLR, 2022, pp. 537–546
2022
Later among the works it cites.
A. Fickinger, S. Cohen, S. Russell, and B. Amos, “Cross-domain imitation learning via optimal transport,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
M. Yang, S. Levine, and O. Nachum, “Trail: Near-optimal imitation learning with suboptimal data,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
R. Ramrakhya, E. Undersander, D. Batra, and A. Das, “Habitat-web: Learning embodied object-search strategies from human demonstrations at scale,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 5173–5183
2022
Later among the works it cites.
A. Deka, C. Liu, and K. P. Sycara, “Arc-actor residual critic for adversarial imitation learning,” in Conference on Robot Learning . PMLR, 2023, pp. 1446–1456
2023
Closest in time.
R. Bhattacharyya, B. Wulfe, D. J. Phillips, A. Kuefler, J. Morton, R. Senanayake, and M. J. Kochenderfer, “Modeling human driving behavior through generative adversarial imitation learning,” IEEE Transactions on Intelligent Transportation Systems , vol. 24, no. 3, pp. 2874–2887, 2023
2023
Closest in time.
J. Van Den Berg, S. Miller, D. Duckworth, H. Hu, A. Wan, X.-Y. Fu, K. Goldberg, and P. Abbeel, “Superhuman performance of surgical tasks by robots using iterative learning from human-guided demonstrations,” in 2010 IEEE International Conference on Robotics and Automation . IEEE, 2010, pp. 2074–2081
2081
Closest in time.