Fetching the paper…
Reading the bibliography…
Behavioural cloning uses a dataset of demonstrations to learn a behavioural policy.
S. Schaal, “Learning from demonstration,” in Advances in Neural Information Processing Systems
1996
Earlier work this paper cites.
A. Y. Ng and S. J. Russell, “Algorithms for inverse reinforcement learning,” in Proceedings of the Seventeenth International Conference on Machine Learning
2000
Earlier work this paper cites.
L. V. D. Maaten and G. Hinton, “Visualizing data using t-sne,” Journal of Machine Learning Research
2008
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” 2016
2016
Earlier work this paper cites.
The MIT Press, second ed., 2018
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction · 2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Minka, R. Cleven, and Y. Zaykov, “Trueskill 2: An improved bayesian skill rating system,” Tech. Rep. MSR-TR-2018-8, Microsoft, March 2018
2018
Earlier work this paper cites.
M. Schilling and A. Melnik, “An approach to hierarchical deep reinforcement learning for a decentralized walking control architecture,” in Biologically Inspired Cognitive Architectures 2018: Proceedings of the Ninth Annual Meeting of the BICA Society
2019
Earlier work this paper cites.
S. K. Saksena, B. Navaneethkrishnan, S. Hegde, P. Raja, and R. M. Vishwanath, “Towards behavioural cloning for autonomous driving,” 2019
2019
Earlier work this paper cites.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, J. Oh, D. Horgan, M. Kroiss, I. Danihelka, A. Huang, L. Sifre, T. Cai, J. P. Agapiou, M. Jaderberg, A. S. Vezhnevets, R. Leblond, T. Pohlen, V. Dalibard, D. Budden, Y. Sulsky, J. Molloy, T. L. Paine, C. Gulcehre, Z. Wang, T. Pfaff, Y. Wu, R. Ring, D. Yogatama, D. Wünsch, K. McKinney, O. Smith, T. Schaul, T. Lillicrap, K. Kavukcuoglu, D. Hassabis, C. Apps, and D. Silver, “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature
2019
Earlier work this paper cites.
P. de Haan, D. Jayaraman, and S. Levine, “Causal confusion in imitation learning,” vol. 32, 2019
2019
Cited alongside, same era.
Penguin, 2019
S. Russell, Human Compatible · 2019
Cited alongside, same era.
W. H. Guss, B. Houghton, N. Topin, P. Wang, C. Codel, M. Veloso, and R. Salakhutdinov, “MinerL: A large-scale dataset of minecraft demonstrations,” in IJCAI International Joint Conference on Artificial Intelligence
2019
Cited alongside, same era.
N. Bach, A. Melnik, M. Schilling, T. Korthals, and H. Ritter, “Learn to move through a combination of policy gradient algorithms: Ddpg, d4pg, and td3,” in International Conference on Machine Learning, Optimization, and Data Science
2020
Cited alongside, same era.
A. Kanervisto, J. Karttunen, and V. Hautamäki, “Playing minecraft with behavioural cloning,” CoRR
2020
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al
2022
Later among the works it cites.
B. Baker, I. Akkaya, P. Zhokov, J. Huizinga, J. Tang, A. Ecoffet, B. Houghton, R. Sampedro, and J. Clune, “Video pretraining (vpt): Learning to act by watching unlabeled online videos,” Advances in Neural Information Processing Systems
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A. Kanervisto, J. Pussinen, and V. Hautamaki, “Benchmarking End-to-End Behavioural Cloning on Video Games,” in IEEE Conference on Computatonal Intelligence and Games, CIG
2020
Cited alongside, same era.
M. Schilling, A. Melnik, F. W. Ohl, H. J. Ritter, and B. Hammer, “Decentralized control and local information for robust and adaptive decentralized deep reinforcement learning,” Neural Networks
2021
Cited alongside, same era.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al
2021
Cited alongside, same era.
2021
Cited alongside, same era.
T. V. Samak, C. V. Samak, and S. Kandhasamy, “Robust behavioral cloning for autonomous vehicles using end-to-end imitation learning,” SAE International Journal of Connected and Automated Vehicles
2021
Cited alongside, same era.
L. Fan, G. Wang, Y. Jiang, A. Mandlekar, Y. Yang, H. Zhu, A. Tang, D.-A. Huang, Y. Zhu, and A. Anandkumar, “Minedojo: Building open-ended embodied agents with internet-scale knowledge,” in Thirty-sixth Conference on Neural Information Processing Systems Datasets and Benchmarks Track
2022
Later among the works it cites.
A. Kanervisto, S. Milani, K. Ramanauskas, N. Topin, Z. Lin, J. Li, J. Shi, D. Ye, Q. Fu, W. Yang, W. Hong, Z. Huang, H. Chen, G. Zeng, Y. Lin, V. Micheli, E. Alonso, F. Fleuret, A. Nikulin, Y. Belousov, O. Svidchenko, and A. Shpilman, “Minerl diamond 2021 competition: Overview, results, and lessons learned,” 2022
2022
Later among the works it cites.
2023
Closest in time.
2023
Closest in time.
S. Milani, A. Kanervisto, K. Ramanauskas, S. Schulhoff, , B. Houghton, S. Mohanty, B. Galbraith, K. Chen, Y. Song, T. Zhou, B. Yu, H. Liu, K. Guan, Y. Hu, T. Lv, F. Malato, F. Leopold, A. Raut, V. Hautamäki, A. Melnik, S. Ishida, J. F. Henriques, R. Klassert, W. Laurito, E. Novoseller, V. G. Goecks, N. Waytowich, D. Watkins, J. Miller, and R. Shah, “A retrospective of the minerl basalt 2022 competition]Towards Solving Fuzzy Tasks with Human Feedback: A Retrospective of the MineRL BASALT 2022 Competition,” arXiv
2023
Closest in time.
Accessed: 2023-05-14
“Video-pre-training.” https://github.com/openai/Video-Pre-Training/tree/main/lib · 2023
Closest in time.