Fetching the paper…
Reading the bibliography…
Sequence models in reinforcement learning require task knowledge to estimate the task policy.
Schaal, S. Is imitation learning the route to humanoid robots?. Trends In Cognitive Sciences
1999
Earlier work this paper cites.
Billard, A., Calinon, S., Dillmann, R. & Schaal, S. Survey: Robot programming by demonstration. (Springrer,2008)
2008
Earlier work this paper cites.
Argall, B., Chernova, S., Veloso, M. & Browning, B. A survey of robot learning from demonstration. Robotics And Autonomous Systems
2009
Earlier work this paper cites.
Ross, S., Gordon, G. & Bagnell, J. A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning. (2011)
2011
Earlier work this paper cites.
Osa, T., Harada, K., Sugita, N. & Mitsuishi, M. Trajectory planning under different initial conditions for surgical task automation by learning from demonstration. 2014 IEEE International Conference On Robotics And Automation (ICRA)
2014
Earlier work this paper cites.
Brys, T., Harutyunyan, A., Suay, H., Chernova, S., Taylor, M. & Nowé, A. Reinforcement Learning from Demonstration through Shaping. Proceedings Of The 24th International Conference On Artificial Intelligence
2015
Earlier work this paper cites.
Brys, T., Harutyunyan, A., Taylor, M. & Nowé, A. Policy Transfer using Reward Shaping.. AAMAS
2015
Earlier work this paper cites.
Suay, H., Brys, T., Taylor, M. & Chernova, S. Learning from Demonstration for Shaping through Inverse Reinforcement Learning. AAMAS
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Hussein, A., Gaber, M., Elyan, E. & Jayne, C. Imitation Learning: A Survey of Learning Methods. ACM Comput. Surv
2017
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A., Kaiser, Ł. & Polosukhin, I. Attention is all you need. Advances In Neural Information Processing Systems
2017
Earlier work this paper cites.
Levy, A., Platt, R. & Saenko, K. Hierarchical actor-critic. ArXiv Preprint ArXiv:1712.00948
2017
Earlier work this paper cites.
Zhu, Z. & Hu, H. Robot Learning from Demonstration in Robotic Assembly: A Survey. Robotics
2018
Cited alongside, same era.
Sutton, R. & Barto, A. Reinforcement Learning: An Introduction. (A Bradford Book,2018)
2018
Cited alongside, same era.
Sieb, M. & Fragkiadaki, K. Data Dreaming for Object Detection: Learning Object-Centric State Representations for Visual Imitation. 2018 IEEE-RAS 18th International Conference On Humanoid Robots (Humanoids)
2018
Cited alongside, same era.
Nachum, O., Gu, S., Lee, H. & Levine, S. Data-efficient hierarchical reinforcement learning. Advances In Neural Information Processing Systems
2018
Cited alongside, same era.
2019
2020
Later among the works it cites.
Jurgenson, T., Avner, O., Groshev, E. & Tamar, A. Sub-Goal Trees a Framework for Goal-Based Reinforcement Learning. International Conference On Machine Learning
2020
Later among the works it cites.
Pertsch, K., Rybkin, O., Ebert, F., Zhou, S., Jayaraman, D., Finn, C. & Levine, S. Long-horizon visual planning with goal-conditioned hierarchical predictors. Advances In Neural Information Processing Systems
2020
Later among the works it cites.
Hansen, N., Su, H. & Wang, X. Stabilizing deep q-learning with convnets and vision transformers under data augmentation. Advances In Neural Information Processing Systems
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Krishnan, S., Garg, A., Liaw, R., Thananjeyan, B., Miller, L., Pokorny, F. & Goldberg, K. SWIRL: A sequential windowed inverse reinforcement learning algorithm for robot tasks with delayed rewards. The International Journal Of Robotics Research
2019
Cited alongside, same era.
Paul, S., Vanbaar, J. & Roy-Chowdhury, A. Learning from trajectories via subgoal discovery. Advances In Neural Information Processing Systems
2019
Cited alongside, same era.
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D. & Sutskever, I. Language Models are Unsupervised Multitask Learners. (2019)
2019
Cited alongside, same era.
Ravichandar, H., Polydoros, A., Chernova, S. & Billard, A. Recent Advances in Robot Learning from Demonstration. Annual Review Of Control, Robotics, And Autonomous Systems
2020
Cited alongside, same era.
Fu, J., Kumar, A., Nachum, O., Tucker, G. & Levine, S. D4RL: Datasets for Deep Data-Driven Reinforcement Learning. (2020)
2020
Cited alongside, same era.
Mandlekar, A., Ramos, F., Boots, B., Savarese, S., Fei-Fei, L., Garg, A. & Fox, D. Iris: Implicit reinforcement without interaction at scale for learning control from offline robot manipulation data. 2020 IEEE International Conference On Robotics And Automation (ICRA)
2020
Cited alongside, same era.
Pertsch, K., Lee, Y. & Lim, J. Accelerating Reinforcement Learning with Learned Skill Priors. Conference On Robot Learning (CoRL)
2020
Cited alongside, same era.
Yu, T., Kumar, A., Chebotar, Y., Hausman, K., Levine, S. & Finn, C. Conservative data sharing for multi-task offline reinforcement learning. Advances In Neural Information Processing Systems
2021
Later among the works it cites.
Chane-Sane, E., Schmid, C. & Laptev, I. Goal-conditioned reinforcement learning with imagined subgoals. International Conference On Machine Learning
2021
Later among the works it cites.
Chen, L., Lu, K., Rajeswaran, A., Lee, K., Grover, A., Laskin, M., Abbeel, P., Srinivas, A. & Mordatch, I. Decision transformer: Reinforcement learning via sequence modeling. Advances In Neural Information Processing Systems
2021
Later among the works it cites.
2021
Later among the works it cites.
Janner, M., Li, Q. & Levine, S. Offline Reinforcement Learning as One Big Sequence Modeling Problem. Advances In Neural Information Processing Systems
2021
Later among the works it cites.