Fetching the paper…
Reading the bibliography…
Skill chaining is a promising approach for synthesizing complex behaviors by sequentially combining previously learned skills.
Temporal credit assignment in reinforcement learning
R. S. Sutton · 1984
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. A. Pomerleau · 1989
Earlier work this paper cites.
Towards compositional learning with dynamic neural networks
J. Schmidhuber · 1990
Earlier work this paper cites.
Learning from demonstration
S. Schaal · 1997
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
R. S. Sutton, D. Precup, and S. Singh · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng and S. J. Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
Skill discovery in continuous reinforcement learning domains using skill chaining
G. Konidaris and A. Barto · 2009
Earlier work this paper cites.
Learning and generalization of motor skills by learning from demonstration
P. Pastor, H. Hoffmann, T. Asfour, and S. Schaal · 2009
Earlier work this paper cites.
Movement templates for learning of hitting and batting
J. Kober, K. Mülling, O. Krömer, C. H. Lampert, B. Schölkopf, and J. Peters · 2010
Earlier work this paper cites.
Robot learning from demonstration by constructing skill trees
G. Konidaris, S. Kuindersma, R. Grupen, and A. Barto · 2012
Earlier work this paper cites.
Learning to select and generalize striking movements in robot table tennis
K. Mülling, J. Kober, O. Kroemer, and J. Peters · 2013
Earlier work this paper cites.
Incremental semantically grounded learning from demonstration
S. Niekum, S. Chitta, A. G. Barto, B. Marthi, and S. Osentoski · 2013
Earlier work this paper cites.
Ikeabot: An autonomous multi-robot coordinated furniture assembly system
R. A. Knepper, T. Layton, J. Romanishin, and D. Rus · 2013
Earlier work this paper cites.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Earlier work this paper cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz · 2015
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Earlier work this paper cites.
A framework for fine robotic assembly
F. Suárez-Ruiz and Q.-C. Pham · 2016
Earlier work this paper cites.
Deep spatial autoencoders for visuomotor learning
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel · 2016
Cited alongside, same era.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Cited alongside, same era.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
T. D. Kulkarni, K. Narasimhan, A. Saeedi, and J. Tenenbaum · 2016
Cited alongside, same era.
The option-critic architecture
P.-L. Bacon, J. Harb, and D. Precup · 2017
Cited alongside, same era.
Zero-shot task generalization with multi-task deep reinforcement learning
J. Oh, S. Singh, H. Lee, and P. Kohli · 2017
Cited alongside, same era.
Modular multitask reinforcement learning with policy sketches
J. Andreas, D. Klein, and S. Levine · 2017
Cited alongside, same era.
Reinforcement and imitation learning for diverse visuomotor skills
Y. Zhu, Z. Wang, J. Merel, A. Rusu, T. Erez, S. Cabi, S. Tunyasuvunakool, J. Kramár, R. Hadsell, N. de Freitas, and N. Heess · 2018
Later among the works it cites.
Learning deep visuomotor policies for dexterous hand manipulation
D. Jain, A. Li, S. Singhal, A. Rajeswaran, V. Kumar, and E. Todorov · 2019
Later among the works it cites.
Composing complex skills by learning transition policies
Y. Lee, S.-H. Sun, S. Somasundaram, E. S. Hu, and J. J. Lim · 2019
Later among the works it cites.
Mcp: Learning composable hierarchical control with multiplicative compositional policies
X. B. Peng, M. Chang, G. Zhang, P. Abbeel, and S. Levine · 2019
Later among the works it cites.
Discriminator-actor-critic: Addressing sample inefficiency and reward bias in adversarial imitation learning
I. Kostrikov, K. K. Agrawal, D. Dwibedi, S. Levine, and J. Tompson · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Cited alongside, same era.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine · 2018
Cited alongside, same era.
Learning to dress: Synthesizing human dressing motion via deep reinforcement learning
A. Clegg, W. Yu, J. Tan, C. K. Liu, and G. Turk · 2018
Cited alongside, same era.
Divide-and-conquer reinforcement learning
D. Ghosh, A. Singh, A. Rajeswaran, V. Kumar, and S. Levine · 2018
Cited alongside, same era.
Learning an embedding space for transferable robot skills
K. Hausman, J. T. Springenberg, Z. Wang, N. Heess, and M. Riedmiller · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Learning multi-level hierarchies with hindsight
A. Levy, G. Konidaris, R. Platt, and K. Saenko · 2019
Later among the works it cites.
Compile: Compositional imitation learning and execution
T. Kipf, Y. Li, H. Dai, V. Zambaldi, A. Sanchez-Gonzalez, E. Grefenstette, P. Kohli, and P. Battaglia · 2019
Later among the works it cites.
Hierarchical visuomotor control of humanoids
J. Merel, A. Ahuja, V. Pham, S. Tunyasuvunakool, S. Liu, D. Tirumala, N. Heess, and G. Wayne · 2019
Later among the works it cites.
Learning to coordinate manipulation skills via skill behavior diversification
Y. Lee, J. Yang, and J. J. Lim · 2020
Later among the works it cites.
Option discovery using deep skill chaining
A. Bagaria and G. Konidaris · 2020
Later among the works it cites.
Task-relevant adversarial imitation learning
K. Zolna, S. Reed, A. Novikov, S. G. Colmenarej, D. Budden, S. Cabi, M. Denil, N. de Freitas, and Z. Wang · 2020
Later among the works it cites.
Transferable task execution from pixels through deep planning domain learning
K. Kase, C. Paxton, H. Mazhar, T. Ogata, and D. Fox · 2020
Later among the works it cites.
Accelerating reinforcement learning with learned skill priors
K. Pertsch, Y. Lee, and J. J. Lim · 2020
Later among the works it cites.
What matters for on-policy deep actor-critic methods? a large-scale study
M. Andrychowicz, A. Raichuk, P. Stańczyk, M. Orsini, S. Girgin, R. Marinier, L. Hussenot, M. Geist, O. Pietquin, M. Michalski, S. Gelly, and O. Bachem · 2021
Closest in time.
IKEA furniture assembly environment for long-horizon complex manipulation tasks
Y. Lee, E. S. Hu, and J. J. Lim · 2021
Closest in time.
Learning task decomposition with ordered memory policy network
Y. Lu, Y. Shen, S. Zhou, A. Courville, J. B. Tenenbaum, and C. Gan · 2021
Closest in time.
Learning geometric reasoning and control for long-horizon tasks from visual input
D. Driess, J.-S. Ha, R. Tedrake, and M. Toussaint · 2021
Closest in time.
Amp: Adversarial motion priors for stylized physics-based character control
X. B. Peng, Z. Ma, P. Abbeel, S. Levine, and A. Kanazawa · 2021
Closest in time.