Fetching the paper…
Reading the bibliography…
We study a novel architecture and training procedure for locomotion tasks.
Die Schreitbewegungen der Neugeborenen [The walking movements of newborns]
A Peiper · 1929
Earlier work this paper cites.
The co-ordination and regulation of movements
N A Bernstein · 1967
Earlier work this paper cites.
A robust layered control system for a mobile robot
R A Brooks · 1986
Earlier work this paper cites.
Feudal reinforcement learning
P Dayan and G E Hinton · 1993
Earlier work this paper cites.
A hierarchical foundation for models of sensorimotor control
G E Loeb, I E Brown, and E J Cheng · 1999
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
R S Sutton, D Precup, and S Singh · 1999
Earlier work this paper cites.
Principles of neural science
E R Kandel, J H Schwartz, T M Jessell, et al · 2000
Earlier work this paper cites.
Motor learning through the combination of primitives
F A Mussa-Ivaldi and E Bizzi · 2000
Earlier work this paper cites.
State abstraction for programmable reinforcement learning agents
D Andre and S J Russell · 2002
Earlier work this paper cites.
Learning attractor landscapes for learning motor primitives
A J Ijspeert, J Nakanishi, and S Schaal · 2002
Earlier work this paper cites.
Relativized options: Choosing the right transformation
B Ravindran and A G Barto · 2003
Cited alongside, same era.
From task parameters to motor synergies: A hierarchical framework for approximately optimal control of redundant manipulators
E Todorov, W Li, and X Pan · 2005
Cited alongside, same era.
Effective control knowledge transfer through learning skill and representation hierarchies
M Asadi and M Huber · 2007
Cited alongside, same era.
Efficient skill learning using abstraction selection
G Konidaris and A G Barto · 2009
Cited alongside, same era.
Locomotor primitives in newborn babies and their development
N Dominici, Y P Ivanenko, G Cappellini, A d’Avella, V Mondì, M Cicchese, A Fabiano, T Silei, A Di Paolo, C Giannini, et al · 2011
Cited alongside, same era.
Motor primitive discovery
P S Thomas and A G Barto · 2012
Hierarchical control using networks trained with higher-level forward models
G Wayne and LF Abbott · 2014
Later among the works it cites.
Learning continuous control policies by stochastic value gradients
N Heess, G Wayne, D Silver, T P Lillicrap, T Erez, and Y Tassa · 2015
Later among the works it cites.
End-to-end training of deep visuomotor policies
S Levine, C Finn, T Darrell, and P Abbeel · 2015
Later among the works it cites.
Continuous control with deep reinforcement learning
T P Lillicrap, J J Hunt, A Pritzel, N Heess, T Erez, Y Tassa, D Silver, and D Wierstra · 2015
Later among the works it cites.
High-dimensional continuous control using generalized advantage estimation
J Schulman, P Moritz, S Levine, M Jordan, and P Abbeel · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Auto-encoding variational Bayes
D P Kingma and M Welling · 2013
Cited alongside, same era.
Learning parameterized motor skills on a humanoid robot
B C Da Silva, G Baldassarre, G Konidaris, and A Barto · 2014
Cited alongside, same era.
Stochastic Backpropagation and Approximate Inference in Deep Generative Models
D Rezende, S Mohamed, and D Wierstra · 2014
Cited alongside, same era.
On the comparative study of disease of the nervous system
J Hughlings Jackson
Cited in the paper.
Later among the works it cites.
A neural circuitry that emphasizes spinal feedback generates diverse behaviours of human locomotion
S Song and H Geyer · 2015
Later among the works it cites.
Continuous deep q-learning with model-based acceleration
S Gu, T P Lillicrap, I Sutskever, and S Levine · 2016
Closest in time.
Asynchronous methods for deep reinforcement learning
V Mnih, A P Badia, M Mirza, A Graves, T P Lillicrap, T Harley, D Silver, and K Kavukcuoglu · 2016
Closest in time.
Strategic attentive writer for learning macro-actions
Alexander Vezhnevets, Volodymyr Mnih, John Agapiou, Simon Osindero, Alex Graves, Oriol Vinyals, and Koray Kavukcuoglu · 2016
Closest in time.