Fetching the paper…
Reading the bibliography…
In the NIPS 2017 Learning to Run challenge, participants were tasked with building a controller for a musculoskeletal model to make it run as fast as possible through an obstacle course.
Physical review 36
Uhlenbeck, G.E., Ornstein, L.S.: On the theory of the brownian motion · 1930
Earlier work this paper cites.
chap. Learning Internal Representations by Error Propagation, pp. 318–362. MIT Press, Cambridge, MA, USA (1986)
Rumelhart, D.E., Hinton, G.E., Williams, R.J.: Parallel distributed processing: Explorations in the microstructure of cognition, vol. 1 · 1986
Earlier work this paper cites.
IFAC Proceedings Volumes 28
Bratko, I., Urbančič, T., Sammut, C.: Behavioural cloning: Phenomena, results and problems · 1995
Earlier work this paper cites.
MIT Press, Cambridge, MA, USA (1997)
Dorigo, M., Colombetti, M.: Robot Shaping: An Experiment in Behavior Engineering · 1997
Earlier work this paper cites.
Wiering, M., Schmidhuber, J.: HQ-learning · 1997
Earlier work this paper cites.
Artificial Intelligence 112
Sutton, R.S., Precup, D., Singh, S.: Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning · 1999
Earlier work this paper cites.
Multiple classifier systems 1857
Dietterich, T.G., et al.: Ensemble methods in machine learning · 2000
Earlier work this paper cites.
In: H. Kimura, K. Tsuchiya, A. Ishiguro, H. Witte (eds.) Adaptive Motion of Animals and Machines, pp. 261–280. Springer Tokyo, Tokyo (2006)
Schaal, S.: Dynamic movement primitives -a framework for motor control in humans and humanoid robotics · 2006
Earlier work this paper cites.
Neural Computation 25
Ijspeert, A., Nakanishi, J., Pastor, P., Hoffmann, H., Schaal, S.: Dynamical movement primitives: Learning attractor models for motor behaviors · 2013
Earlier work this paper cites.
IEEE transactions on pattern analysis and machine intelligence 35
Ji, S., Xu, W., Yang, M., Yu, K.: 3d convolutional neural networks for human action recognition · 2013
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Graves, A., Antonoglou, I., Wierstra, D., Riedmiller, M.A.: Playing atari with deep reinforcement learning · 2013
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A method for stochastic optimization · 2014
Earlier work this paper cites.
In: Proceedings of the 31st International Conference on Machine Learning (ICML-14), pp. 387–395 (2014)
Silver, D., Lever, G., Heess, N., Degris, T., Wierstra, D., Riedmiller, M.: Deterministic policy gradient algorithms · 2014
Earlier work this paper cites.
URL https://www.tensorflow.org/
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G.S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M., Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., Zheng, X.: TensorFlow: Large-scale machine learning on heterogeneous systems (2015) · 2015
Cited alongside, same era.
arXiv preprint arXiv:1511.07289 (2015)
Clevert, D.A., Unterthiner, T., Hochreiter, S.: Fast and accurate deep network learning by exponential linear units (elus) · 2015
Cited alongside, same era.
arXiv preprint arXiv:1509.02971 (2015)
Lillicrap, T.P., Hunt, J.J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., Wierstra, D.: Continuous control with deep reinforcement learning · 2015
Cited alongside, same era.
Nature 518
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., et al.: Human-level control through deep reinforcement learning · 2015
Cited alongside, same era.
arXiv preprint arXiv:1710.02298 (2017)
Hessel, M., Modayil, J., Van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., Silver, D.: Rainbow: Combining improvements in deep reinforcement learning · 2017
Later among the works it cites.
arXiv preprint arXiv:1706.02515 (2017)
Klambauer, G., Unterthiner, T., Mayr, A., Hochreiter, S.: Self-normalizing neural networks · 2017
Later among the works it cites.
ArXiv e-prints (2017)
Pavlov, M., Kolesnikov, S., Plis, S.M.: Run, skeleton, run: skeletal model in a physics-based simulation · 2017
Later among the works it cites.
arXiv preprint arXiv:1706.01905 (2) (2017)
Plappert, M., Houthooft, R., Dhariwal, P., Sidor, S., Chen, R.Y., Chen, X., Asfour, T., Abbeel, P., Andrychowicz, M.: Parameter space noise for exploration · 2017
Later among the works it cites.
ArXiv e-prints (2017)
Salimans, T., Ho, J., Chen, X., Sidor, S., Sutskever, I.: Evolution Strategies as a Scalable Alternative to Reinforcement Learning · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
arXiv preprint arXiv:1511.05952 (2015)
Schaul, T., Quan, J., Antonoglou, I., Silver, D.: Prioritized experience replay · 2015
Cited alongside, same era.
arXiv preprint arXiv:1607.06450 (2016)
Ba, J.L., Kiros, J.R., Hinton, G.E.: Layer normalization · 2016
Cited alongside, same era.
Heess, N., Wayne, G., Tassa, Y., Lillicrap, T.P., Riedmiller, M.A., Silver, D.: Learning and transfer of modulated locomotor controllers · 2016
Cited alongside, same era.
In: Advances in Neural Information Processing Systems, pp. 4026–4034 (2016)
Osband, I., Blundell, C., Pritzel, A., Van Roy, B.: Deep exploration via bootstrapped dqn · 2016
Cited alongside, same era.
https://github.com/matthiasplappert/keras-rl (2016)
Plappert, M.: keras-rl · 2016
Cited alongside, same era.
Rusu, A.A., Rabinowitz, N.C., Desjardins, G., Soyer, H., Kirkpatrick, J., Kavukcuoglu, K., Pascanu, R., Hadsell, R.: Progressive neural networks · 2016
Cited alongside, same era.
https://github.com/openai/baselines (2017)
Dhariwal, P., Hesse, C., Plappert, M., Radford, A., Schulman, J., Sidor, S., Wu, Y.: OpenAI Baselines · 2017
Cited alongside, same era.
arXiv preprint arXiv:1707.02286 (2017)
Heess, N., Sriram, S., Lemmon, J., Merel, J., Wayne, G., Tassa, Y., Erez, T., Wang, Z., Eslami, A., Riedmiller, M., et al.: Emergence of locomotion behaviours in rich environments · 2017
Cited alongside, same era.
Later among the works it cites.
https://github.com/openai/evolution-strategies-starter (2017)
Salimans, T., Ho, J., Chen, X., Sidor, S., Sutskever, I.: Starter code for evolution strategies · 2017
Later among the works it cites.
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., Klimov, O.: Proximal policy optimization algorithms · 2017
Later among the works it cites.
ArXiv e-prints (2017)
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., Klimov, O.: Proximal Policy Optimization Algorithms · 2017
Later among the works it cites.
https://github.com/AdamStelmaszczyk/learning2run (2017)
Stelmaszczyk, A., Jarosik, P.: Our NIPS 2017: Learning to Run source code · 2017
Later among the works it cites.
International Conference on Learning Representations (2018)
Anonymous: Distributional policy gradients · 2018
Closest in time.
In: S. Escalera, M. Weimer (eds.) NIPS 2017 Competition Book. Springer, Springer (2018)
Jaśkowski, W., Lykkebø, O.R., Toklu, N.E., Trifterer, F., Buk, Z., Koutník, J., Gomez, F.: Reinforcement Learning to Run… Fast · 2018
Closest in time.
In: S. Escalera, M. Weimer (eds.) NIPS 2017 Competition Book. Springer, Springer (2018)
Kidziński, Ł., Sharada, M.P., Ong, C., Hicks, J., Francis, S., Levine, S., Salathé, M., Delp, S.: Learning to run challenge: Synthesizing physiologically accurate motion using deep reinforcement learning · 2018
Closest in time.