Fetching the paper…
Reading the bibliography…
We address the longstanding challenge of producing flexible, realistic humanoid character controllers that can perform diverse whole-body tasks involving object interactions.
Animation of dynamic legged locomotion. In Proceedings of the 18th annual conference on Computer graphics and interactive techniques . 349–358
Marc H Raibert and Jessica K Hodgins. 1991 · 1991
Earlier work this paper cites.
Sensor-actuator networks. In Proceedings of the 20th annual conference on Computer graphics and interactive techniques . ACM, 335–342
Michiel Van de Panne and Eugene Fiume. 1993 · 1993
Earlier work this paper cites.
Animat vision: Active vision in artificial animals. In Proceedings of IEEE International Conference on Computer Vision . IEEE, 801–808
Demetri Terzopoulos and Tamer F Rabie. 1995 · 1995
Earlier work this paper cites.
Robot learning from demonstration. In ICML , Vol. 97. Citeseer, 12–20
Christopher G Atkeson and Stefan Schaal. 1997 · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Learning from demonstration. In Advances in neural information processing systems . 1040–1046
Stefan Schaal. 1997 · 1997
Earlier work this paper cites.
Behavior-based robotics
Ronald C Arkin, Ronald C Arkin, et al · 1998
Earlier work this paper cites.
Algorithms for Inverse Reinforcement Learning. In Proceedings of the Seventeenth International Conference on Machine Learning . Morgan Kaufmann Publishers Inc., 663–670
Andrew Y Ng and Stuart J Russell. 2000 · 2000
Earlier work this paper cites.
Composable Controllers for Physics-Based Character Animation. In Proceedings of the 28th Annual Conference on Computer Graphics and Interactive Techniques . Association for Computing Machinery
Petros Faloutsos, Michiel van de Panne, and Demetri Terzopoulos. 2001 · 2001
Earlier work this paper cites.
Understanding intelligence
Rolf Pfeifer and Christian Scheier. 2001 · 2001
Earlier work this paper cites.
Interactive motion generation from examples. In ACM Transactions on Graphics (TOG) , Vol. 21. ACM, 483–490
Okan Arikan and David A Forsyth. 2002 · 2002
Earlier work this paper cites.
Interactive control of avatars animated with human motion data. In ACM Transactions on Graphics (ToG) , Vol. 21. ACM, 491–500
Jehee Lee, Jinxiang Chai, Paul SA Reitsma, Jessica K Hodgins, and Nancy S Pollard. 2002 · 2002
Earlier work this paper cites.
Effective reinforcement learning for mobile robots. In Proceedings 2002 IEEE International Conference on Robotics and Automation (Cat. No. 02CH37292) , Vol. 4. IEEE, 3404–3410
William D Smart and L Pack Kaelbling. 2002 · 2002
Earlier work this paper cites.
Automated derivation of behavior vocabularies for autonomous humanoid motion. In Proceedings of the second international joint conference on Autonomous agents and multiagent systems . ACM, 225–232
Odest Chadwicke Jenkins and Maja J Mataric. 2003 · 2003
Earlier work this paper cites.
Computational approaches to motor learning by imitation
Stefan Schaal, Auke Ijspeert, and Aude Billard. 2003 · 2003
Earlier work this paper cites.
A model of the upper extremity for simulating musculoskeletal surgery and analyzing neuromuscular control
Katherine RS Holzbaur, Wendy M Murray, and Scott L Delp. 2005 · 2005
Earlier work this paper cites.
Synthesis of whole-body behaviors through hierarchical control of behavioral primitives
Luis Sentis and Oussama Khatib. 2005 · 2005
Earlier work this paper cites.
Reinforcement learning for imitating constrained reaching movements
Florent Guenter, Micha Hersch, Sylvain Calinon, and Aude Billard. 2007 · 2007
Earlier work this paper cites.
Modeling embodied visual behaviors
Nathan Sprague, Dana Ballard, and Al Robinson. 2007 · 2007
Earlier work this paper cites.
Simbicon: Simple biped locomotion control. In ACM Transactions on Graphics (TOG) , Vol. 26. ACM, 105
KangKang Yin, Kevin Loken, and Michiel Van de Panne. 2007 · 2007
Earlier work this paper cites.
Learning for control from multiple demonstrations. In Proceedings of the 25th international conference on Machine learning . ACM, 144–151
Adam Coates, Pieter Abbeel, and Andrew Y Ng. 2008 · 2008
Earlier work this paper cites.
Motion graphs. In ACM SIGGRAPH 2008 classes . ACM, 51
Lucas Kovar, Michael Gleicher, and Frédéric Pighin. 2008 · 2008
Earlier work this paper cites.
Reinforcement learning of motor skills with policy gradients
Jan Peters and Stefan Schaal. 2008 · 2008
Earlier work this paper cites.
Policy search for motor primitives in robotics. In Advances in neural information processing systems . 849–856
Jens Kober and Jan R Peters. 2009 · 2009
Earlier work this paper cites.
Generalized biped walking control. In ACM Transactions on Graphics (TOG) , Vol. 29. ACM, 130
Stelian Coros, Philippe Beaudoin, and Michiel Van de Panne. 2010 · 2010
Earlier work this paper cites.
Sampling-based contact-rich motion control
Libin Liu, KangKang Yin, Michiel van de Panne, Tianjia Shao, and Weiwei Xu. 2010 · 2010
Earlier work this paper cites.
Space-time planning with parameterized locomotion controllers
Sergey Levine, Yongjoon Lee, Vladlen Koltun, and Zoran Popović. 2011 · 2011
Cited alongside, same era.
Skill learning and task outcome prediction for manipulation. In 2011 IEEE International Conference on Robotics and Automation . IEEE, 3828–3834
Peter Pastor, Mrinal Kalakrishnan, Sachin Chitta, Evangelos Theodorou, and Stefan Schaal. 2011 · 2011
Cited alongside, same era.
Terrain runner: control, parameterization, composition, and planning for highly dynamic motions
Libin Liu, KangKang Yin, Michiel van de Panne, and Baining Guo. 2012 · 2012
Cited alongside, same era.
Discovery of complex behaviors through contact-invariant optimization
Igor Mordatch, Emanuel Todorov, and Zoran Popović. 2012 · 2012
Cited alongside, same era.
Learning and generalization of complex tasks from unstructured demonstrations. In 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 5239–5246
Scott Niekum, Sarah Osentoski, George Konidaris, and Andrew G Barto. 2012 · 2012
Maximum a posteriori policy optimisation. In International Conference on Learning Representations
Abbas Abdolmaleki, Jost Tobias Springenberg, Yuval Tassa, Remi Munos, Nicolas Heess, and Martin Riedmiller. 2018 · 2018
Later among the works it cites.
Learning dexterous in-hand manipulation
Marcin Andrychowicz, Bowen Baker, Maciek Chociej, Rafal Jozefowicz, Bob McGrew, Jakub Pachocki, Arthur Petron, Matthias Plappert, Glenn Powell, Alex Ray, et al · 2018
Later among the works it cites.
Physics-based motion capture imitation with deep reinforcement learning. In Proceedings of the 11th Annual International Conference on Motion, Interaction, and Games . 1–10
Nuttapong Chentanez, Matthias Müller, Miles Macklin, Viktor Makoviychuk, and Stefan Jeschke. 2018 · 2018
Later among the works it cites.
Learning manipulation skills from a single demonstration
Peter Englert and Marc Toussaint. 2018 · 2018
Later among the works it cites.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures. In International Conference on Machine Learning . 1406–1415
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
MuJoCo: A physics engine for model-based control. In 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 5026–5033
Emanuel Todorov, Tom Erez, and Yuval Tassa. 2012 · 2012
Cited alongside, same era.
Eyecatch: Simulating visuomotor coordination for object interception
Sang Hoon Yeo, Martin Lesmana, Debanga R Neog, and Dinesh K Pai. 2012 · 2012
Cited alongside, same era.
STAC: Simultaneous tracking and calibration. In 2013 13th IEEE-RAS International Conference on Humanoid Robots (Humanoids) . IEEE, 469–476
Tingfan Wu, Yuval Tassa, Vikash Kumar, Javier Movellan, and Emanuel Todorov. 2013 · 2013
Cited alongside, same era.
Catching objects in flight
Seungsu Kim, Ashwini Shukla, and Aude Billard. 2014 · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
A review of eye gaze in virtual agents, social robotics and hci: Behaviour generation, user interaction and perception. In Computer graphics forum , Vol. 34. Wiley Online Library, 299–326
Kerstin Ruhland, Christopher E Peters, Sean Andrist, Jeremy B Badler, Norman I Badler, Michael Gleicher, Bilge Mutlu, and Rachel McDonnell. 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Volodymyr Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, et al · 2018
Later among the works it cites.
Learning basketball dribbling skills using trajectory optimization and deep reinforcement learning
Libin Liu and Jessica Hodgins. 2018 · 2018
Later among the works it cites.
Deep learning of biomimetic sensorimotor control for biomechanical human animation
Masaki Nakada, Tao Zhou, Honglin Chen, Tomer Weiss, and Demetri Terzopoulos. 2018 · 2018
Later among the works it cites.
DeepMimic: Example-guided deep reinforcement learning of physics-based character skills
Xue Bin Peng, Pieter Abbeel, Sergey Levine, and Michiel van de Panne. 2018 · 2018
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, et al · 2018
Later among the works it cites.
Sim-to-real: Learning agile locomotion for quadruped robots
Jie Tan, Tingnan Zhang, Erwin Coumans, Atil Iscen, Yunfei Bai, Danijar Hafner, Steven Bohez, and Vincent Vanhoucke. 2018 · 2018
Later among the works it cites.
Yuval Tassa, Yotam Doron, Alistair Muldal, Tom Erez, Yazhe Li, Diego de Las Casas, David Budden, Abbas Abdolmaleki, Josh Merel, Andrew Lefrancq, Timothy Lillicrap, and Martin Riedmiller. 2018 · 2018
Later among the works it cites.
Reinforcement and imitation learning for diverse visuomotor skills. In Robotics: Science and Systems
Yuke Zhu, Ziyu Wang, Josh Merel, Andrei Rusu, Tom Erez, Serkan Cabi, Saran Tunyasuvunakool, János Kramár, Raia Hadsell, Nando de Freitas, and Nicolas Heess. 2018 · 2018
Later among the works it cites.
DReCon: data-driven responsive control of physics-based characters
Kevin Bergamin, Simon Clavet, Daniel Holden, and James Richard Forbes. 2019 · 2019
Closest in time.
Learning to sit: Synthesizing human-chair interactions via hierarchical control
Yu-Wei Chao, Jimei Yang, Weifeng Chen, and Jia Deng. 2019 · 2019
Closest in time.
Model Predictive Control with a Visuomotor System for Physics-based Character Animation
Haegwang Eom, Daseong Han, Joseph S Shin, and Junyong Noh. 2019 · 2019
Closest in time.
Learning agile and dynamic motor skills for legged robots
Jemin Hwangbo, Joonho Lee, Alexey Dosovitskiy, Dario Bellicoso, Vassilios Tsounis, Vladlen Koltun, and Marco Hutter. 2019 · 2019
Closest in time.
Data-driven Gaze Animation using Recurrent Neural Networks
Alex Klein, Zerrin Yumak, Arjen Beij, and A Frank van der Stappen. 2019 · 2019
Closest in time.
Scalable muscle-actuated human simulation and control
Seunghwan Lee, Moonseok Park, Kyoungmin Lee, and Jehee Lee. 2019 · 2019
Closest in time.
Hierarchical motor control in mammals and machines
Josh Merel, Matthew Botvinick, and Greg Wayne. 2019b · 2019
Closest in time.
Learning predict-and-simulate policies from unorganized human motion data
Soohwan Park, Hoseok Ryu, Seyoung Lee, Sunmin Lee, and Jehee Lee. 2019 · 2019
Closest in time.
MCP: Learning Composable Hierarchical Control with Multiplicative Compositional Policies
Xue Bin Peng, Michael Chang, Grace Zhang, Pieter Abbeel, and Sergey Levine. 2019 · 2019
Closest in time.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito. 2019 · 2019
Closest in time.
Iterative reinforcement learning based design of dynamic locomotion skills for cassie
Zhaoming Xie, Patrick Clary, Jeremy Dao, Pedro Morais, Jonathan Hurst, and Michiel van de Panne. 2019 · 2019
Closest in time.
TossingBot: Learning to Throw Arbitrary Objects with Residual Physics
Andy Zeng, Shuran Song, Johnny Lee, Alberto Rodriguez, and Thomas Funkhouser. 2019 · 2019
Closest in time.
V-MPO: On-Policy Maximum a Posteriori Policy Optimization for Discrete and Continuous Control. In International Conference on Learning Representations
H Francis Song, Abbas Abdolmaleki, Jost Tobias Springenberg, Aidan Clark, Hubert Soyer, Jack W Rae, Seb Noury, Arun Ahuja, Siqi Liu, Dhruva Tirumala, et al · 2020
Closest in time.