Fetching the paper…
Reading the bibliography…
The physical design of a robot and the policy that controls its motion are inherently coupled, and should be determined according to the task and environment.
T. McGeer, “Passive dynamic walking,”
1990
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,”
1992
Earlier work this paper cites.
J.-H. Park and H. Asada, “Concurrent design optimization of mechanical structure and control for high speed robots,”
1994
Earlier work this paper cites.
K. Sims, “Evolving virtual creatures,” in
1994
Earlier work this paper cites.
A. C. Pil and H. H. Asada, “Integrated structure/control design of mechatronic systems using a recursive experimental optimization method,”
1996
Earlier work this paper cites.
A. Goswami, B. Thuilot, and B. Espiau, “A study of the passive gait of a compass-like biped robot: Symmetry and chaos,”
1998
Earlier work this paper cites.
D. R. Jones, M. Schonlau, and W. J. Welch, “Efficient global optimization of expensive black-box functions,”
1998
Earlier work this paper cites.
H. Lipson and J. B. Pollack, “Automatic design and manufacture of robotic lifeforms,”
2000
Earlier work this paper cites.
R. S. Sutton, D. A. McAllester, S. P. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in
2000
Earlier work this paper cites.
S. H. Collins, M. Wisse, and A. Ruina, “A three-dimensional passive-dynamic walking robot with two legs and knees,”
2001
Earlier work this paper cites.
C. Paul and J. C. Bongard, “The road less travelled: Morphology in the optimization of biped robot locomotion,” in
2001
Earlier work this paper cites.
J. A. Reyer and P. Y. Papalambros, “Combined optimal design and control with application to an electric DC motor,”
2002
Earlier work this paper cites.
M. Riedmiller, “Neural fitted Q iteration-first experiences with a data efficient neural reinforcement learning method,” in
2005
Earlier work this paper cites.
P. Wawrzyński, “Real-time reinforcement learning by sequential actor-critics and experience replay,”
2009
Earlier work this paper cites.
F. Sehnke, C. Osendorfer, T. Rückstieß, A. Graves, J. Peters, and J. Schmidhuber, “Parameter-exploring policy gradients,”
2010
Earlier work this paper cites.
J. Bongard, “Morphological change in machines accelerates the evolution of robust behavior,”
2011
Earlier work this paper cites.
I. Mordatch, E. Todorov, and Z. Popović, “Discovery of complex behaviors through contact-invariant optimization,”
2012
Cited alongside, same era.
Y. Tassa, T. Erez, and E. Todorov, “Synthesis and stabilization of complex behaviors through online trajectory optimization,” in
2012
Cited alongside, same era.
M. G. Villarreal-Cervantes, C. A. Cruz-Villar, J. Alvarez-Gallegos, and E. A. Portilla-Flores, “Robust structure-control design approach for mechatronic systems,”
2013
Cited alongside, same era.
K. Wampler, J. Popović, and Z. Popović, “Animal locomotion controllers from scratch,” in
2013
Cited alongside, same era.
S. Coros, B. Thomaszewski, G. Noris, S. Sueda, M. Forberg, R. W. Sumner, W. Matusik, and B. Bickel, “Computational design of mechanical characters,”
2013
Cited alongside, same era.
M. Posa, S. Kuindersma, and R. Tedrake, “Optimization and stabilization of trajectories for constrained dynamical systems,” in
2016
Later among the works it cites.
B. Griffin and J. Grizzle, “Nonholonomic virtual constraints and gait optimization for robust walking control,”
2016
Later among the works it cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,”
2016
Later among the works it cites.
X. B. Peng, G. Berseth, and M. Van de Panne, “Terrain-adaptive locomotion skills using deep reinforcement learning,”
2016
Later among the works it cites.
A. Chakrabarti, “Learning sensor multiplexing design through backpropagation,” in
2016
Later among the works it cites.
T. G. Authors, “GPyOpt: A Bayesian optimization framework in Python,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
K. M. Digumarti, C. Gehring, S. Coros, J. Hwangbo, and R. Siegwart, “Concurrent optimization of mechanical design and locomotion control of a legged robot,” in
2014
Cited alongside, same era.
S. Levine and P. Abbeel, “Learning neural network policies with guided policy search under unknown dynamics,” in
2014
Cited alongside, same era.
I. Mordatch and E. Todorov, “Combining the benefits of function approximation and trajectory optimization,” in
2014
Cited alongside, same era.
A. M. Mehta, J. DelPreto, B. Shaya, and D. Rus, “Cogeneration of mechanical, electrical, and software designs for printable robots from structural specifications,” in
2014
Cited alongside, same era.
I. Mordatch, K. Lowrey, G. Andrew, Z. Popović, and E. V. Todorov, “Interactive control of diverse complex characters with neural networks,” in
2015
Cited alongside, same era.
J. Schulman, S. Levine, P. Moritz, M. I. Jordan, and P. Abbeel, “Trust region policy optimization,”
2015
Cited alongside, same era.
2016
Later among the works it cites.
S. Ha, S. Coros, A. Alspach, J. Kim, and K. Yamane, “Joint optimization of robot design and motion parameters using the implicit function theorem,” in
2017
Later among the works it cites.
A. Spielberg, B. Araki, C. Sung, R. Tedrake, and D. Rus, “Functional co-optimization of articulated robots,” in
2017
Later among the works it cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in
2017
Later among the works it cites.
A. Rajeswaran, K. Lowrey, E. V. Todorov, and S. M. Kakade, “Towards generalization and simplicity in continuous control,” in
2017
Later among the works it cites.
A. Censi, “A class of co-design problems with cyclic constraints and their solution,”
2017
Later among the works it cites.
C. Schaff, D. Yunis, A. Chakrabarti, and M. R. Walter, “Jointly optimizing placement and inference for beacon-based localization,” in
2017
Later among the works it cites.
2017
Later among the works it cites.
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,”
2018
Closest in time.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” in
2018
Closest in time.