Fetching the paper…
Reading the bibliography…
As learning-based approaches progress towards automating robot controllers design, transferring learned policies to new domains with different dynamics (e.g.
M. Gautier and W. Khalil, “Exciting trajectories for the identification of base inertial parameters of robots,” The International Journal of Robotics Research , vol. 11, no. 4, pp. 362–375, 1992. [Online]. Available: https://doi.org/10.1177/027836499201100408
1992
Earlier work this paper cites.
A. Farchy, S. Barrett, P. MacAlpine, and P. Stone, “Humanoid robots learning to walk faster: From the real world to simulation and back,” in Proc. of 12th Int. Conf. on Autonomous Agents and Multiagent Systems (AAMAS) , May 2013. [Online]. Available: http://www.cs.utexas.edu/users/ai-lab?AAMAS13-Farchy
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
K. Ayusawa, G. Venture, and Y. Nakamura, “Identifiability and identification of inertial parameters using the underactuated base-link dynamics for legged multibody systems,” The International Journal of Robotics Research , vol. 33, no. 3, pp. 446–468, 2014. [Online]. Available: https://doi.org/10.1177/0278364913495932
2014
Earlier work this paper cites.
J. Tan, Z. Xie, B. Boots, and C. K. Liu, “Simulation-based design of dynamic controllers for humanoid balancing,” in 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2016, pp. 2729–2736
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in Advances in neural information processing systems , 2016, pp. 4565–4573
2016
Earlier work this paper cites.
K.-T. Yu, M. Bauza, N. Fazeli, and A. Rodriguez, “More than a million ways to be pushed. a high-fidelity experimental dataset of planar pushing,” in 2016 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 2016, pp. 30–37
2016
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 , 2017, pp. 2817–2826
2017
Earlier work this paper cites.
A. A. Rusu, M. Večerík, T. Rothörl, N. Heess, R. Pascanu, and R. Hadsell, “Sim-to-real robot learning from pixels with progressive nets,” ser. Proceedings of Machine Learning Research, S. Levine, V. Vanhoucke, and K. Goldberg, Eds., vol. 78. PMLR, 13–15 Nov 2017, pp. 262–270. [Online]. Available: http://proceedings.mlr.press/v78/rusu17a.html
2017
Earlier work this paper cites.
A. Rajeswaran, S. Ghotra, B. Ravindran, and S. Levine, “Epopt: Learning robust neural network policies using model ensembles,” 2017
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 23–30
2017
Earlier work this paper cites.
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige, S. Levine, and V. Vanhoucke, “Using simulation and domain adaptation to improve efficiency of deep robotic grasping,” 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation in robotics, games and machine learning,” 2017
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) , May 2018, pp. 1–8
2018
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” 2018
2018
Cited alongside, same era.
A. Nagabandi, I. Clavera, S. Liu, R. S. Fearing, P. Abbeel, S. Levine, and C. Finn, “Learning to adapt in dynamic, real-world environments through meta-reinforcement learning,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
N. Hansen, Y. Akimoto, and P. Baudis, “CMA-ES/pycma on Github,” Zenodo, DOI:10.5281/zenodo.2559634, Feb. 2019. [Online]. Available: https://doi.org/10.5281/zenodo.2559634
2019
Later among the works it cites.
X. B. Peng, E. Coumans, T. Zhang, T.-W. E. Lee, J. Tan, and S. Levine, “Learning agile robotic locomotion skills by imitating animals,” in Robotics: Science and Systems , 07 2020
2020
Later among the works it cites.
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros, “Unpaired image-to-image translation using cycle-consistent adversarial networks,” 2020
2020
Later among the works it cites.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, et al. , “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , vol. 39, no. 1, pp. 3–20, 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Zhu, A. Kimmel, K. E. Bekris, and A. Boularias, “Fast model identification via physics engines for data-efficient policy search,” 2018
2018
Cited alongside, same era.
I. Kostrikov, K. K. Agrawal, D. Dwibedi, S. Levine, and J. Tompson, “Discriminator-actor-critic: Addressing sample inefficiency and reward bias in adversarial imitation learning,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
Unitree, “Laikago: Let’s challenge new possibilities,” 2018. [Online]. Available: http://www.unitree.cc/
2018
Cited alongside, same era.
2019
Cited alongside, same era.
B. Mehta, M. Diaz, F. Golemo, C. J. Pal, and L. Paull, “Active domain randomization,” 2019
2019
Cited alongside, same era.
W. Yu, V. C. Kumar, G. Turk, and C. K. Liu, “Sim-to-real transfer for biped locomotion,” 2019
2019
Cited alongside, same era.
Y. Chebotar, A. Handa, V. Makoviychuk, M. Macklin, J. Issac, N. Ratliff, and D. Fox, “Closing the sim-to-real loop: Adapting simulation randomization with real world experience,” 2019
2019
Cited alongside, same era.
R. Jeong, J. Kay, F. Romano, T. Lampe, T. Rothorl, A. Abdolmaleki, T. Erez, Y. Tassa, and F. Nori, “Modelling generalized forces with reinforcement learning for sim-to-real transfer,” 2019
2019
Cited alongside, same era.
Y. Yang, K. Caluwaerts, A. Iscen, T. Zhang, J. Tan, and V. Sindhwani, “Data efficient reinforcement learning for legged robots,” in Conference on Robot Learning . PMLR, 2020, pp. 1–10
2020
Later among the works it cites.
S. Desai, H. Karnan, J. P. Hanna, G. Warnell, and P. Stone, “Stochastic grounded action transformation for robot learning in simulation,” in IEEE/RSJ International Conference on Intelligent Robots and Systems(IROS 2020) , October 2020
2020
Later among the works it cites.
Y. Song, A. Mavalankar, W. Sun, and S. Gao, “Provably efficient model-based policy adaptation,” 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
W. Yu, J. Tan, Y. Bai, E. Coumans, and S. Ha, “Learning fast adaptation with meta strategy optimization,” 2020
2020
Later among the works it cites.
B. Eysenbach, S. Asawa, S. Chaudhari, R. Salakhutdinov, and S. Levine, “Off-dynamics reinforcement learning: Training for transfer with domain classifiers,” 2020
2020
Later among the works it cites.
M. Jegorova, J. Smith, M. Mistry, and T. Hospedales, “Adversarial generation of informative trajectories for dynamics system identification,” 2020
2020
Later among the works it cites.
F. Muratore, C. Eilers, M. Gienger, and J. Peters, “Bayesian domain randomization for sim-to-real transfer,” 2020
2020
Later among the works it cites.
S. Desai, I. Durugkar, H. Karnan, G. Warnell, J. Hanna, and P. Stone, “An imitation from observation approach to transfer learning with dynamics mismatch,” in NeurIPS , 2020
2020
Later among the works it cites.
K. Morse, N. Das, Y. Lin, A. S. Wang, A. Rai, and F. Meier, “Learning state-dependent losses for inverse dynamics learning,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2020, pp. 5261–5268
2020
Later among the works it cites.
K. Rao, C. Harris, A. Irpan, S. Levine, J. Ibarz, and M. Khansari, “Rl-cyclegan: Reinforcement learning aware simulation-to-real,” 2020
2020
Later among the works it cites.