Fetching the paper…
Reading the bibliography…
Training agents to autonomously learn how to use anthropomorphic robotic hands has the potential to lead to systems capable of performing a multitude of complex manipulation tasks in unstructured and uncertain environments.
Contact-invariant optimization for hand manipulation
Mordatch, I., Popović, Z., and Todorov, E · 2012
Earlier work this paper cites.
Real-time behaviour synthesis for dynamic hand-manipulation
Kumar, V., Tassa, Y., Erez, T., and Todorov, E · 2014
Earlier work this paper cites.
Model predictive path integral control using covariance variable importance sampling
Williams, G., Aldrich, A., and Theodorou, E · 2015
Earlier work this paper cites.
Openai gym, 2016
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Earlier work this paper cites.
Learning dexterous manipulation for a soft robotic hand from human demonstrations
Gupta, A., Eppner, C., Levine, S., and Abbeel, P · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
Ho, J. and Ermon, S · 2016
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Lillicrap, T. P., Hunt, J. J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., and Wierstra, D · 2016
Earlier work this paper cites.
Hindsight experience replay
Andrychowicz, M., Wolski, F., Ray, A., Schneider, J., Fong, R., Welinder, P., McGrew, B., Tobin, J., Pieter Abbeel, O., and Zaremba, W · 2017
Earlier work this paper cites.
Learning from demonstrations for real world reinforcement learning
Hester, T., Vecerík, M., Pietquin, O., Lanctot, M., Schaul, T., Piot, B., Sendonaris, A., Dulac-Arnold, G., Osband, I., Agapiou, J., Leibo, J. Z., and Gruslys, A · 2017
Cited alongside, same era.
Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards
Vecerík, M., Hester, T., Scholz, J., Wang, F., Pietquin, O., Piot, B., Heess, N., Rothörl, T., Lampe, T., and Riedmiller, M. A · 2017
Cited alongside, same era.
Learning dexterous in-hand manipulation
Andrychowicz, M., Baker, B., Chociej, M., Józefowicz, R., McGrew, B., Pachocki, J. W., Pachocki, J., Petron, A., Plappert, M., Powell, G., Ray, A., Schneider, J., Sidor, S., Tobin, J., Welinder, P., Weng, L., and Zaremba, W · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Fujimoto, S., Van Hoof, H., and Meger, D · 2018
Cited alongside, same era.
Goal-conditioned imitation learning
Ding, Y., Florensa, C., Abbeel, P., and Phielipp, M · 2019
Later among the works it cites.
Go-explore: a new approach for hard-exploration problems
Ecoffet, A., Huizinga, J., Lehman, J., Stanley, K. O., and Clune, J · 2019
Later among the works it cites.
Soft actor-critic algorithms and applications, 2019
Haarnoja, T., Zhou, A., Hartikainen, K., Tucker, G., Ha, S., Tan, J., Kumar, V., Zhu, H., Gupta, A., Abbeel, P., and Levine, S · 2019
Later among the works it cites.
Learning latent dynamics for planning from pixels
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., and Davidson, J · 2019
Later among the works it cites.
Competitive experience replay
Liu, H., Trott, A., Socher, R., and Xiong, C · 2019
Later among the works it cites.
Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Overcoming exploration in reinforcement learning with demonstrations
Nair, A., McGrew, B., Andrychowicz, M., Zaremba, W., and Abbeel, P · 2018
Cited alongside, same era.
Multi-goal reinforcement learning: Challenging robotics environments and request for research
Plappert, M., Andrychowicz, M., Ray, A., McGrew, B., Baker, B., Powell, G., Schneider, J., Tobin, J., Chociej, M., Welinder, P., Kumar, V., and Zaremba, W · 2018
Cited alongside, same era.
Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Rajeswaran, A., Kumar, V., Gupta, A., Vezzani, G., Schulman, J., Todorov, E., and Levine, S · 2018
Cited alongside, same era.
Solving rubik’s cube with a robot hand, 2019
Akkaya, I., Andrychowicz, M., Chociej, M., Litwin, M., McGrew, B., Petron, A., Paino, A., Plappert, M., Powell, G., Ribas, R., Schneider, J., Tezak, N., Tworek, J., Welinder, P., Weng, L., Yuan, Q., Zaremba, W., and Zhang, L · 2019
Cited alongside, same era.
Lowrey, K., Rajeswaran, A., Kakade, S., Todorov, E., and Mordatch, I · 2019
Later among the works it cites.
Deep Dynamics Models for Learning Dexterous Manipulation
Nagabandi, A., Konoglie, K., Levine, S., and Kumar, V · 2019
Later among the works it cites.
Soft hindsight experience replay, 2020
He, Q., Zhuang, L., and Li, H · 2020
Closest in time.