Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Rajeswaran, A., Kumar, V., Gupta, A., Vezzani, G., Schulman, J., Todorov, E., and Levine, S. (2018) · 2018
Later among the works it cites.
Few-shot goal inference for visuomotor learning and planning
Xie, A., Singh, A., Levine, S., and Finn, C. (2018) · 2018
Later among the works it cites.
TF-Agents: A library for reinforcement learning in tensorflow
Guadarrama, S., Korattikara, A., Ramirez, O., Castro, P., Holly, E., Fishman, S., Wang, K., Gonina, E., Wu, N., Kokiopoulou, E., Sbaiz, L., Smith, J., Bartók, G., Berent, J., Harris, C., Vanhoucke, V., and Brevdo, E. (2018) · 2019
Later among the works it cites.
Off-policy evaluation via off-policy classification
Irpan, A., Rao, K., Bousmalis, K., Harris, C., Ibarz, J., and Levine, S. (2019) · 2019
Later among the works it cites.
Imitation learning via off-policy distribution matching
Kostrikov, I., Nachum, O., and Tompson, J. (2019) · 2019
Later among the works it cites.
End-to-end robotic reinforcement learning without reward engineering
Singh, A., Yang, L., Finn, C., and Levine, S. (2019) · 2019
Later among the works it cites.
A practical approach to insertion with variable socket position using deep reinforcement learning
Vecerik, M., Sushkov, O., Barker, D., Rothörl, T., Hester, T., and Scholz, J. (2019) · 2019
Later among the works it cites.
Curl: Contrastive unsupervised representations for reinforcement learning
Laskin, M., Srinivas, A., and Abbeel, P. (2020) · 2020
Later among the works it cites.
Multifingered grasp planning via inference in deep neural networks: Outperforming sampling by learning differentiable models
Lu, Q., Van der Merwe, M., Sundaralingam, B., and Hermans, T. (2020) · 2020
Later among the works it cites.
DisCo RL: Distribution-Conditioned Reinforcement Learning for General-Purpose Policies
Nasiriany, S. (2020) · 2020
Later among the works it cites.
{SQIL}: Imitation learning via reinforcement learning with sparse rewards
Reddy, S., Dragan, A. D., and Levine, S. (2020) · 2020
Later among the works it cites.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Yu, T., Quillen, D., He, Z., Julian, R., Hausman, K., Finn, C., and Levine, S. (2020) · 2020
Later among the works it cites.
C-learning: Learning to achieve goals via recursive classification
Eysenbach, B., Salakhutdinov, R., and Levine, S. (2021) · 2021
Closest in time.