Fetching the paper…
Reading the bibliography…
Reinforcement learning is still struggling with solving long-horizon surgical robot tasks which involve multiple steps over an extended duration of time due to the policy exploration challenge.
G. Konidaris and A. Barto, “Skill discovery in continuous reinforcement learning domains using skill chaining,”
2009
Earlier work this paper cites.
G. Konidaris, S. Kuindersma, R. Grupen, and A. Barto, “Robot learning from demonstration by constructing skill trees,”
2012
Earlier work this paper cites.
T. Schaul, D. Horgan, K. Gregor, and D. Silver, “Universal value function approximators,” in
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in
2015
Earlier work this paper cites.
B. Thananjeyan, A. Garg, S. Krishnan, C. Chen, L. Miller, and K. Goldberg, “Multilateral surgical pattern cutting in 2d orthotropic gauze with deep reinforcement learning policies for tensioning,” in
2017
Earlier work this paper cites.
J. Andreas, D. Klein, and S. Levine, “Modular multitask reinforcement learning with policy sketches,” in
2017
Earlier work this paper cites.
M. Riedmiller, R. Hafner, T. Lampe, M. Neunert, J. Degrave, T. Wiele, V. Mnih, N. Heess, and J. T. Springenberg, “Learning by playing solving sparse reward tasks from scratch,” in
2018
Earlier work this paper cites.
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming exploration in reinforcement learning with demonstrations,” in
2018
Earlier work this paper cites.
A. Clegg, W. Yu, J. Tan, C. K. Liu, and G. Turk, “Learning to dress: Synthesizing human dressing motion via deep reinforcement learning,”
2018
Earlier work this paper cites.
D. Ghosh, A. Singh, A. Rajeswaran, V. Kumar, and S. Levine, “Divide-and-conquer reinforcement learning,” in
2018
Earlier work this paper cites.
J. Oh, Y. Guo, S. Singh, and H. Lee, “Self-imitation learning,” in
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in
2018
Earlier work this paper cites.
M. Yip and N. Das, “Robot autonomy for surgery,” in
2019
Earlier work this paper cites.
Y. Lee, S.-H. Sun, S. Somasundaram, E. S. Hu, and J. J. Lim, “Composing complex skills by learning transition policies,” in
2019
Cited alongside, same era.
T. Nguyen, N. D. Nguyen, F. Bello, and S. Nahavandi, “A new tensioning method using deep reinforcement learning for surgical pattern cutting,” in
2019
Cited alongside, same era.
N. D. Nguyen, T. Nguyen, S. Nahavandi, A. Bhatti, and G. Guest, “Manipulating soft tissues by deep reinforcement learning for autonomous robotic surgery,” in
2019
Cited alongside, same era.
M. Ginesi, D. Meli, H. Nakawala, A. Roberti, and P. Fiorini, “A knowledge-based framework for task automation in surgery,” in
2019
Cited alongside, same era.
S. Nasiriany, V. Pong, S. Lin, and S. Levine, “Planning with goal-conditioned policies,” 2019
2019
Cited alongside, same era.
J. Ibarz, J. Tan, C. Finn, M. Kalakrishnan, P. Pastor, and S. Levine, “How to train your robot with deep reinforcement learning: lessons we have learned,”
2021
Later among the works it cites.
Y. Lee, J. J. Lim, A. Anandkumar, and Y. Zhu, “Adversarial skill chaining for long-horizon robot manipulation via terminal state regularization,” in
2021
Later among the works it cites.
J. Xu, B. Li, B. Lu, Y.-H. Liu, Q. Dou, and P.-A. Heng, “Surrol: An open-source reinforcement learning centered and dvrk compatible platform for surgical robot learning,” in
2021
Later among the works it cites.
A. Pore, D. Corsi, E. Marchesini, D. Dall’Alba, A. Casals, A. Farinelli, and P. Fiorini, “Safe reinforcement learning using formal verification for tissue retraction in autonomous robotic-assisted surgery,”
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Ding, C. Florensa, P. Abbeel, and M. Phielipp, “Goal-conditioned imitation learning,” in
2019
Cited alongside, same era.
Y. Lee, J. Yang, and J. J. Lim, “Learning to coordinate manipulation skills via skill behavior diversification,” in
2020
Cited alongside, same era.
E. Tagliabue, A. Pore, D. Dall’Alba, E. Magnabosco, M. Piccinelli, and P. Fiorini, “Soft tissue simulation environment to learn manipulation tasks in autonomous robotic surgery,” in
2020
Cited alongside, same era.
M. Ginesi, D. Meli, A. Roberti, N. Sansonetto, and P. Fiorini, “Autonomous task planning and situation awareness in robotic surgery,” in
2020
Cited alongside, same era.
V. M. Varier, D. K. Rajamani, N. Goldfarb, F. Tavakkolmoghaddam, A. Munawar, and G. S. Fischer, “Collaborative suturing: A reinforcement learning approach to automate hand-off task in suturing for surgical robots,”
2020
Cited alongside, same era.
A. Bagaria and G. Konidaris, “Option discovery using deep skill chaining,” in
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
K. L. Schwaner, I. Iturrate, J. K. Andersen, P. T. Jensen, and T. R. Savarimuthu, “Autonomous bi-manual surgical suturing based on skills learned from demonstration,” in
2021
Later among the works it cites.
K. L. Schwaner, D. Dall’Alba, P. T. Jensen, P. Fiorini, and T. R. Savarimuthu, “Autonomous needle manipulation for robotic surgical suturing based on skills learned from demonstration,” in
2021
Later among the works it cites.
A. Bagaria, J. Senthil, M. Slivinski, and G. Konidaris, “Robustly learning composable options in deep reinforcement learning,” in
2021
Later among the works it cites.
J.-S. BYUN and A. Perrault, “Training transition policies via distribution matching for complex tasks,” in
2022
Later among the works it cites.
D. Zhang, Z. Wu, J. Chen, R. Zhu, A. Munawar, B. Xiao, Y. Guan, H. Su, W. Hong, Y. Guo,
2022
Later among the works it cites.
A. Wilcox, J. Kerr, B. Thananjeyan, J. Ichnowski, M. Hwang, S. Paradis, D. Fer, and K. Goldberg, “Learning to localize, grasp, and hand over unmodified surgical needles,” in
2022
Later among the works it cites.
T. Huang, K. Chen, B. Li, Y.-H. Liu, and Q. Dou, “Guided reinforcement learning with efficient exploration for task automation of surgical robot,” in
2023
Closest in time.