Fetching the paper…
Reading the bibliography…
Surgical robot task automation has recently attracted great attention due to its potential to benefit both surgeons and patients.
C. J. Watkins and P. Dayan, “Q-Learning,” Machine Learning , vol. 8, no. 3-4, pp. 279–292, 1992
1992
Earlier work this paper cites.
A. Y. Ng and S. J. Russell, “Algorithms for Inverse Reinforcement Learning,” ICML , pp. 663–670, 2000
2000
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum Entropy Inverse Reinforcement Learning,” AAAI Conference on Artificial Intelligence (AAAI-08) , pp. 1433–1438, 2008
2008
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A Survey of Robot Learning from Demonstration,” Robotics and Autonomous Systems , vol. 57, no. 5, pp. 469–483, 2009
2009
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “ImageNet classification with deep convolutional neural networks,” Advances in Neural Information Processing Systems , vol. 25, 2012
2012
Earlier work this paper cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement Learning in Robotics: A Survey,” The International Journal of Robotics Research , vol. 32, no. 11, pp. 1238–1274, 2013
2013
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget, M. Mirza, and et al, “Generative adversarial networks,” NeurIPS , vol. 27, pp. 2672–2680, 2014
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” ICLR , 2015
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, and et al, “Human-Level Control Through Deep Reinforcement Learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: a method for stochastic optimization,” International Conference on Learning Representations , May 2015
2015
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative Adversarial Imitation Learning,” Advances in Neural Information Processing Systems , vol. 29, 2016
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous Control with Deep Reinforcement Learning,” ICLR May 2-4 , 2016
2016
Earlier work this paper cites.
H. v. Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double Q-Learning,” AAAI , p. 2094–2100, 2016
2016
Earlier work this paper cites.
M. Vecerik, C. Hesse, T. Wang, and et al, “Integrating Behavior Cloning and Reinforcement Learning for Improved Performance in Dense and Sparse Reward Environments,” Advances in Neural Information Processing Systems , pp. 5308–5317, 2017
2017
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation Learning: A Survey of Learning Methods,” ACM Computing Surveys (CSUR) , vol. 50, no. 2, pp. 1–35, 2017
2017
Earlier work this paper cites.
M. Andrychowicz, D. Crow, A. Ray, and et al, “Hindsight Experience Replay,” NIPS , pp. 5048–5058, 2017
2017
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction , 2nd ed. MIT Press, 2018
2018
Earlier work this paper cites.
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming Exploration in Reinforcement Learning with Demonstrations,” ICRA , 2018
2018
Earlier work this paper cites.
J. Oh, Y. Guo, S. Singh, and H. Lee, “Self-Imitation Learning,” ICML , 2018
2018
Earlier work this paper cites.
F. Torabi, G. Warnell, and P. Stone, “Behavioral Cloning from Observation,” IJCAI , 2018
2018
Earlier work this paper cites.
S. Fujimoto, H. van Hoof, and D. Meger, “Addressing Function Approximation Error in Actor-Critic Methods,” ICML , vol. 80, pp. 1587–1596, 2018
2018
Earlier work this paper cites.
S. Fujimoto, D. Meger, and D. Precup, “Off-Policy Deep Reinforcement Learning without Exploration,” ICML , 2018
2018
Cited alongside, same era.
T. Osa, J. Pajarinen, and G. Neumann, An Algorithmic Perspective on Imitation Learning . Hanover, MA, USA: Now Publishers Inc., 2018
2018
Cited alongside, same era.
X. Tan, C.-B. Chng, Y. Su, and et al, “Robot-Assisted Training in Laparoscopy Using Deep Reinforcement Learning,” IEEE Robotics and Automation Letters , vol. 4, no. 2, pp. 485–492, 2019
2019
Cited alongside, same era.
F. Torabi, G. Warnell, and P. Stone, “Recent Advances in Imitation Learning from Observation,” IJCAI , pp. 6325–6331, 7 2019
2019
Cited alongside, same era.
F. Torabi, G. Warnell, and P. Stone, “Generative adversarial imitation from observation,” ICML Workshop , June 2019
2019
Cited alongside, same era.
T. Joshi, S. Makker, H. Kodamana, and H. Kandath, “Twin actor Twin Delayed Deep Deterministic Policy Gradient (TATD3) Learning for Batch Process Control,” Computers & Chemical Engineering , vol. 155, p. 107527, 2021
2021
Later among the works it cites.
I. Kostrikov, R. Fergus, J. Tompson, and O. Nachum, “Offline Reinforcement Learning with Fisher Divergence Critic Regularization,” Proceedings of the 38th International Conference on Machine Learning, 18-24 July 2021 , vol. 139, pp. 5774–5783, 2021
2021
Later among the works it cites.
J. Xu, B. Li, B. Lu, and et al, “SurRoL: An Open-source Reinforcement Learning Centered and dVRK Compatible Platform for Surgical Robot Learning,” IROS , 2021
2021
Later among the works it cites.
A. Plaat, Deep Reinforcement Learning . Springer, 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. J. Loftus, A. C. Filiberto, Y. Li, and et al, “Decision Analysis and Reinforcement Learning in Surgical Decision-Making,” Surgery , vol. 168, no. 2, pp. 273–278, 2020
2020
Cited alongside, same era.
Z. Yuan, D. Wu, X. Lin, M. Lin, Z. Zhang, S. Ma, H. Zhang, and B. Zhou, “SEABO: A Simple Search-Based Method for Offline Imitation Learning,” NeurIPS , pp. 18 032–18 042, 2020
2020
Cited alongside, same era.
H. Su, Y. Hu, Z. Li, and et al, “Reinforcement Learning Based Manipulation Skill Transferring for Robot-Assisted Minimally Invasive Surgery,” ICRA , pp. 2203–2208, 2020
2020
Cited alongside, same era.
S. Reddy, A. D. Dragan, and S. Levine, “SQIL: Imitation Learning via Reinforcement Learning with Sparse Rewards,” ICLR , 2020
2020
Cited alongside, same era.
Y. Pan, C.-A. Cheng, K. Saigol, and et al, “Imitation Learning for Agile Autonomous Driving,” The International Journal of Robotics Research , vol. 39, no. 2-3, pp. 286–302, 2020
2020
Cited alongside, same era.
S. Fujimoto, D. Meger, and D. Precup, “AWAC: Accelerating Online Reinforcement Learning with Offline Datasets,” ICML , 2020
2020
Cited alongside, same era.
K. Wei, M. J. Bansal, S. M. Kakade, C. Daskalakis, and A. A. Rusu, “Optimal Transport for Offline Imitation Learning,” Advances in Neural Information Processing Systems , pp. 14 694–14 704, 2020
2020
Cited alongside, same era.
2022
Later among the works it cites.
G.-H. Kim, J. Lee, Y. Jang, H. Yang, and K.-E. Kim, “LobsDICE: Offline Learning from Observation via Stationary Distribution Correction Estimation,” NeurIPS , 2022
2022
Later among the works it cites.
S. Arora, P. Doshi, and H. Younger, “A Survey of Inverse Reinforcement Learning: Challenges, Methods and Progress,” Artificial Intelligence Review , vol. 55, p. 4307–4346, 2022
2022
Later among the works it cites.
A. Segato, M. D. Marzo, S. Zucchelli, S. Galvan, R. Secoli, and E. De Momi, “Inverse Reinforcement Learning Intra-Operative Path Planning for Steerable Needle,” IEEE Transactions on Biomedical Engineering , vol. 69, no. 6, pp. 1995–2005, 2022
2022
Later among the works it cites.
C. D.Ettorre, S. Zirino, N. N.Dei, A. Stilli, E. Momi, and D. Stoyanov, “Learning Intraoperative Organ Manipulation with Context-Based Reinforcement Learning,” International Journal of Computer Assisted Radiology and Surgery , 2022
2022
Later among the works it cites.
A. T. Bourdillon, A. Garg, H. Wang, and et al, “Integration of Reinforcement Learning in a Virtual Robotic Surgical Simulation,” Surgical innovation , pp. 94–102, 2023
2023
Later among the works it cites.
Y. Ou and M. Tavakoli, “Sim-to-Real Surgical Robot Learning and Autonomous Planning for Internal Tissue Points Manipulation Using Reinforcement Learning,” IEEE Robotics and Automation Letters , vol. 8, no. 5, pp. 2502–2509, 2023
2023
Later among the works it cites.
P. M. Scheikl, E. Tagliabue, B. Gyenes, and et al, “Sim-to-Real Transfer for Visual Reinforcement Learning of Deformable Object Manipulation for Robot-Assisted Surgery,” IEEE Robotics and Automation Letters , vol. 8, no. 2, pp. 560–567, 2023
2023
Later among the works it cites.
Y. Ou and M. Tavakoli, “Towards safe and efficient reinforcement learning for surgical robots using real-time human supervision and demonstration,” ISMR , pp. 1–7, 2023
2023
Later among the works it cites.
D. Tarasov, V. Kurenkov, A. Nikulin, and S. Kolesnikov, “Revisiting the Minimalist Approach to Offline Reinforcement Learning,” Advances in Neural Information Processing Systems 36 (NeurIPS) , 2023
2023
Later among the works it cites.
Y. Long, W. Wei, T. Huang, Y. Wang, and Q. Dou, “Human-in-the-loop Embodied Intelligence with Interactive Simulation Environment for Surgical Robot Learning,” RAL , 2023
2023
Later among the works it cites.
R. Bendikas, V. Modugno, D. Kanoulas, F. Vasconcelos1, and D. Stoyanov1, “Learning Needle Pick-And-Place without Expert Demonstrations,” IEEE Robotics and Automation Letters , 2023
2023
Later among the works it cites.
T. Huang, K. Chen, W. Wei, J. Li, Y. Long, and Q. Dou, “Value-Informed Skill Chaining for Policy Learning of Long-Horizon Tasks with Surgical Robot,” IROS , 2023
2023
Later among the works it cites.
T. Huang, K. Chen, B. Li, and et al, “Demonstration-Guided Reinforcement Learning with Efficient Exploration for Task Automation of Surgical Robot,” IEEE International Conference on Robotics and Automation (ICRA) , 2023
2023
Later among the works it cites.
S. Schmidgall, J. W. Kim, A. Kuntz, A. E. Ghazi, and A. Krieger, “General-purpose foundation models for increased autonomy in robot-assisted surgery,” 2024
2024
Closest in time.