Fetching the paper…
Reading the bibliography…
In endoscopic procedures, autonomous tracking of abnormal regions and following circumferential cutting markers can significantly reduce the cognitive burden on endoscopists.
Endoscopic submucosal dissection
J. T. Maple, B. K. A. Dayyeh, S. S. Chauhan, J. H. Hwang, S. Komanduri, M. Manfredi, V. Konda, F. M. Murad, U. D. Siddiqui, and S. Banerjee · 2015
Earlier work this paper cites.
You only look once: Unified, real-time object detection
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi · 2016
Earlier work this paper cites.
Evaluation and stability analysis of video-based navigation system for functional endoscopic sinus surgery on in vivo clinical data
S. Leonard, A. Sinha, A. Reiter, M. Ishii, G. L. Gallia, R. H. Taylor, and G. D. Hager · 2018
Earlier work this paper cites.
Learning where to look while tracking instruments in robot-assisted surgery
M. Islam, Y. Li, and H. Ren · 2019
Earlier work this paper cites.
Learning curve for endoscopic submucosal dissection with an untutored, prevalence-based approach in the united states
X. Zhang, E. K. Ly, S. Nithyanand, R. J. Modayil, D. O. Khodorskiy, S. Neppala, S. Bhumi, M. DeMaria, J. L. Widmer, D. M. Friedel, et al · 2020
Earlier work this paper cites.
Applying depth-sensing to automated surgical manipulation with a da vinci robot
M. Hwang, D. Seita, B. Thananjeyan, J. Ichnowski, S. Paradis, D. Fer, T. Low, and K. Goldberg · 2020
Earlier work this paper cites.
easyendo robotic endoscopy system: Development and usability test in a randomized controlled trial with novices and physicians
D.-H. Lee, B. Cheon, J. Kim, and D.-S. Kwon · 2021
Earlier work this paper cites.
The future of endoscopic navigation: a review of advanced endoscopic vision technology
Z. Fu, Z. Jin, C. Zhang, Z. He, Z. Zha, C. Hu, T. Gan, Q. Yan, P. Wang, and X. Ye · 2021
Earlier work this paper cites.
Surrol: An open-source reinforcement learning centered and dvrk compatible platform for surgical robot learning
J. Xu, B. Li, B. Lu, Y.-H. Liu, Q. Dou, and P.-A. Heng · 2021
Earlier work this paper cites.
Minimally invasive gastrointestinal surgery: from past to the future
R. Rudiman · 2021
Earlier work this paper cites.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, W. Chen, et al · 2022
Earlier work this paper cites.
A survey on in-context learning
Q. Dong, L. Li, D. Dai, C. Zheng, J. Ma, R. Li, H. Xia, J. Xu, Z. Wu, T. Liu, et al · 2022
Earlier work this paper cites.
End-to-end learning of deep visuomotor policy for needle picking
H. Lin, B. Li, X. Chu, Q. Dou, Y. Liu, and K. W. S. Au · 2023
Earlier work this paper cites.
Human-in-the-loop embodied intelligence with interactive simulation environment for surgical robot learning
Y. Long, W. Wei, T. Huang, Y. Wang, and Q. Dou · 2023
Cited alongside, same era.
Guided reinforcement learning with efficient exploration for task automation of surgical robot
T. Huang, K. Chen, B. Li, Y.-H. Liu, and Q. Dou · 2023
Cited alongside, same era.
J. Achiam, S. Adler, S. Agarwal, L. Ahmad, I. Akkaya, F. L. Aleman, D. Almeida, J. Altenschmidt, S. Altman, S. Anadkat, et al · 2023
Cited alongside, same era.
Visual instruction tuning
H. Liu, C. Li, Q. Wu, and Y. J. Lee · 2023
Cited alongside, same era.
Rt-2: Vision-language-action models transfer web knowledge to robotic control
B. Zitkovich, T. Yu, S. Xu, P. Xu, T. Xiao, F. Xia, J. Wu, P. Wohlhart, S. Welker, A. Wahid, et al · 2023
Cited alongside, same era.
Openvla: An open-source vision-language-action model
M. J. Kim, K. Pertsch, S. Karamcheti, T. Xiao, A. Balakrishna, S. Nair, R. Rafailov, E. Foster, G. Lam, P. Sanketi, et al · 2024
Later among the works it cites.
Learn from safe experience: Safe reinforcement learning for task automation of surgical robot
K. Fan, Z. Chen, G. Ferrigno, and E. De Momi · 2024
Later among the works it cites.
Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0
A. O’Neill, A. Rehman, A. Maddukuri, A. Gupta, A. Padalkar, A. Lee, A. Pooley, A. Gupta, A. Mandlekar, A. Jain, et al · 2024
Later among the works it cites.
General-purpose foundation models for increased autonomy in robot-assisted surgery
S. Schmidgall, J. W. Kim, A. Kuntz, A. E. Ghazi, and A. Krieger · 2024
Later among the works it cites.
π 0 \displaystyle\pi_{0} : A vision-language-action flow model for general robot control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Hu, Q. Xie, V. Jain, J. Francis, J. Patrikar, N. Keetha, S. Kim, Y. Xie, T. Zhang, H.-S. Fang, et al · 2023
Cited alongside, same era.
Perceiver-actor: A multi-task transformer for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2023
Cited alongside, same era.
Rt-2: Vision-language-action models transfer web knowledge to robotic control
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, X. Chen, K. Choromanski, T. Ding, D. Driess, A. Dubey, C. Finn, et al · 2023
Cited alongside, same era.
Unsloth, 2023
M. H. Daniel Han and U. team · 2023
Cited alongside, same era.
Transendoscopic flexible parallel continuum robotic mechanism for bimanual endoscopic submucosal dissection
H. Gao, X. Yang, X. Xiao, X. Zhu, T. Zhang, C. Hou, H. Liu, M. Q.-H. Meng, L. Sun, X. Zuo, et al · 2024
Cited alongside, same era.
Robonurse-vla: Robotic scrub nurse system based on vision-language-action model
S. Li, J. Wang, R. Dai, W. Ma, W. Y. Ng, Y. Hu, and Z. Li · 2024
Cited alongside, same era.
Sufia: language-guided augmented dexterity for robotic surgical assistants
M. Moghani, L. Doorenbos, W. C.-H. Panitch, S. Huver, M. Azizian, K. Goldberg, and A. Garg · 2024
Cited alongside, same era.
K. Black, N. Brown, D. Driess, A. Esmail, M. Equi, C. Finn, N. Fusai, L. Groom, K. Hausman, B. Ichter, et al · 2024
Later among the works it cites.
Surgical task automation using actor-critic frameworks and self-supervised imitation learning
J. Liu, A. Andres, Y. Jiang, X. Luo, W. Shu, and S. A. Tsaftaris · 2024
Later among the works it cites.
Surgical robot transformer (srt): Imitation learning for surgical tasks
J. W. Kim, T. Z. Zhao, S. Schmidgall, A. Deguet, M. Kobilarov, C. Finn, and A. Krieger · 2024
Later among the works it cites.
Deepseekmath: Pushing the limits of mathematical reasoning in open language models
Z. Shao, P. Wang, Q. Zhu, R. Xu, J. Song, X. Bi, H. Zhang, M. Zhang, Y. Li, Y. Wu, et al · 2024
Later among the works it cites.
Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution
P. Wang, S. Bai, S. Tan, S. Wang, Z. Fan, J. Bai, K. Chen, X. Liu, J. Wang, W. Ge, et al · 2024
Later among the works it cites.
Endochat: Grounded multimodal large language model for endoscopic surgery
G. Wang, L. Bai, J. Wang, K. Yuan, Z. Li, T. Jiang, X. He, J. Wu, Z. Chen, Z. Lei, et al · 2025
Closest in time.
Covla: Comprehensive vision-language-action dataset for autonomous driving
H. Arai, K. Miwa, K. Sasaki, K. Watanabe, Y. Yamaguchi, S. Aoki, and I. Yamamoto · 2025
Closest in time.
Deep reinforcement learning in surgical robotics: enhancing the automation level
C. Qian and H. Ren · 2025
Closest in time.
Vlm-r1: A stable and generalizable r1-style large vision-language model
H. Shen, P. Liu, J. Li, C. Fang, Y. Ma, J. Liao, Q. Shen, Z. Zhang, K. Zhao, Q. Zhang, R. Xu, and T. Zhao · 2025
Closest in time.