Back

Deep learning-based recognition model for surgical phases of minimally invasive hysterectomy: A multicentre retrospective study

Koike, R.; Takenaka, S.; Suzuki, Y.; Matsuzaki, H.; Harada, Y.; Nakabayashi, M.; Hirose, Y.; Chikazawa, K.; Shimada, K.; Yoshiizumi, E.; Komatsu, H.; Tanabe, H.; Matsumoto, K.

2026-05-17 obstetrics and gynecology
10.64898/2026.05.13.26353100 medRxiv
Show abstract

Objective: To develop and validate a robust deep-learning model capable of fine-grained phase recognition in total hysterectomy, particularly the complex periuterine dissection phase. Design: Multicentre retrospective observational study. Setting: Japan. Sample: Surgical videos (n = 764) from 43 institutions. Methods: We developed a robust and generalisable deep-learning model for surgical phase recognition in total hysterectomy, applicable to laparoscopic and robot-assisted procedures. Overall, 1,591,334 still images were annotated across nine surgical phases. A convolutional neural network (Xception architecture) was trained on 200 cases using four-fold cross-validation, with institutional separation between training and testing sets. Main outcome measures: Model performance was assessed using accuracy, precision, recall, and F1 score. Subgroup analysis and logistic regression evaluated the association between background clinical factors and recognition accuracy. Results: The model achieved an overall phase recognition accuracy of 0.78 (95% CI: 0.74--0.80), with a precision of 0.75 (95% CI: 0.72--0.78) and a recall of 0.76 (95% CI: 0.74--0.78). Performance was consistent across laparoscopic and robot-assisted procedures and across most surgical phases. Accuracy plateaued after training on 120 cases. No clinical factors significantly impacted performance. Trends toward lower accuracy were observed for cases with cervical myoma and pouch of Douglas adhesions. Conclusions: This model demonstrated high accuracy across diverse institutions and patient backgrounds. Its potential applications include surgical education, real-time intraoperative support, and training efficiency enhancement.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

1
npj Digital Medicine
118 papers in training set
Top 0.4%
15.6%
2
PLOS ONE
5266 papers in training set
Top 13%
13.5%
3
Scientific Reports
3612 papers in training set
Top 2%
13.1%
4
Bioengineering
29 papers in training set
Top 0.1%
12.3%
50% of probability mass above
5
Diagnostics
50 papers in training set
Top 0.1%
8.1%
6
Medical Image Analysis
35 papers in training set
Top 0.2%
3.5%
7
Nature Communications
5641 papers in training set
Top 38%
2.7%
8
Computational and Structural Biotechnology Journal
242 papers in training set
Top 2%
2.5%
9
International Journal of Medical Informatics
26 papers in training set
Top 0.6%
2.2%
10
Healthcare
17 papers in training set
Top 0.2%
2.2%
11
BMC Medical Education
21 papers in training set
Top 0.3%
2.0%
12
iScience
1154 papers in training set
Top 16%
1.8%
13
Journal of the American Medical Informatics Association
71 papers in training set
Top 2%
1.4%
14
BMJ Open
601 papers in training set
Top 11%
1.2%
15
Nature Medicine
125 papers in training set
Top 2%
1.2%
16
BMC Medicine
176 papers in training set
Top 3%
1.2%
17
JAMIA Open
42 papers in training set
Top 1%
1.1%
18
eLife
5828 papers in training set
Top 60%
1.1%
19
Journal of Pathology Informatics
15 papers in training set
Top 0.2%
1.0%
20
Journal of Clinical Medicine
97 papers in training set
Top 4%
1.0%
21
British Journal of Anaesthesia
17 papers in training set
Top 0.3%
1.0%
22
JAMA Network Open
130 papers in training set
Top 4%
0.9%
23
Kidney International Reports
15 papers in training set
Top 0.3%
0.9%