• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

使用增强卷积和递归神经网络监测手术视频中的工具使用情况。

Monitoring tool usage in surgery videos using boosted convolutional and recurrent neural networks.

机构信息

Inserm, UMR 1101, Brest F-29200, France.

Inserm, UMR 1101, Brest F-29200, France; Univ Bretagne Occidentale, Brest F-29200, France.

出版信息

Med Image Anal. 2018 Jul;47:203-218. doi: 10.1016/j.media.2018.05.001. Epub 2018 May 9.

DOI:10.1016/j.media.2018.05.001
PMID:29778931
Abstract

This paper investigates the automatic monitoring of tool usage during a surgery, with potential applications in report generation, surgical training and real-time decision support. Two surgeries are considered: cataract surgery, the most common surgical procedure, and cholecystectomy, one of the most common digestive surgeries. Tool usage is monitored in videos recorded either through a microscope (cataract surgery) or an endoscope (cholecystectomy). Following state-of-the-art video analysis solutions, each frame of the video is analyzed by convolutional neural networks (CNNs) whose outputs are fed to recurrent neural networks (RNNs) in order to take temporal relationships between events into account. Novelty lies in the way those CNNs and RNNs are trained. Computational complexity prevents the end-to-end training of "CNN+RNN" systems. Therefore, CNNs are usually trained first, independently from the RNNs. This approach is clearly suboptimal for surgical tool analysis: many tools are very similar to one another, but they can generally be differentiated based on past events. CNNs should be trained to extract the most useful visual features in combination with the temporal context. A novel boosting strategy is proposed to achieve this goal: the CNN and RNN parts of the system are simultaneously enriched by progressively adding weak classifiers (either CNNs or RNNs) trained to improve the overall classification accuracy. Experiments were performed in a dataset of 50 cataract surgery videos, where the usage of 21 surgical tools was manually annotated, and a dataset of 80 cholecystectomy videos, where the usage of 7 tools was manually annotated. Very good classification performance are achieved in both datasets: tool usage could be labeled with an average area under the ROC curve of A=0.9961 and A=0.9939, respectively, in offline mode (using past, present and future information), and A=0.9957 and A=0.9936, respectively, in online mode (using past and present information only).

摘要

本文研究了手术过程中工具使用的自动监测,其潜在应用包括报告生成、手术培训和实时决策支持。考虑了两种手术:白内障手术,这是最常见的手术程序,和胆囊切除术,这是最常见的消化系统手术之一。通过显微镜(白内障手术)或内窥镜(胆囊切除术)记录的视频来监测工具使用情况。按照最新的视频分析解决方案,视频的每一帧都由卷积神经网络(CNN)进行分析,其输出被馈送到递归神经网络(RNN)中,以考虑事件之间的时间关系。新颖之处在于训练这些 CNN 和 RNN 的方式。计算复杂性阻止了“CNN+RNN”系统的端到端训练。因此,CNN 通常首先独立于 RNN 进行训练。这种方法对于手术工具分析显然是次优的:许多工具彼此非常相似,但它们通常可以根据过去的事件进行区分。CNN 应该经过训练以结合时间上下文提取最有用的视觉特征。提出了一种新的增强策略来实现这一目标:系统的 CNN 和 RNN 部分通过逐步添加弱分类器(CNN 或 RNN)来同时丰富,这些弱分类器经过训练可以提高整体分类准确性。在 50 个白内障手术视频的数据集和 80 个胆囊切除术视频的数据集上进行了实验,其中手动注释了 21 种手术工具的使用情况和 7 种工具的使用情况。在两个数据集上都实现了非常好的分类性能:在离线模式(使用过去、现在和未来的信息)下,工具使用情况可以用平均 ROC 曲线下的面积 A=0.9961 和 A=0.9939 进行标记,在在线模式(仅使用过去和现在的信息)下,分别为 A=0.9957 和 A=0.9936。

相似文献

1
Monitoring tool usage in surgery videos using boosted convolutional and recurrent neural networks.使用增强卷积和递归神经网络监测手术视频中的工具使用情况。
Med Image Anal. 2018 Jul;47:203-218. doi: 10.1016/j.media.2018.05.001. Epub 2018 May 9.
2
Assessment of Automated Identification of Phases in Videos of Cataract Surgery Using Machine Learning and Deep Learning Techniques.使用机器学习和深度学习技术评估白内障手术视频中的相位自动识别。
JAMA Netw Open. 2019 Apr 5;2(4):e191860. doi: 10.1001/jamanetworkopen.2019.1860.
3
Surgical tool detection in cataract surgery videos through multi-image fusion inside a convolutional neural network.通过卷积神经网络内的多图像融合在白内障手术视频中进行手术工具检测。
Annu Int Conf IEEE Eng Med Biol Soc. 2017 Jul;2017:2002-2005. doi: 10.1109/EMBC.2017.8037244.
4
SV-RCNet: Workflow Recognition From Surgical Videos Using Recurrent Convolutional Network.SV-RCNet:基于递归卷积网络的手术视频工作流程识别
IEEE Trans Med Imaging. 2018 May;37(5):1114-1126. doi: 10.1109/TMI.2017.2787657.
5
CATARACTS: Challenge on automatic tool annotation for cataRACT surgery.白内障:白内障手术自动工具标注挑战。
Med Image Anal. 2019 Feb;52:24-41. doi: 10.1016/j.media.2018.11.008. Epub 2018 Nov 16.
6
Smart data augmentation for surgical tool detection on the surgical tray.用于手术托盘上手术工具检测的智能数据增强
Annu Int Conf IEEE Eng Med Biol Soc. 2017 Jul;2017:4407-4410. doi: 10.1109/EMBC.2017.8037833.
7
EndoNet: A Deep Architecture for Recognition Tasks on Laparoscopic Videos.EndoNet:腹腔镜视频识别任务的深度架构。
IEEE Trans Med Imaging. 2017 Jan;36(1):86-97. doi: 10.1109/TMI.2016.2593957. Epub 2016 Jul 22.
8
Real-time segmentation and recognition of surgical tasks in cataract surgery videos.白内障手术视频中手术任务的实时分割与识别。
IEEE Trans Med Imaging. 2014 Dec;33(12):2352-60. doi: 10.1109/TMI.2014.2340473. Epub 2014 Jul 18.
9
Adaptive detrending to accelerate convolutional gated recurrent unit training for contextual video recognition.自适应去趋势以加速上下文视频识别的卷积门控循环单元训练。
Neural Netw. 2018 Sep;105:356-370. doi: 10.1016/j.neunet.2018.05.009. Epub 2018 May 22.
10
A hybrid model based on neural networks for biomedical relation extraction.基于神经网络的生物医学关系抽取混合模型。
J Biomed Inform. 2018 May;81:83-92. doi: 10.1016/j.jbi.2018.03.011. Epub 2018 Mar 27.

引用本文的文献

1
Use of artificial intelligence in the analysis of digital videos of invasive surgical procedures: scoping review.人工智能在侵入性外科手术数字视频分析中的应用:范围综述。
BJS Open. 2025 Jul 1;9(4). doi: 10.1093/bjsopen/zraf073.
2
CatSkill: Artificial Intelligence-Based Metrics for the Assessment of Surgical Skill Level from Intraoperative Cataract Surgery Video Recordings.CatSkill:基于人工智能的术中白内障手术视频记录评估手术技能水平的指标
Ophthalmol Sci. 2025 Mar 14;5(4):100764. doi: 10.1016/j.xops.2025.100764. eCollection 2025 Jul-Aug.
3
Challenges in multi-centric generalization: phase and step recognition in Roux-en-Y gastric bypass surgery.
多中心泛化的挑战:Roux-en-Y 胃旁路手术中的相位和步骤识别。
Int J Comput Assist Radiol Surg. 2024 Nov;19(11):2249-2257. doi: 10.1007/s11548-024-03166-3. Epub 2024 May 18.
4
Depth over RGB: automatic evaluation of open surgery skills using depth camera.深度信息超越 RGB:利用深度相机自动评估开放手术技能。
Int J Comput Assist Radiol Surg. 2024 Jul;19(7):1349-1357. doi: 10.1007/s11548-024-03158-3. Epub 2024 May 15.
5
Artificial Intelligence in Cataract Surgery: A Systematic Review.人工智能在白内障手术中的应用:系统评价。
Transl Vis Sci Technol. 2024 Apr 2;13(4):20. doi: 10.1167/tvst.13.4.20.
6
Where do we stand in AI for endoscopic image analysis? Deciphering gaps and future directions.我们在人工智能用于内镜图像分析方面处于什么位置?解读差距与未来方向。
NPJ Digit Med. 2022 Dec 20;5(1):184. doi: 10.1038/s41746-022-00733-3.
7
Gauze Detection and Segmentation in Minimally Invasive Surgery Video Using Convolutional Neural Networks.基于卷积神经网络的微创手术视频中纱布检测与分割。
Sensors (Basel). 2022 Jul 11;22(14):5180. doi: 10.3390/s22145180.
8
Open surgery tool classification and hand utilization using a multi-camera system.多摄像机系统的开放性手术工具分类与手部使用
Int J Comput Assist Radiol Surg. 2022 Aug;17(8):1497-1505. doi: 10.1007/s11548-022-02691-3. Epub 2022 Jun 27.
9
Video-based fully automatic assessment of open surgery suturing skills.基于视频的开放式手术缝合技能全自动评估。
Int J Comput Assist Radiol Surg. 2022 Mar;17(3):437-448. doi: 10.1007/s11548-022-02559-6. Epub 2022 Feb 1.
10
Application of artificial intelligence in cataract management: current and future directions.人工智能在白内障治疗中的应用:现状与未来方向。
Eye Vis (Lond). 2022 Jan 7;9(1):3. doi: 10.1186/s40662-021-00273-z.