Real-time instance segmentation of surgical instruments using attention and multi-scale feature fusion

被引:23
作者
Angeles Ceron, Juan Carlos [1 ]
Ochoa Ruiz, Gilberto [1 ]
Chang, Leonardo [1 ]
Ali, Sharib [2 ,3 ,4 ]
机构
[1] Tecnol Monterrey, Escuela Ingn & Ciencias, Monterrey, NL, Mexico
[2] Univ Oxford, Inst Biomed Engn IBME, Dept Engn Sci, Oxford, England
[3] Univ Oxford, Oxford NIHR Biomed Res Ctr, Oxford, England
[4] Univ Leeds, Sch Comp, Leeds, W Yorkshire, England
关键词
Deep learning; MIS instance segmentation; Real-time; Single-stage; Attention; Multi-scale feature fusion;
D O I
10.1016/j.media.2022.102569
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Precise instrument segmentation aids surgeons to navigate the body more easily and increases patient safety. While accurate tracking of surgical instruments in real-time plays a crucial role in minimally invasive computer -assisted surgeries, it is a challenging task to achieve, mainly due to: (1) a complex surgical environment, and (2) model design trade-off in terms of both optimal accuracy and speed. Deep learning gives us the opportunity to learn complex environment from large surgery scene environments and placements of these instruments in real world scenarios. The Robust Medical Instrument Segmentation 2019 challenge (ROBUST-MIS) provides more than 10,000 frames with surgical tools in different clinical settings. In this paper, we propose a light-weight single stage instance segmentation model complemented with a convolutional block attention module for achieving both faster and accurate inference. We further improve accuracy through data augmentation and optimal anchor localization strategies. To our knowledge, this is the first work that explicitly focuses on both real-time performance and improved accuracy. Our approach out-performed top team performances in the most recent edition of ROBUST-MIS challenge with over 44% improvement on area-based multi-instance dice metric MI_DSC and 39% on distance-based multi-instance normalized surface dice MI_NSD. We also demonstrate real-time performance (> 60 frames-per-second) with different but competitive variants of our final approach.
引用
收藏
页数:12
相关论文
共 50 条
  • [21] Lightweight real-time vehicle collision warning based on deep learning multi-scale feature fusion
    Qiu, Chengqun
    Tang, Hao
    Xu, Xixi
    Peng, Yu
    Liu, Xin
    Ji, Jie
    Meng, Mingyu
    Tang, Yunqing
    [J]. PROCEEDINGS OF THE INSTITUTION OF MECHANICAL ENGINEERS PART D-JOURNAL OF AUTOMOBILE ENGINEERING, 2024,
  • [22] An Improved Multi-Scale Feature Fusion for Skin Lesion Segmentation
    Liu, Luzhou
    Zhang, Xiaoxia
    Li, Yingwei
    Xu, Zhinan
    [J]. APPLIED SCIENCES-BASEL, 2023, 13 (14):
  • [23] MFANet: Multi-scale feature fusion network with attention mechanism
    Wang, Gaihua
    Gan, Xin
    Cao, Qingcheng
    Zhai, Qianyu
    [J]. VISUAL COMPUTER, 2023, 39 (07) : 2969 - 2980
  • [24] MFANet: Multi-scale feature fusion network with attention mechanism
    Gaihua Wang
    Xin Gan
    Qingcheng Cao
    Qianyu Zhai
    [J]. The Visual Computer, 2023, 39 : 2969 - 2980
  • [25] Liver segmentation network based on detail enhancement and multi-scale feature fusion
    Lu, Tinglan
    Qin, Jun
    Qin, Guihe
    Shi, Weili
    Zhang, Wentao
    [J]. SCIENTIFIC REPORTS, 2025, 15 (01):
  • [26] Real-time detection network for tiny traffic sign using multi-scale attention module
    TingTing Yang
    Chao Tong
    [J]. Science China Technological Sciences, 2022, 65 : 396 - 406
  • [27] Real-time detection network for tiny traffic sign using multi-scale attention module
    Yang TingTing
    Tong Chao
    [J]. SCIENCE CHINA-TECHNOLOGICAL SCIENCES, 2022, 65 (02) : 396 - 406
  • [28] LRTransDet: A Real-Time SAR Ship-Detection Network with Lightweight ViT and Multi-Scale Feature Fusion
    Feng, Kunyu
    Lun, Li
    Wang, Xiaofeng
    Cui, Xiaoxin
    [J]. REMOTE SENSING, 2023, 15 (22)
  • [29] MASSD: Multi-scale attention single shot detector for surgical instruments
    Yu, Lingtao
    Wang, Pengcheng
    Yan, Yusheng
    Xia, Yongqiang
    Cao, Wei
    [J]. COMPUTERS IN BIOLOGY AND MEDICINE, 2020, 123
  • [30] DMPNet: Distributed Multi-Scale Pyramid Network for Real-Time Semantic Segmentation
    Atif, Nadeem
    Mazhar, Saquib
    Ahamed, Shaik Rafi
    Bhuyan, M. K.
    Alfarhood, Sultan
    Safran, Mejdl
    [J]. IEEE ACCESS, 2024, 12 : 16573 - 16585