Real-time instance segmentation of surgical instruments using attention and multi-scale feature fusion

被引:23
|
作者
Angeles Ceron, Juan Carlos [1 ]
Ochoa Ruiz, Gilberto [1 ]
Chang, Leonardo [1 ]
Ali, Sharib [2 ,3 ,4 ]
机构
[1] Tecnol Monterrey, Escuela Ingn & Ciencias, Monterrey, NL, Mexico
[2] Univ Oxford, Inst Biomed Engn IBME, Dept Engn Sci, Oxford, England
[3] Univ Oxford, Oxford NIHR Biomed Res Ctr, Oxford, England
[4] Univ Leeds, Sch Comp, Leeds, W Yorkshire, England
关键词
Deep learning; MIS instance segmentation; Real-time; Single-stage; Attention; Multi-scale feature fusion;
D O I
10.1016/j.media.2022.102569
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Precise instrument segmentation aids surgeons to navigate the body more easily and increases patient safety. While accurate tracking of surgical instruments in real-time plays a crucial role in minimally invasive computer -assisted surgeries, it is a challenging task to achieve, mainly due to: (1) a complex surgical environment, and (2) model design trade-off in terms of both optimal accuracy and speed. Deep learning gives us the opportunity to learn complex environment from large surgery scene environments and placements of these instruments in real world scenarios. The Robust Medical Instrument Segmentation 2019 challenge (ROBUST-MIS) provides more than 10,000 frames with surgical tools in different clinical settings. In this paper, we propose a light-weight single stage instance segmentation model complemented with a convolutional block attention module for achieving both faster and accurate inference. We further improve accuracy through data augmentation and optimal anchor localization strategies. To our knowledge, this is the first work that explicitly focuses on both real-time performance and improved accuracy. Our approach out-performed top team performances in the most recent edition of ROBUST-MIS challenge with over 44% improvement on area-based multi-instance dice metric MI_DSC and 39% on distance-based multi-instance normalized surface dice MI_NSD. We also demonstrate real-time performance (> 60 frames-per-second) with different but competitive variants of our final approach.
引用
收藏
页数:12
相关论文
共 50 条
  • [1] A hybrid attention multi-scale fusion network for real-time semantic segmentation
    Ye, Baofeng
    Xue, Renzheng
    Wu, Qianlong
    SCIENTIFIC REPORTS, 2025, 15 (01):
  • [2] Real-time segmentation method of billet infrared image based on multi-scale feature fusion
    Lixin Zhang
    Qingrong Nan
    Shengqin Bian
    Tao Liu
    Zhengguang Xu
    Scientific Reports, 12
  • [3] Real-time segmentation method of billet infrared image based on multi-scale feature fusion
    Zhang, Lixin
    Nan, Qingrong
    Bian, Shengqin
    Liu, Tao
    Xu, Zhengguang
    SCIENTIFIC REPORTS, 2022, 12 (01)
  • [4] Real-Time Robotic Grasp Detection with Multi-Scale Feature Fusion
    Ma, Hao
    Yuan, Ding
    Cao, Zhe
    Yin, Jihao
    2020 IEEE INTERNATIONAL CONFERENCE ON REAL-TIME COMPUTING AND ROBOTICS (IEEE-RCAR 2020), 2020, : 140 - 145
  • [5] BFMNet: Bilateral feature fusion network with multi-scale context aggregation for real-time semantic segmentation
    Liu, Jin
    Zhang, Fangyu
    Zhou, Ziyin
    Wang, Jiajun
    NEUROCOMPUTING, 2023, 521 : 27 - 40
  • [6] A novel lightweight multi-scale feature fusion segmentation algorithm for real-time cervical lesion screening
    Yang, Jiahui
    Zhang, Ying
    Fan, Wenlong
    Wang, Jie
    Zhang, Xinhe
    Liu, Chunhui
    Liu, Shuang
    Xue, Linyan
    SCIENTIFIC REPORTS, 2025, 15 (01):
  • [7] Using multi-scale feature predictions for FPN architecture based real-time semantic segmentation
    Quyen, Van Toan
    Kim, Min Young
    2024 FIFTEENTH INTERNATIONAL CONFERENCE ON UBIQUITOUS AND FUTURE NETWORKS, ICUFN 2024, 2024, : 4 - 9
  • [8] Multi-scale feature fusion network with local attention for lung segmentation
    Xie, Yinghua
    Zhou, Yuntong
    Wang, Chen
    Ma, Yanshan
    Yang, Ming
    SIGNAL PROCESSING-IMAGE COMMUNICATION, 2023, 119
  • [9] BOUNDARY CORRECTED MULTI-SCALE FUSION NETWORK FOR REAL-TIME SEMANTIC SEGMENTATION
    Jiang, Tianjiao
    Jin, Yi
    Liang, Tengfei
    Wang, Xu
    Li, Yidong
    2022 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP, 2022, : 1886 - 1890
  • [10] Efficient Parallel Branch Network With Multi-Scale Feature Fusion for Real-Time Overhead Power Line Segmentation
    Gao, Zishu
    Yang, Guodong
    Li, En
    Liang, Zize
    Guo, Rui
    IEEE SENSORS JOURNAL, 2021, 21 (10) : 12220 - 12227