Real-time instance segmentation of surgical instruments using attention and multi-scale feature fusion

被引:23
|
作者
Angeles Ceron, Juan Carlos [1 ]
Ochoa Ruiz, Gilberto [1 ]
Chang, Leonardo [1 ]
Ali, Sharib [2 ,3 ,4 ]
机构
[1] Tecnol Monterrey, Escuela Ingn & Ciencias, Monterrey, NL, Mexico
[2] Univ Oxford, Inst Biomed Engn IBME, Dept Engn Sci, Oxford, England
[3] Univ Oxford, Oxford NIHR Biomed Res Ctr, Oxford, England
[4] Univ Leeds, Sch Comp, Leeds, W Yorkshire, England
关键词
Deep learning; MIS instance segmentation; Real-time; Single-stage; Attention; Multi-scale feature fusion;
D O I
10.1016/j.media.2022.102569
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Precise instrument segmentation aids surgeons to navigate the body more easily and increases patient safety. While accurate tracking of surgical instruments in real-time plays a crucial role in minimally invasive computer -assisted surgeries, it is a challenging task to achieve, mainly due to: (1) a complex surgical environment, and (2) model design trade-off in terms of both optimal accuracy and speed. Deep learning gives us the opportunity to learn complex environment from large surgery scene environments and placements of these instruments in real world scenarios. The Robust Medical Instrument Segmentation 2019 challenge (ROBUST-MIS) provides more than 10,000 frames with surgical tools in different clinical settings. In this paper, we propose a light-weight single stage instance segmentation model complemented with a convolutional block attention module for achieving both faster and accurate inference. We further improve accuracy through data augmentation and optimal anchor localization strategies. To our knowledge, this is the first work that explicitly focuses on both real-time performance and improved accuracy. Our approach out-performed top team performances in the most recent edition of ROBUST-MIS challenge with over 44% improvement on area-based multi-instance dice metric MI_DSC and 39% on distance-based multi-instance normalized surface dice MI_NSD. We also demonstrate real-time performance (> 60 frames-per-second) with different but competitive variants of our final approach.
引用
收藏
页数:12
相关论文
共 50 条
  • [31] Real-time airplane detection using multi-dimensional attention and feature fusion
    Li L.
    Peng N.
    Li B.
    Liu H.
    PeerJ Computer Science, 2023, 9 : 1 - 20
  • [32] SSD with multi-scale feature fusion and attention mechanism
    Liu, Qiang
    Dong, Lijun
    Zeng, Zhigao
    Zhu, Wenqiu
    Zhu, Yanhui
    Meng, Chen
    SCIENTIFIC REPORTS, 2023, 13 (01):
  • [33] SSD with multi-scale feature fusion and attention mechanism
    Qiang Liu
    Lijun Dong
    Zhigao Zeng
    Wenqiu Zhu
    Yanhui Zhu
    Chen Meng
    Scientific Reports, 13 (1)
  • [34] Multi-Scale Mixed Attention Tea Shoot Instance Segmentation Model
    Chen, Dongmei
    Cao, Peipei
    Yan, Lijie
    Chen, Huidong
    Lin, Jia
    Li, Xin
    Yuan, Lin
    Wu, Kaihua
    PHYTON-INTERNATIONAL JOURNAL OF EXPERIMENTAL BOTANY, 2024, 93 (02) : 261 - 275
  • [35] Multi-scale YOLACT for instance segmentation
    Zeng, Jiexian
    Ouyang, Huan
    Liu, Min
    Leng, Lu
    Fu, Xiang
    JOURNAL OF KING SAUD UNIVERSITY-COMPUTER AND INFORMATION SCIENCES, 2022, 34 (10) : 9419 - 9427
  • [36] Real-Time Instance Segmentation Method Based on Location Attention
    Liu, Li
    Kong, Yuqi
    KSII TRANSACTIONS ON INTERNET AND INFORMATION SYSTEMS, 2024, 18 (09): : 2483 - 2494
  • [37] LithoSegNet: Regional attention-based deep fusion of multi-scale and cross-stage features for real-time lithology segmentation
    Xu, Zhenhao
    Shi, Heng
    Lin, Peng
    Li, Shan
    INTERNATIONAL JOURNAL OF ROCK MECHANICS AND MINING SCIENCES, 2024, 180
  • [38] Mamba-UAV-SegNet: A Multi-Scale Adaptive Feature Fusion Network for Real-Time Semantic Segmentation of UAV Aerial Imagery
    Huang, Longyang
    Tan, Jintao
    Chen, Zhonghui
    DRONES, 2024, 8 (11)
  • [39] Attention-Guided Lightweight Network for Real-Time Segmentation of Robotic Surgical Instruments
    Ni, Zhen-Liang
    Bian, Gui-Bin
    Hou, Zeng-Guang
    Zhou, Xiao-Hu
    Xie, Xiao-Liang
    Li, Zhen
    2020 IEEE INTERNATIONAL CONFERENCE ON ROBOTICS AND AUTOMATION (ICRA), 2020, : 9939 - 9945
  • [40] MFCPNet: Real time medical image segmentation network via multi-scale feature fusion and channel pruning
    Hou, Linlin
    Yan, Zishen
    Desrosiers, Christian
    Liu, Hui
    BIOMEDICAL SIGNAL PROCESSING AND CONTROL, 2025, 100