Multi-task approach based on combined CNN-transformer for efficient segmentation and classification of breast tumors in ultrasound images

被引:0
作者
Jaouad Tagnamas
Hiba Ramadan
Ali Yahyaouy
Hamid Tairi
机构
[1] University of Sidi Mohamed Ben Abdellah,Department of Informatics, Faculty of Sciences Dhar El Mahraz
来源
Visual Computing for Industry, Biomedicine, and Art | / 7卷
关键词
Breast Ultrasound segmentation and classification; Breast tumors; Convolutional Neural Networks; Self-Attention; MLP-Mixer; Channel Attention;
D O I
暂无
中图分类号
学科分类号
摘要
Nowadays, inspired by the great success of Transformers in Natural Language Processing, many applications of Vision Transformers (ViTs) have been investigated in the field of medical image analysis including breast ultrasound (BUS) image segmentation and classification. In this paper, we propose an efficient multi-task framework to segment and classify tumors in BUS images using hybrid convolutional neural networks (CNNs)-ViTs architecture and Multi-Perceptron (MLP)-Mixer. The proposed method uses a two-encoder architecture with EfficientNetV2 backbone and an adapted ViT encoder to extract tumor regions in BUS images. The self-attention (SA) mechanism in the Transformer encoder allows capturing a wide range of high-level and complex features while the EfficientNetV2 encoder preserves local information in image. To fusion the extracted features, a Channel Attention Fusion (CAF) module is introduced. The CAF module selectively emphasizes important features from both encoders, improving the integration of high-level and local information. The resulting feature maps are reconstructed to obtain the segmentation maps using a decoder. Then, our method classifies the segmented tumor regions into benign and malignant using a simple and efficient classifier based on MLP-Mixer, that is applied for the first time, to the best of our knowledge, for the task of lesion classification in BUS images. Experimental results illustrate the outperformance of our framework compared to recent works for the task of segmentation by producing 83.42% in terms of Dice coefficient as well as for the classification with 86% in terms of accuracy.
引用
收藏
相关论文
共 188 条
[71]  
Zanjani FG(undefined)undefined undefined undefined undefined-undefined
[72]  
Singh VK(undefined)undefined undefined undefined undefined-undefined
[73]  
Abdel-Nasser M(undefined)undefined undefined undefined undefined-undefined
[74]  
Akram F(undefined)undefined undefined undefined undefined-undefined
[75]  
Rashwan HA(undefined)undefined undefined undefined undefined-undefined
[76]  
Sarker MMK(undefined)undefined undefined undefined undefined-undefined
[77]  
Pandey N(undefined)undefined undefined undefined undefined-undefined
[78]  
Lei BY(undefined)undefined undefined undefined undefined-undefined
[79]  
Huang S(undefined)undefined undefined undefined undefined-undefined
[80]  
Li R(undefined)undefined undefined undefined undefined-undefined