CycleGAN-based speech enhancement for the unpaired training data

被引:0
作者
Yuan, Jing [1 ]
Bao, Changchun [1 ]
机构
[1] Beijing Univ Technol, Beijing, Peoples R China
来源
2019 ASIA-PACIFIC SIGNAL AND INFORMATION PROCESSING ASSOCIATION ANNUAL SUMMIT AND CONFERENCE (APSIPA ASC) | 2019年
基金
中国国家自然科学基金;
关键词
D O I
10.1109/apsipaasc47483.2019.9023072
中图分类号
TP31 [计算机软件];
学科分类号
081202 ; 0835 ;
摘要
Speech enhancement is an important task of improving speech quality in noise scenario. Many speech enhancement methods have achieved remarkable success based on the paired data. However, for many tasks, the paired training data is not available. In this paper, we present a speech enhancement method for the unpaired data based on cycle-consistent generative adversarial network (CycleGAN) that can minimize the reconstruction loss as much as possible. The proposed model employs two discriminators and two generators to preserve speech components and reduce noise so that the network could map features better for the unseen noise. In this method, the generators are used to generate the enhanced speech, and two discriminators are employed to discriminate real inputs and the outputs of the generators. The experimental results showed that the proposed method effectively improved the performance compared to traditional deep neural network (DNN) and the recent GAN-based speech enhancement methods.
引用
收藏
页码:878 / 883
页数:6
相关论文
共 50 条
[31]   Enhanced Night-to-Day Image Conversion Using CycleGAN-Based Base-Detail Paired Training [J].
Son, Dong-Min ;
Kwon, Hyuk-Ju ;
Lee, Sung-Hak .
MATHEMATICS, 2023, 11 (14)
[32]   Speech Enhancement Using Augmented SSL CycleGAN [J].
Popovic, Branislav ;
Krstanovic, Lidija ;
Janev, Marko ;
Suzic, Sinisa ;
Nosek, Tijana ;
Galic, Jovan .
2022 30TH EUROPEAN SIGNAL PROCESSING CONFERENCE (EUSIPCO 2022), 2022, :1155-1159
[33]   CycleGAN-based deep learning technique for artifact reduction in fundus photography [J].
Yoo, Tae Keun ;
Choi, Joon Yul ;
Kim, Hong Kyu .
GRAEFES ARCHIVE FOR CLINICAL AND EXPERIMENTAL OPHTHALMOLOGY, 2020, 258 (08) :1631-1637
[34]   Voice Privacy Through x-vector and CycleGAN-based Anonymization [J].
Prajapati, Gauri P. ;
Singh, Dipesh K. ;
Amin, Preet P. ;
Patil, Hemant A. .
INTERSPEECH 2021, 2021, :1684-1688
[35]   CYCLEGAN-BASED SAR IMAGE EXPANSION TO ENHANCE TARGET DETECTION PERFORMANCE [J].
Zhou, Fang ;
Liu, Rong ;
Yang, Tingting ;
Yang, Xiyu .
IGARSS 2024-2024 IEEE INTERNATIONAL GEOSCIENCE AND REMOTE SENSING SYMPOSIUM, IGARSS 2024, 2024, :2601-2604
[36]   CycleGAN-Based SAR-Optical Image Fusion for Target Recognition [J].
Sun, Yuchuang ;
Yan, Kaijia ;
Li, Wangzhe .
REMOTE SENSING, 2023, 15 (23)
[37]   CycleGAN-Based Clutter Suppression and Pipeline Positioning Method for GPR Image [J].
Wang, Jiachun ;
Lin, Yun ;
Ma, Deyun ;
Wang, Yanping ;
Ye, Shengbo .
IEEE GEOSCIENCE AND REMOTE SENSING LETTERS, 2025, 22
[38]   Unpaired Speech Enhancement by Acoustic and Adversarial Supervision for Speech Recognition [J].
Kim, Geonmin ;
Lee, Hwaran ;
Kim, Bo-Kyeong ;
Oh, Sang-Hoon ;
Lee, Soo-Young .
IEEE SIGNAL PROCESSING LETTERS, 2019, 26 (01) :159-163
[39]   Application of CycleGAN-based image style transfer algorithm in visual communication design [J].
Zhao, Ying .
JOURNAL OF COMPUTATIONAL METHODS IN SCIENCES AND ENGINEERING, 2025, 25 (04) :3152-3164
[40]   Improved CycleGAN-based feature recognition in young children and preschool education research [J].
Han J. ;
Yuchi X. ;
Su Y. .
Applied Mathematics and Nonlinear Sciences, 2024, 9 (01)