A Priori SNR Estimation Using Air- and Bone-Conduction Microphones

被引:16
作者
Shin, Ho Seon [1 ]
Fingscheidt, Tim [2 ]
Kang, Hong-Goo [1 ]
机构
[1] Yonsei Univ, Dept Elect & Elect Engn, Seoul 120749, South Korea
[2] Tech Univ Carolo Wilhelmina Braunschweig, Inst Commun Technol, D-38160 Braunschweig, Germany
关键词
A priori signal-to-noise ratio (SNR); bone-conduction (BC) microphone; speech enhancement; SPEECH ENHANCEMENT;
D O I
10.1109/TASLP.2015.2446202
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
This paper proposes an a priori signal-to-noise ratio (SNR) estimator using an air-conduction (AC) and a bone-conduction (BC) microphone. Among various ways of combining AC and BC microphones for speech enhancement, it is shown that the total enhancement performance can be maximized if the BC microphone is utilized for estimating the power spectral density (PSD) of the desired speech signal. Considering the fact that a small deviation in the speech PSD estimation process brings severe spectral distortion, this paper focuses on controlling weighting factors while estimating the a priori SNR with the decision-directed approach framework. The time-frequency varying weighting factor that is determined by taking a minimum mean square error criterion improves the capability of eliminating residual noise and minimizing speech distortion. Since the weighting factors are also adjusted by measuring the usefulness of the AC and BC microphones, the proposed approach is suitable for tracking the parameter even if the characteristic of environment changes rapidly. The simulation results confirm the superiority of the proposed algorithm to conventional algorithms in high noise environments.
引用
收藏
页码:2015 / 2025
页数:11
相关论文
共 28 条
[1]  
[Anonymous], 2013, Speech Enhancement: Theory and Practice
[2]  
[Anonymous], 2008, DHSTRPSC0805
[3]  
Atkinson D. J., 2013, INTELLIGIBILITY ADAP, P13
[4]   Analysis of the Decision-Directed SNR Estimator for Speech Enhancement With Respect to Low-SNR and Transient Conditions [J].
Breithaupt, Colin ;
Martin, Rainer .
IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2011, 19 (02) :277-289
[5]   Elimination of the Musical Noise Phenomenon with the Ephraim and Malah Noise Suppressor [J].
Cappe, Olivier .
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING, 1994, 2 (02) :345-349
[6]   Impact of SNR and gain-function over- and under-estimation on speech intelligibility [J].
Chen, Fei ;
Loizou, Philipos C. .
SPEECH COMMUNICATION, 2012, 54 (02) :272-281
[7]   Noise estimation by minima controlled recursive averaging for robust speech enhancement [J].
Cohen, I ;
Berdugo, B .
IEEE SIGNAL PROCESSING LETTERS, 2002, 9 (01) :12-15
[8]   Body Conducted Speech Enhancement by Equalization and Signal Fusion [J].
Dekens, Tomas ;
Verhelst, Werner .
IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2013, 21 (12) :2481-2492
[9]   SPEECH ENHANCEMENT USING A MINIMUM MEAN-SQUARE ERROR SHORT-TIME SPECTRAL AMPLITUDE ESTIMATOR [J].
EPHRAIM, Y ;
MALAH, D .
IEEE TRANSACTIONS ON ACOUSTICS SPEECH AND SIGNAL PROCESSING, 1984, 32 (06) :1109-1121
[10]   A modified A priori SNR for speech enhancement using spectral subtraction rules [J].
Hasan, MK ;
Salahuddin, S ;
Khan, MR .
IEEE SIGNAL PROCESSING LETTERS, 2004, 11 (04) :450-453