Pitch- and Formant-Based Order Adaptation of the Fractional Fourier Transform and Its Application to Speech Recognition

被引:2
作者
Yin, Hui [1 ,2 ]
Nadeu, Climent [1 ]
Hohmann, Volker [1 ,3 ]
机构
[1] Univ Politecn Cataluna, TALP Res Ctr, ES-08034 Barcelona, Spain
[2] Beijing Inst Technol, Dept Elect Engn, Beijing 100081, Peoples R China
[3] Carl von Ossietzky Univ Oldenburg, D-26111 Oldenburg, Germany
来源
EURASIP JOURNAL ON AUDIO SPEECH AND MUSIC PROCESSING | 2009年
关键词
FREQUENCY; AMPLITUDE; SIGNALS;
D O I
10.1155/2009/304579
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
Fractional Fourier transform(FrFT) has been proposed to improve the time-frequency resolution in signal analysis and processing. However, selecting the FrFT transform order for the proper analysis of multicomponent signals like speech is still debated. In this work, we investigated several order adaptation methods. Firstly, FFT-and FrFT-based spectrograms of an artificially-generated vowel are compared to demonstrate the methods. Secondly, an acoustic feature set combining MFCC and FrFT is proposed, and the transform orders for the FrFT are adaptively set according to various methods based on pitch and formants. A tonal vowel discrimination test is designed to compare the performance of these methods using the feature set. The results show that the FrFT-MFCC yields a better discriminability of tones and also of vowels, especially by using multitransform-order methods. Thirdly, speech recognition experiments were conducted on the clean intervocalic English consonants provided by the Consonant Challenge. Experimental results show that the proposed features with different order adaptation methods can obtain slightly higher recognition rates compared to the reference MFCC-based recognizer. Copyright (c) 2009 Hui Yin et al.
引用
收藏
页数:14
相关论文
共 45 条