AUTOMATIC PHONETIC SEGMENTATION IN MANDARIN CHINESE: BOUNDARY MODELS, GLOTTAL FEATURES AND TONE

被引:0
作者
Yuan, Jiahong [1 ]
Ryant, Neville [1 ]
Liberman, Mark [1 ]
机构
[1] Univ Penn, Philadelphia, PA 19104 USA
来源
2014 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP) | 2014年
关键词
Forced alignment; boundary model; glottal features; tone; Mandarin Chinese; PHONATION;
D O I
暂无
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
We conducted experiments on forced alignment in Mandarin Chinese. A corpus of 7,849 utterances was created for the purpose of the study. Systems differing in their use of explicit phone boundary models, glottal features, and tone information were trained and evaluated on the corpus. Results showed that employing special one-state phone boundary HMM models significantly improved forced alignment accuracy, even when no manual phonetic segmentation was available for training. Spectral features extracted from glottal waveforms (by performing glottal inverse filtering from the speech waveforms) also improved forced alignment accuracy. Tone dependent models only slightly outperformed tone independent models. The best system achieved 93.1% agreement (of phone boundaries) within 20 ms compared to manual segmentation without boundary correction.
引用
收藏
页数:5
相关论文
共 35 条