Study of Wavelet Packet Energy Entropy for Emotion Classification in Speech and Glottal Signals

被引:4
作者
He, Ling [1 ]
Lech, Margaret [2 ]
Zhang, Jing [1 ]
Ren, Xiaomei [1 ]
Deng, Lihua [1 ]
机构
[1] Sichuan Univ, Sch Elect Engn & Informat, Chengdu, Peoples R China
[2] RMIT Univ, Sch Elect & Comp Engn, Melbourne, Vic, Australia
来源
FIFTH INTERNATIONAL CONFERENCE ON DIGITAL IMAGE PROCESSING (ICDIP 2013) | 2013年 / 8878卷
基金
美国国家科学基金会;
关键词
Emotion recognition; feature extraction; wavelet packet energy entropy; perceptual wavelet packet; RECOGNITION; DEPRESSION; FEATURES;
D O I
10.1117/12.2030929
中图分类号
O43 [光学];
学科分类号
070207 ; 0803 ;
摘要
The automatic speech emotion recognition has important applications in human-machine communication. Majority of current research in this area is focused on finding optimal feature parameters. In recent studies, several glottal features were examined as potential cues for emotion differentiation. In this study, a new type of feature parameter is proposed, which calculates energy entropy on values within selected Wavelet Packet frequency bands. The modeling and classification tasks are conducted using the classical GMM algorithm. The experiments use two data sets: the Speech Under Simulated Emotion (SUSE) data set annotated with three different emotions (angry, neutral and soft) and Berlin Emotional Speech (BES) database annotated with seven different emotions (angry, bored, disgust, fear, happy, sad and neutral). The average classification accuracy achieved for the SUSE data (74%-76%) is significantly higher than the accuracy achieved for the BES data (51%-54%). In both cases, the accuracy was significantly higher than the respective random guessing levels (33% for SUSE and 14.3% for BES).
引用
收藏
页数:6
相关论文
共 18 条
  • [1] [Anonymous], 1997, P 5 EUROPEAN C SPEEC, DOI DOI 10.21437/EUROSPEECH.1997-494
  • [2] [Anonymous], 8 INT C SIGN PROC 16
  • [3] [Anonymous], 2001, Discrete-Time Speech Signal Processing:Principles and Practice
  • [4] [Anonymous], P INT 2005
  • [5] Emotion recognition in human-computer interaction
    Cowie, R
    Douglas-Cowie, E
    Tsapatsoulis, N
    Votsis, G
    Kollias, S
    Fellenz, W
    Taylor, JG
    [J]. IEEE SIGNAL PROCESSING MAGAZINE, 2001, 18 (01) : 32 - 80
  • [6] Cummings K.E., 1993, AC SPEECH SIGN PROC
  • [7] Cummings K.E., 1990, AC SPEECH SIGN PROC
  • [8] He L., 2008, INT 2008
  • [9] Emotion recognition in speech using inter-sentence Glottal statistics
    Iliev, Alexander I.
    Scordilis, Michael S.
    [J]. PROCEEDINGS OF IWSSIP 2008: 15TH INTERNATIONAL CONFERENCE ON SYSTEMS, SIGNALS AND IMAGE PROCESSING, 2008, : 465 - 468
  • [10] Moore E.I.I., 2004, ENG MED BIOL SOC 200