Audiovisual emotion recognition in wild

被引:1
作者
Egils Avots
Tomasz Sapiński
Maie Bachmann
Dorota Kamińska
机构
[1] University of Tartu,Institute of Technology
[2] Tallinn University of Technology,Department of Health Technologies, School of Information Technologies
[3] Łodz University of Technology,Institute of Mechatronics and Information Systems
来源
Machine Vision and Applications | 2019年 / 30卷
关键词
Emotion recognition; Audio signal processing; Facial expression; Deep learning;
D O I
暂无
中图分类号
学科分类号
摘要
People express emotions through different modalities. Utilization of both verbal and nonverbal communication channels allows to create a system in which the emotional state is expressed more clearly and therefore easier to understand. Expanding the focus to several expression forms can facilitate research on emotion recognition as well as human–machine interaction. This article presents analysis of audiovisual information to recognize human emotions. A cross-corpus evaluation is done using three different databases as the training set (SAVEE, eNTERFACE’05 and RML) and AFEW (database simulating real-world conditions) as a testing set. Emotional speech is represented by commonly known audio and spectral features as well as MFCC coefficients. The SVM algorithm has been used for classification. In case of facial expression, faces in key frames are found using Viola–Jones face recognition algorithm and facial image emotion classification done by CNN (AlexNet). Multimodal emotion recognition is based on decision-level fusion. The performance of emotion recognition algorithm is compared with the validation of human decision makers.
引用
收藏
页码:975 / 985
页数:10
相关论文
共 50 条
[21]   Evaluation of Deep Convolutional Neural Network architectures for Emotion Recognition in the Wild [J].
Talipu, A. ;
Generosi, A. ;
Mengoni, M. ;
Giraldi, L. .
2019 IEEE 23RD INTERNATIONAL SYMPOSIUM ON CONSUMER TECHNOLOGIES (ISCT), 2019, :25-27
[22]   A Probabilistic Fusion Strategy for Audiovisual Emotion Recognition of Sparse and Noisy Data [J].
Lin, Jen-Chun ;
Wu, Chung-Hsien ;
Wei, Wen-Li .
1ST INTERNATIONAL CONFERENCE ON ORANGE TECHNOLOGIES (ICOT 2013), 2013, :278-281
[23]   Emotion Recognition Using EEG Signals and Audiovisual Features with Contrastive Learning [J].
Lee, Ju-Hwan ;
Kim, Jin-Young ;
Kim, Hyoung-Gook .
BIOENGINEERING-BASEL, 2024, 11 (10)
[24]   Survey on audiovisual emotion recognition: databases, features, and data fusion strategies [J].
Wu, Chung-Hsien ;
Lin, Jen-Chun ;
Wei, Wen-Li .
APSIPA TRANSACTIONS ON SIGNAL AND INFORMATION PROCESSING, 2014, 3 (03)
[25]   Audiovisual emotion recognition in schizophrenia: Reduced integration of facial and vocal affect [J].
de Jong, J. J. ;
Hodiamont, P. P. G. ;
Van den Stock, J. ;
de Geldera, B. .
SCHIZOPHRENIA RESEARCH, 2009, 107 (2-3) :286-293
[26]   HEU Emotion: a large-scale database for multimodal emotion recognition in the wild [J].
Jing Chen ;
Chenhui Wang ;
Kejun Wang ;
Chaoqun Yin ;
Cong Zhao ;
Tao Xu ;
Xinyi Zhang ;
Ziqiang Huang ;
Meichen Liu ;
Tao Yang .
Neural Computing and Applications, 2021, 33 :8669-8685
[27]   Multi-cue fusion for emotion recognition in the wild [J].
Yan, Jingwei ;
Zheng, Wenming ;
Cui, Zhen ;
Tang, Chuangao ;
Zhang, Tong ;
Zong, Yuan .
NEUROCOMPUTING, 2018, 309 :27-35
[28]   Emotion Recognition in the Wild from Videos using Images [J].
Bargal, Sarah Adel ;
Barsoum, Emad ;
Ferrer, Cristian Canton ;
Zhang, Cha .
ICMI'16: PROCEEDINGS OF THE 18TH ACM INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION, 2016, :433-436
[29]   HEU Emotion: a large-scale database for multimodal emotion recognition in the wild [J].
Chen, Jing ;
Wang, Chenhui ;
Wang, Kejun ;
Yin, Chaoqun ;
Zhao, Cong ;
Xu, Tao ;
Zhang, Xinyi ;
Huang, Ziqiang ;
Liu, Meichen ;
Yang, Tao .
NEURAL COMPUTING & APPLICATIONS, 2021, 33 (14) :8669-8685
[30]   Emotion Recognition in the Wild via Convolutional Neural Networks and Mapped Binary Patterns [J].
Levi, Gil ;
Hassner, Tal .
ICMI'15: PROCEEDINGS OF THE 2015 ACM INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION, 2015, :503-510