Audiovisual emotion recognition in wild

被引:1
作者
Egils Avots
Tomasz Sapiński
Maie Bachmann
Dorota Kamińska
机构
[1] University of Tartu,Institute of Technology
[2] Tallinn University of Technology,Department of Health Technologies, School of Information Technologies
[3] Łodz University of Technology,Institute of Mechatronics and Information Systems
来源
Machine Vision and Applications | 2019年 / 30卷
关键词
Emotion recognition; Audio signal processing; Facial expression; Deep learning;
D O I
暂无
中图分类号
学科分类号
摘要
People express emotions through different modalities. Utilization of both verbal and nonverbal communication channels allows to create a system in which the emotional state is expressed more clearly and therefore easier to understand. Expanding the focus to several expression forms can facilitate research on emotion recognition as well as human–machine interaction. This article presents analysis of audiovisual information to recognize human emotions. A cross-corpus evaluation is done using three different databases as the training set (SAVEE, eNTERFACE’05 and RML) and AFEW (database simulating real-world conditions) as a testing set. Emotional speech is represented by commonly known audio and spectral features as well as MFCC coefficients. The SVM algorithm has been used for classification. In case of facial expression, faces in key frames are found using Viola–Jones face recognition algorithm and facial image emotion classification done by CNN (AlexNet). Multimodal emotion recognition is based on decision-level fusion. The performance of emotion recognition algorithm is compared with the validation of human decision makers.
引用
收藏
页码:975 / 985
页数:10
相关论文
共 50 条
  • [21] A Probabilistic Fusion Strategy for Audiovisual Emotion Recognition of Sparse and Noisy Data
    Lin, Jen-Chun
    Wu, Chung-Hsien
    Wei, Wen-Li
    1ST INTERNATIONAL CONFERENCE ON ORANGE TECHNOLOGIES (ICOT 2013), 2013, : 278 - 281
  • [22] Survey on audiovisual emotion recognition: databases, features, and data fusion strategies
    Wu, Chung-Hsien
    Lin, Jen-Chun
    Wei, Wen-Li
    APSIPA TRANSACTIONS ON SIGNAL AND INFORMATION PROCESSING, 2014, 3 (03)
  • [23] Audiovisual emotion recognition in schizophrenia: Reduced integration of facial and vocal affect
    de Jong, J. J.
    Hodiamont, P. P. G.
    Van den Stock, J.
    de Geldera, B.
    SCHIZOPHRENIA RESEARCH, 2009, 107 (2-3) : 286 - 293
  • [24] Emotion Recognition Using EEG Signals and Audiovisual Features with Contrastive Learning
    Lee, Ju-Hwan
    Kim, Jin-Young
    Kim, Hyoung-Gook
    BIOENGINEERING-BASEL, 2024, 11 (10):
  • [25] HEU Emotion: a large-scale database for multimodal emotion recognition in the wild
    Jing Chen
    Chenhui Wang
    Kejun Wang
    Chaoqun Yin
    Cong Zhao
    Tao Xu
    Xinyi Zhang
    Ziqiang Huang
    Meichen Liu
    Tao Yang
    Neural Computing and Applications, 2021, 33 : 8669 - 8685
  • [26] Emotion Recognition in the Wild from Videos using Images
    Bargal, Sarah Adel
    Barsoum, Emad
    Ferrer, Cristian Canton
    Zhang, Cha
    ICMI'16: PROCEEDINGS OF THE 18TH ACM INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION, 2016, : 433 - 436
  • [27] Emotion Recognition in the Wild via Convolutional Neural Networks and Mapped Binary Patterns
    Levi, Gil
    Hassner, Tal
    ICMI'15: PROCEEDINGS OF THE 2015 ACM INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION, 2015, : 503 - 510
  • [28] HEU Emotion: a large-scale database for multimodal emotion recognition in the wild
    Chen, Jing
    Wang, Chenhui
    Wang, Kejun
    Yin, Chaoqun
    Zhao, Cong
    Xu, Tao
    Zhang, Xinyi
    Huang, Ziqiang
    Liu, Meichen
    Yang, Tao
    NEURAL COMPUTING & APPLICATIONS, 2021, 33 (14) : 8669 - 8685
  • [29] An Assessment of In-the-Wild Datasets for Multimodal Emotion Recognition
    Aguilera, Ana
    Mellado, Diego
    Rojas, Felipe
    SENSORS, 2023, 23 (11)
  • [30] Multi-cue fusion for emotion recognition in the wild
    Yan, Jingwei
    Zheng, Wenming
    Cui, Zhen
    Tang, Chuangao
    Zhang, Tong
    Zong, Yuan
    NEUROCOMPUTING, 2018, 309 : 27 - 35