Significance of Phase-based Features for Person Recognition Using Humming

被引:0
|
作者
Sailor, Hardik B. [1 ]
Madhavi, Maulik C. [1 ]
Patil, Hemant A. [1 ]
机构
[1] DA IICT, Gandhinagar, Gujarat, India
关键词
Humming; person recognition; Modified Group Delay Function (MODGDF); polynomial classifier;
D O I
10.1145/2708463.2709035
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper presents use of hum of a person as a biometric cue for person recognition task. Mel Frequency Cepstral Coefficients (MFCC) is found to be state-of-the-art in the voice biometrics. However, it is magnitude-based features and ignores the phase information. This paper shows the effectiveness of phase-based information extracted via Modified Group Delay Function (MODGDF). The features developed by Mel filtering of MODGDF spectrum are called Modified Group Delay Cepstral Coefficients (MGDCC). The paper demonstrates two types of fusion strategies, viz., score-level and feature-level. The experimental results show that overall performance is improved by 3 % if a score-level fusion is employed between MFCC and MGDCC and 19.78 % by feature-level fusion in terms of % Equal Error Rate (EER). These experimental results clearly indicate that incorporating phase information along with magnitude-based features can effectively captures person-specific characteristics in humming.
引用
收藏
页码:99 / 103
页数:5
相关论文
共 50 条
  • [1] Person Recognition using Humming, Singing and Speech
    Patil, Hemant A.
    Madhavi, Maulik C.
    Chhayani, Nirav H.
    2012 INTERNATIONAL CONFERENCE ON ASIAN LANGUAGE PROCESSING (IALP 2012), 2012, : 149 - 152
  • [2] Exploitation of Phase-Based Features for Whispered Speech Emotion Recognition
    Deng, Jun
    Xu, Xinzhou
    Zhang, Zixing
    Fruehholz, Sascha
    Schuller, Bjoern
    IEEE ACCESS, 2016, 4 : 4299 - 4309
  • [3] Combining evidences from magnitude and phase information using VTEO for person recognition using humming
    Patil, Hemant A.
    Madhavi, Maulik C.
    COMPUTER SPEECH AND LANGUAGE, 2018, 52 : 225 - 256
  • [4] Development of Corpora for Person Recognition using Humming, Singing and Speech
    Chhayani, Nirav H.
    Patil, Hemant A.
    2013 INTERNATIONAL CONFERENCE ORIENTAL COCOSDA HELD JOINTLY WITH 2013 CONFERENCE ON ASIAN SPOKEN LANGUAGE RESEARCH AND EVALUATION (O-COCOSDA/CASLRE), 2013,
  • [5] Phase-based local features
    Carneiro, G
    Jepson, AD
    COMPUTER VISON - ECCV 2002, PT 1, 2002, 2350 : 282 - 296
  • [6] A palmprint recognition algorithm using phase-based image matching
    Ito, Koichi
    Aoki, Takafumi
    Nakajima, Hiroshi
    Kobayashi, Koji
    Higuchi, Tatsuo
    2006 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP 2006, PROCEEDINGS, 2006, : 2669 - +
  • [7] A PALMPRINT RECOGNITION ALGORITHM USING PHASE-BASED CORRESPONDENCE MATCHING
    Ito, Koichi
    Iitsuka, Satoshi
    Aoki, Takafumi
    2009 16TH IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, VOLS 1-6, 2009, : 1977 - 1980
  • [8] An iris recognition system using phase-based image matching
    Miyazawa, Kazuyuki
    Ito, Koichi
    Aoki, Takafumi
    Kobayashi, Koji
    Katsumata, Atsushi
    2006 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP 2006, PROCEEDINGS, 2006, : 325 - +
  • [9] A phase-based iris recognition algorithm
    Miyazawa, K
    Ito, K
    Aoki, T
    Kobayashi, K
    Nakajima, H
    ADVANCES IN BIOMETRICS, PROCEEDINGS, 2006, 3832 : 356 - 365
  • [10] Static and dynamic information derived from source and system features for person recognition from humming
    Patil, Hemant
    Madhavi, Maulik
    Parhi, Keshab
    INTERNATIONAL JOURNAL OF SPEECH TECHNOLOGY, 2012, 15 (03) : 393 - 406