SPARSE GAUSSIAN PROCESS AUDIO SOURCE SEPARATION USING SPECTRUM PRIORS IN THE TIME-DOMAIN

被引:0
|
作者
Alvarado, Pablo A. [1 ]
Alvarez, Mauricio A. [1 ,2 ]
Stowell, Dan [1 ]
机构
[1] Queen Mary Univ London, Ctr Digital Mus, London, England
[2] Univ Sheffield, Dept Comp Sci, Sheffield, S Yorkshire, England
来源
2019 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP) | 2019年
基金
英国工程与自然科学研究理事会;
关键词
Time-domain source separation; Gaussian processes; spectral mixture kernels; variational inference;
D O I
暂无
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
Gaussian process (GP) audio source separation is a time-domain approach that circumvents the inherent phase approximation issue of spectrogram based methods. Furthermore, through its kernel, GPs elegantly incorporate prior knowledge about the sources into the separation model. Despite these compelling advantages, the computational complexity of GP inference scales cubically with the number of audio samples. As a result, source separation GP models have been restricted to the analysis of short audio frames. We introduce an efficient application of GPs to time-domain audio source separation, without compromising performance. For this purpose, we used GP regression, together with spectral mixture kernels, and variational sparse GPs. We compared our method with LD-PSDTF (positive semi-definite tensor factorization), KL-NMF (Kullback-Leibler non-negative matrix factorization), and IS-NMF (Itakura-Saito NMF). Results show that the proposed method outperforms these techniques.
引用
收藏
页码:995 / 999
页数:5
相关论文
共 50 条
  • [1] TIME-DOMAIN AUDIO SOURCE SEPARATION BASED ON GAUSSIAN PROCESSES WITH DEEP KERNEL LEARNING
    Nugraha, Aditya Arie
    Di Carlo, Diego
    Bando, Yoshiaki
    Fontaine, Mathieu
    Yoshii, Kazuyoshi
    2023 IEEE WORKSHOP ON APPLICATIONS OF SIGNAL PROCESSING TO AUDIO AND ACOUSTICS, WASPAA, 2023,
  • [2] Time-Domain Blind Audio Source Separation using Advanced ICA Methods
    Koldovsky, Zbynek
    Tichavsky, Petr
    INTERSPEECH 2007: 8TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION, VOLS 1-4, 2007, : 1861 - +
  • [3] Time-domain blind audio source separation using advanced component clustering and reconstruction
    Koldovsky, Zbynek
    Tichavsky, Petr
    2008 HANDS-FREE SPEECH COMMUNICATION AND MICROPHONE ARRAYS, 2008, : 216 - +
  • [4] Time-Domain Audio Source Separation With Neural Networks Based on Multiresolution Analysis
    Nakamura, Tomohiko
    Kozuka, Shihori
    Saruwatari, Hiroshi
    IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2021, 29 : 1687 - 1701
  • [5] Unsupervised Audio Source Separation using Generative Priors
    Narayanaswamy, Vivek
    Thiagarajan, Jayaraman J.
    Anirudh, Rushil
    Spanias, Andreas
    INTERSPEECH 2020, 2020, : 2657 - 2661
  • [6] Spatial location priors for Gaussian model based reverberant audio source separation
    Duong, Ngoc Q. K.
    Vincent, Emmanuel
    Gribonval, Remi
    EURASIP JOURNAL ON ADVANCES IN SIGNAL PROCESSING, 2013,
  • [7] Spatial location priors for Gaussian model based reverberant audio source separation
    Ngoc Q K Duong
    Emmanuel Vincent
    Rémi Gribonval
    EURASIP Journal on Advances in Signal Processing, 2013
  • [8] PHASE SHIFTED BEDROSIAN FILTERBANK: AN INTERPRETABLE AUDIO FRONT-END FOR TIME-DOMAIN AUDIO SOURCE SEPARATION
    Mathieu, Felix
    Courtat, Thomas
    Richard, Gael
    Peeters, Geoffroy
    2022 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP), 2022, : 531 - 535
  • [9] ROBUST UNDERDETERMINED BLIND AUDIO SOURCE SEPARATION OF SPARSE SIGNALS IN THE TIME-FREQUENCY DOMAIN
    Sbai, Si Mohamed Aziz
    Aissa-El-Bey, Abdeldjalil
    Pastor, Dominique
    2011 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, 2011, : 3716 - 3719
  • [10] Fuzzy Clustering of Independent Components within Time-Domain Blind Audio Source Separation Method
    Malek, Jiri
    Koldovsky, Zbynek
    2011 10TH INTERNATIONAL WORKSHOP ON ELECTRONICS, CONTROL, MEASUREMENT AND SIGNALS (ECMS), 2011, : 44 - 49