Skip to Main content Skip to Navigation
New interface
Journal articles

A Variational EM Algorithm for the Separation of Time-Varying Convolutive Audio Mixtures

Dionyssos Kounades-Bastian 1 Laurent Girin 2, 1 Xavier Alameda-Pineda 3 Sharon Gannot 4 Radu Horaud 1 
1 PERCEPTION - Interpretation and Modelling of Images and Videos
Inria Grenoble - Rhône-Alpes, Grenoble INP - Institut polytechnique de Grenoble - Grenoble Institute of Technology, LJK - Laboratoire Jean Kuntzmann
Abstract : This paper addresses the problem of separating audio sources from time-varying convolutive mixtures. We propose a probabilistic framework based on the local complex-Gaussian model combined with non-negative matrix factorization. The time-varying mixing filters are modeled by a continuous temporal stochastic process. We present a variational expectation-maximization (VEM) algorithm that employs a Kalman smoother to estimate the time-varying mixing matrix, and that jointly estimate the source parameters. The sound sources are then separated by Wiener filters constructed with the estimators provided by the VEM algorithm. Extensive experiments on simulated data show that the proposed method outperforms a block-wise version of a state-of-the-art baseline method.
Complete list of metadata

Cited literature [56 references]  Display  Hide  Download
Contributor : Perception team Connect in order to contact the contributor
Submitted on : Tuesday, April 12, 2016 - 6:09:47 PM
Last modification on : Tuesday, October 25, 2022 - 4:19:27 PM
Long-term archiving on: : Tuesday, November 15, 2016 - 2:06:18 AM


Files produced by the author(s)



Dionyssos Kounades-Bastian, Laurent Girin, Xavier Alameda-Pineda, Sharon Gannot, Radu Horaud. A Variational EM Algorithm for the Separation of Time-Varying Convolutive Audio Mixtures. IEEE/ACM Transactions on Audio, Speech and Language Processing, 2016, 24 (8), pp.1408-1423. ⟨10.1109/TASLP.2016.2554286⟩. ⟨hal-01301762⟩



Record views


Files downloads