Coding-based informed source separation: Nonnegative tensor factorization approach

Alexey Ozerov; Antoine Liutkus; Roland Badeau; Gael Richard

doi:10.1109/TASL.2013.2260153

Article Dans Une Revue IEEE Transactions on Audio, Speech and Language Processing Année : 2013

Coding-based informed source separation: Nonnegative tensor factorization approach

(1) , (2) , (2) , (2, 3)

1
2
3

Alexey Ozerov

Fonction : Auteur
PersonId : 930358

Technicolor R & I [Cesson Sévigné]

Antoine Liutkus

Fonction : Auteur
PersonId : 2740
IdHAL : antoine-liutkus
ORCID : 0000-0002-3458-6498
IdRef : 167600419

Laboratoire Traitement et Communication de l'Information

Roland Badeau

Fonction : Auteur
PersonId : 1121
IdHAL : rbadeau
ORCID : 0000-0002-9630-6877
IdRef : 106938134

Laboratoire Traitement et Communication de l'Information

Gael Richard

Fonction : Auteur
PersonId : 14146
IdHAL : gael-richard
IdRef : 094977208

Laboratoire Traitement et Communication de l'Information

Département Traitement du Signal et des Images

Résumé

Informed source separation (ISS) aims at reliably recovering sources from a mixture. To this purpose, it relies on the assumption that the original sources are available during an encoding stage. Given both sources and mixture, a sideinformation may be computed and transmitted along with the mixture, whereas the original sources are not available any longer. During a decoding stage, both mixture and side-information are processed to recover the sources. ISS is motivated by a number of specific applications including active listening and remixing of music, karaoke, audio gaming, etc. Most ISS techniques proposed so far rely on a source separation strategy and cannot achieve better results than oracle estimators. In this study, we introduce Coding-based ISS (CISS) and draw the connection between ISS and source coding. CISS amounts to encode the sources using not only a model as in source coding but also the observation of the mixture. This strategy has several advantages over conventional ISS methods. First, it can reach any quality, provided sufficient bandwidth is available as in source coding. Second, it makes use of the mixture in order to reduce the bitrate required to transmit the sources, as in classical ISS. Furthermore, we introduce Nonnegative Tensor Factorization as a very efficient model for CISS and report rate-distortion results that strongly outperform the state of the art.

Mots clés

Informed source separation spatial audio object coding source coding constrained entropy quantization probabilistic model nonnegative tensor factorization

Domaines

Traitement du signal et de l'image [eess.SP] Traitement du signal et de l'image [eess.SP]

Fichier principal

Ozerov_et_al_IEEE_TASLP-v12.pdf (2.16 Mo)

Origine : Fichiers produits par l'(les) auteur(s)

Alexey Ozerov : Connectez-vous pour contacter le contributeur

https://inria.hal.science/hal-00869603

Soumis le : jeudi 3 octobre 2013-16:58:50

Dernière modification le : lundi 9 octobre 2023-12:49:39

Archivage à long terme le : samedi 4 janvier 2014-07:45:21

Dates et versions

hal-00869603 , version 1 (03-10-2013)

Identifiants

HAL Id : hal-00869603 , version 1
DOI : 10.1109/TASL.2013.2260153

Citer

Alexey Ozerov, Antoine Liutkus, Roland Badeau, Gael Richard. Coding-based informed source separation: Nonnegative tensor factorization approach. IEEE Transactions on Audio, Speech and Language Processing, 2013, 21 (8), pp.1699-1712. ⟨10.1109/TASL.2013.2260153⟩. ⟨hal-00869603⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

INSTITUT-TELECOM CNRS PARISTECH LTCI IDS S2A

333 Consultations

501 Téléchargements

Coding-based informed source separation: Nonnegative tensor factorization approach

Résumé

Mots clés

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager