Evaluating Voice Conversion-based Privacy Protection against Informed Attackers

Speech data conveys sensitive speaker attributes like identity or accent. With a small amount of found data, such attributes can be inferred and exploited for malicious purposes: voice cloning, spoofing, etc. Anonymization aims to make the data unlinkable, i.e., ensure that no utterance can be linked to its original speaker. In this paper, we investigate anonymization methods based on voice conversion. In contrast to prior work, we argue that various linkage attacks can be designed depending on the attackers' knowledge about the anonymization scheme. We compare two frequency warping-based conversion methods and a deep learning based method in three attack scenarios. The utility of converted speech is measured via the word error rate achieved by automatic speech recognition, while privacy protection is assessed by the increase in equal error rate achieved by state-of-the-art i-vector or x-vector based speaker verification. Our results show that voice conversion schemes are unable to effectively protect against an attacker that has extensive knowledge of the type of conversion and how it has been applied, but may provide some protection against less knowledgeable attackers.

Mots clés

linkage attack privacy speaker verification attacker speech recognition voice conversion

Domaines

Apprentissage [cs.LG] Informatique et langage [cs.CL]

Fichier principal

ppvc_final.pdf (427.37 Ko)

Origine : Fichiers produits par l'(les) auteur(s)

Brij Mohan Lal Srivastava : Connectez-vous pour contacter le contributeur

https://inria.hal.science/hal-02355115

Soumis le : jeudi 13 février 2020-19:26:44

Dernière modification le : jeudi 1 février 2024-10:05:38

Archivage à long terme le : jeudi 14 mai 2020-18:21:36

Dates et versions

hal-02355115 , version 1 (08-11-2019)

hal-02355115 , version 2 (13-02-2020)

Identifiants

HAL Id : hal-02355115 , version 2

Citer

Brij Mohan Lal Srivastava, Nathalie Vauquier, Md Sahidullah, Aurélien Bellet, Marc Tommasi, et al.. Evaluating Voice Conversion-based Privacy Protection against Informed Attackers. ICASSP 2020 - 45th International Conference on Acoustics, Speech, and Signal Processing, IEEE Signal Processing Society, May 2020, Barcelona, Spain. pp.2802-2806. ⟨hal-02355115v2⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

UNIV-RENNES1 CNRS INRIA IRISA GRID5000 CRISTAL UNIV-LORRAINE INRIA2 CRISTAL-MAGNET LORIA LORIA-NLPKD UR1-MATH-STIC UR1-UFR-ISTIC UNIV-RENNES UNIV-LILLE SILECS ANR UR1-MATH-NUM

322 Consultations

543 Téléchargements