conference-paper

Generating gender-ambiguous voices for privacy-preserving speech recognition

  • Interspeech 2022
Research footprint

At a glance

Citations
13
References
29
Comments
0
Paper overview

Abstract

Our voice encodes a uniquely identifiable pattern which can be used to infer private attributes, such as gender or identity, that an individual might wish not to reveal when using a speech recognition service.To prevent attribute inference attacks alongside speech recognition tasks, we present a generative adversarial network, GenGAN, that synthesises voices that conceal the gender or identity of a speaker.The proposed network includes a generator with a U-Net architecture that learns to fool a discriminator.We condition the generator only on gender information and use an adversarial loss between signal distortion and privacy preservation.We show that GenGAN improves the tradeoff between privacy and utility compared to privacy-preserving representation learning methods that consider gender information as a sensitive attribute to protect.

Record transparency

Publication details

DOI
10.21437/interspeech.2022-11322
OpenAlex
W4297841750
Document type
conference-paper
Language
EN
Source
Interspeech 2022
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.