Wen-Chin Huang
3 papers in the PaperMetrix corpus
Papers by this author
-
Non-Autoregressive Sequence-To-Sequence Voice Conversion
2021
This paper proposes a novel voice conversion (VC) method based on non-autoregressive sequence-to-sequence (NAR-S2S) models. Inspired by the great success of NAR-S2S models such as FastSpeech in text-to-speech (TTS), we extend the FastSpeech2 model for …
-
On Prosody Modeling for ASR+TTS Based Voice Conversion
2021 · 2021 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU)
In voice conversion (VC), an approach showing promising results in the latest voice conversion challenge (VCC) 2020 is to first use an automatic speech recognition (ASR) model to transcribe the source speech into the underlying …
-
A Comparative Study of Self-supervised Speech Representation Based Voice Conversion
2022 · arXiv (Cornell University)
We present a large-scale comparative study of self-supervised speech representation (S3R)-based voice conversion (VC). In the context of recognition-synthesis VC, S3Rs are attractive owing to their potential to replace expensive supervised representations such as phonetic …