ملف الباحث
Xiang Yin
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
A Unified Sequence-to-Sequence Front-End Model for Mandarin Text-to-Speech Synthesis
2020
In Mandarin text-to-speech (TTS) system, the front-end text processing module significantly influences the intelligibility and naturalness of synthesized speech. Building a typical pipeline-based front-end which consists of multiple individual components requires extensive efforts. In this …
-
RefXVC: Cross-Lingual Voice Conversion With Enhanced Reference Leveraging
2024 · IEEE/ACM Transactions on Audio Speech and Language Processing
This paper proposes RefXVC, a method for cross-lingual voice conversion (XVC) that leverages reference information to improve conversion performance. Previous XVC works generally take an average speaker embedding to condition the speaker identity, which does …