ملف الباحث
Cheng-I Lai
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019 · arXiv (Cornell University)
While speaker adaptation for end-to-end speech synthesis using speaker embeddings can produce good speaker similarity for speakers seen during training, there remains a gap for zero-shot adaptation to unseen speakers. We investigate multi-speaker modeling for …
-
SUPERB: Speech Processing Universal PERformance Benchmark
2021
Self-supervised learning (SSL) has proven vital for advancing research in natural language processing (NLP) and computer vision (CV).The paradigm pretrains a shared model on large volumes of unlabeled data and achieves state-of-the-art (SOTA) for various …