ملف الباحث
Yusuke Yasuda
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Initial investigation of an encoder-decoder end-to-end TTS framework using marginalization of monotonic hard latent alignments
2019 · arXiv (Cornell University)
End-to-end text-to-speech (TTS) synthesis is a method that directly converts input text to output acoustic features using a single network. A recent advance of end-to-end TTS is due to a key technique called attention mechanisms, …
-
Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019 · arXiv (Cornell University)
While speaker adaptation for end-to-end speech synthesis using speaker embeddings can produce good speaker similarity for speakers seen during training, there remains a gap for zero-shot adaptation to unseen speakers. We investigate multi-speaker modeling for …