ملف الباحث
J. J. Miller
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Deep Voice: Real-time Neural Text-to-Speech
2017 · arXiv (Cornell University)
We present Deep Voice, a production-quality text-to-speech system constructed entirely from deep neural networks. Deep Voice lays the groundwork for truly end-to-end neural speech synthesis. The system comprises five major building blocks: a segmentation model …
-
Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence\n Learning
2017 · arXiv (Cornell University)
We present Deep Voice 3, a fully-convolutional attention-based neural\ntext-to-speech (TTS) system. Deep Voice 3 matches state-of-the-art neural\nspeech synthesis systems in naturalness while training ten times faster. We\nscale Deep Voice 3 to data set sizes unprecedented …