ملف الباحث

Mirco Ravanelli

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. SoundChoice: Grapheme-to-Phoneme Models with Semantic Disambiguation

    2022 · Interspeech 2022

    End-to-end speech synthesis models directly convert the input characters into an audio representation (e.g., spectrograms).Despite their impressive performance, such models have difficulty disambiguating the pronunciations of identically spelled words.To mitigate this issue, a separate Grapheme-to-Phoneme …