Junichi Yamagishi
5 papers in the PaperMetrix corpus
Papers by this author
-
A Deep Generative Architecture for Postfiltering in Statistical Parametric Speech Synthesis
2015 · IEEE/ACM Transactions on Audio Speech and Language Processing
The generated speech of hidden Markov model (HMM)-based statistical parametric speech synthesis still sounds “muffled.” One cause of this degradation in speech quality may be the loss of fine spectral structures. In this paper, we …
-
An autoregressive recurrent mixture density network for parametric speech synthesis
2017
Neural-network-based generative models, such as mixture density networks, are potential solutions for speech synthesis. In this paper we follow this path and propose a recurrent mixture density network that incorporates a trainable autoregressive model. An …
-
Initial investigation of an encoder-decoder end-to-end TTS framework using marginalization of monotonic hard latent alignments
2019 · arXiv (Cornell University)
End-to-end text-to-speech (TTS) synthesis is a method that directly converts input text to output acoustic features using a single network. A recent advance of end-to-end TTS is due to a key technique called attention mechanisms, …
-
Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019 · arXiv (Cornell University)
While speaker adaptation for end-to-end speech synthesis using speaker embeddings can produce good speaker similarity for speakers seen during training, there remains a gap for zero-shot adaptation to unseen speakers. We investigate multi-speaker modeling for …
-
ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge
2015
An increasing number of independent studies have con-firmed the vulnerability of automatic speaker verification (ASV) technology to spoofing. However, in comparison to that involving other biometric modalities, spoofing and countermea-sure research for ASV is still …