Zhizheng Wu
8 papers in the PaperMetrix corpus
Papers by this author
-
Investigating gated recurrent neural networks for speech synthesis
2016 · arXiv (Cornell University)
Recently, recurrent neural networks (RNNs) as powerful sequence models have re-emerged as a potential acoustic model for statistical parametric speech synthesis (SPSS). The long short-term memory (LSTM) architecture is particularly attractive because it addresses the …
-
Spoofing detection under noisy conditions: a preliminary investigation and an initial database
2016 · arXiv (Cornell University)
Spoofing detection for automatic speaker verification (ASV), which is to discriminate between live speech and attacks, has received increasing attentions recently. However, all the previous studies have been done on the clean data without significant …
-
Listening test materials for "A study of speaker adaptation for DNN-based speech synthesis"
2015 · Edinburgh Research Explorer (University of Edinburgh)
The dataset contains the testing stimuli and listeners' MUSHRA test responses for the Interspeech 2015 paper, "A study of speaker adaptation for DNN-based speech synthesis". In this paper, we conduct an experimental analysis of speaker …
-
Investigating gated recurrent networks for speech synthesis
2016
Recently, recurrent neural networks (RNNs) as powerful sequence models have re-emerged as a potential acoustic model for statistical parametric speech synthesis (SPSS). The long short-term memory (LSTM) architecture is particularly attractive because it addresses the …
-
SpMis: An Investigation of Synthetic Spoken Misinformation Detection
2024 · arXiv (Cornell University)
In recent years, speech generation technology has advanced rapidly, fueled by generative models and large-scale training techniques. While these developments have enabled the production of high-quality synthetic speech, they have also raised concerns about the …
-
Deep neural networks employing Multi-Task Learning and stacked bottleneck features for speech synthesis
2015
Deep neural networks (DNNs) use a cascade of hidden representations to enable the learning of complex mappings from input to output features. They are able to learn the complex mapping from text-based linguistic features to …
-
ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge
2015
An increasing number of independent studies have con-firmed the vulnerability of automatic speaker verification (ASV) technology to spoofing. However, in comparison to that involving other biometric modalities, spoofing and countermea-sure research for ASV is still …
-
Merlin: An Open Source Neural Network Speech Synthesis System
2016
We introduce the Merlin speech synthesis toolkit for neural network-based speech synthesis. The system takes linguistic features as input, and employs neural networks to predict acoustic features, which are then passed to a vocoder to …