Xu Li
4 papers in the PaperMetrix corpus
Papers by this author
-
Applying Multitask Learning to Acoustic-Phonemic Model for Mispronunciation Detection and Diagnosis in L2 English Speech
2018
For mispronunciation detection and diagnosis (MDD), nowadays approaches generally treat the phonemes in correct and mispronunciations as the same despite the fact they may actually carry different characteristics. Furthermore, serious data imbalance issue between correct …
-
Channel-wise Gated Res2Net: Towards Robust Detection of Synthetic Speech Attacks
2021 · arXiv (Cornell University)
Existing approaches for anti-spoofing in automatic speaker verification (ASV) still lack generalizability to unseen attacks. The Res2Net approach designs a residual-like connection between feature groups within one block, which increases the possible receptive fields and …
-
A Hierarchical Speaker Representation Framework for One-shot Singing Voice Conversion
2022 · Interspeech 2022
Typically, singing voice conversion (SVC) depends on an embedding vector, extracted from either a speaker lookup table (LUT) or a speaker recognition network (SRN), to model speaker identity.However, singing contains more expressive speaker characteristics than …
-
Enhancing the vocal range of single-speaker singing voice synthesis with melody-unsupervised pre-training
2023 · arXiv (Cornell University)
The single-speaker singing voice synthesis (SVS) usually underperforms at pitch values that are out of the singer's vocal range or associated with limited training samples. Based on our previous work, this work proposes a melody-unsupervised …