Researcher profile

Yuxuan Wang

7 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. A Unified Sequence-to-Sequence Front-End Model for Mandarin Text-to-Speech Synthesis

    2020

    In Mandarin text-to-speech (TTS) system, the front-end text processing module significantly influences the intelligibility and naturalness of synthesized speech. Building a typical pipeline-based front-end which consists of multiple individual components requires extensive efforts. In this …

  2. Network-Level Adversaries in Federated Learning

    2022 · arXiv (Cornell University)

    Federated learning is a popular strategy for training models on distributed, sensitive data, while preserving data privacy. Prior work identified a range of security threats on federated learning protocols that poison the data or the …

  3. Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions

    2017 · arXiv (Cornell University)

    This paper describes Tacotron 2, a neural network architecture for speech synthesis directly from text. The system is composed of a recurrent sequence-to-sequence feature prediction network that maps character embeddings to mel-scale spectrograms, followed by …

  4. Towards Better

    2018 · Proceedings of the

    This paper describes our system (HIT-SCIR) submitted to the CoNLL 2018 shared task on Multilingual Parsing from Raw Text to Universal Dependencies. We base our submission on Stanford's winning system for the CoNLL 2017 shared …

  5. Tacotron: Towards End-to-End Speech Synthesis

    2017

    A text-to-speech synthesis system typically consists of multiple stages, such as a text analysis frontend, an acoustic model and an audio synthesis module.Building these components often requires extensive domain expertise and may contain brittle design …

  6. Natural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions

    2018

    This paper describes Tacotron 2, a neural network architecture for speech synthesis directly from text. The system is composed of a recurrent sequence-to-sequence feature prediction network that maps character embeddings to mel-scale spectrograms, followed by …

  7. Cross-Lingual BERT Transformation for Zero-Shot Dependency Parsing

    2019

    Yuxuan Wang, Wanxiang Che, Jiang Guo, Yijia Liu, Ting Liu. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). 2019.