Lei He
7 papers in the PaperMetrix corpus
Papers by this author
-
Construction of Evaluation and Tracking System for Undergraduation Students
2015 · Guangzhou Chemical Industry
The information tracking system and its construction were introduced in the current paper. The ASP technology, access 2003, dreamweaver 8, and photoshop CS2 software were applied as tools to assemble the students information tracking system. …
-
Joint Pre-Training with Speech and Bilingual Text for Direct Speech to Speech Translation
2022 · arXiv (Cornell University)
Direct speech-to-speech translation (S2ST) is an attractive research topic with many advantages compared to cascaded S2ST. However, direct S2ST suffers from the data scarcity problem because the corpora from speech of the source language to …
-
Expressive Attribute-Based Proxy Signature Scheme for UAV Networks
2025 · Sensors
Unmanned aerial vehicle (UAV) networks have become an essential component of modern civilian and military infrastructures. However, the communication channels between UAVs and their control entities remain vulnerable to spoofing and message tampering attacks. Although …
-
Part-of-Speech Tagging with Bidirectional Long Short-Term Memory Recurrent Neural Network
2015 · arXiv (Cornell University)
Bidirectional Long Short-Term Memory Recurrent Neural Network (BLSTM-RNN) has been shown to be very effective for tagging sequential data, e.g. speech utterances or handwritten documents. While word embedding has been demoed as a powerful representation …
-
A Unified Tagging Solution: Bidirectional LSTM Recurrent Neural Network with Word Embedding
2015 · arXiv (Cornell University)
Bidirectional Long Short-Term Memory Recurrent Neural Network (BLSTM-RNN) has been shown to be very effective for modeling and predicting sequential data, e.g. speech utterances or handwritten documents. In this study, we propose to use BLSTM-RNN …
-
Speaker and language factorization in DNN-based TTS synthesis
2016
We have successfully proposed to use multi-speaker modelling in DNN-based TTS synthesis for improved voice quality with limited available data from a speaker. In this paper, we propose a new speaker and language factorized DNN, …
-
Learning Latent Representations for Style Control and Transfer in End-to-end Speech Synthesis
2019
In this paper, we introduce the Variational Autoencoder (VAE) to an end-to-end speech synthesis model, to learn the latent representation of speaking styles in an unsupervised manner. The style representation learned through VAE shows good …