Chunlei Zhang
6 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
Semi-supervised Learning with Generative Adversarial Networks for Arabic Dialect Identification
2019
Dialect Identification (DID) refers to the process of identifying different dialects within the same language class. Compared with more general language identification (LID), DID is a more challenging task because of the substantial similarity between …
-
Towards Robust Speaker Verification with Target Speaker Enhancement
2021 · arXiv (Cornell University)
This paper proposes the target speaker enhancement based speaker verification network (TASE-SVNet), an all neural model that couples target speaker enhancement and speaker embedding extraction for robust speaker verification (SV). Specifically, an enrollment speaker conditioned …
-
Towards end-to-end Speaker Diarization with Generalized Neural Speaker Clustering
2022 · ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Speaker diarization consists of many components, e.g., front-end processing, speech activity detection (SAD), overlapped speech detection (OSD) and speaker segmentation/clustering. Conventionally, most of the involved components are separately developed and optimized. The resulting speaker diarization …
-
Research on the Construction Method of University Graduates Employment Quality Analysis Database
2022
research-article Share on Research on the Construction Method of University Graduates Employment Quality Analysis Database Authors: Chunlei Zhang Heilongjiang Bayi Agricultural University, China Heilongjiang Bayi Agricultural University, ChinaSearch about this author , Guoxin Han School …
-
Towards Improved Zero-shot Voice Conversion with Conditional DSVAE
2022 · arXiv (Cornell University)
Disentangling content and speaking style information is essential for zero-shot non-parallel voice conversion (VC). Our previous study investigated a novel framework with disentangled sequential variational autoencoder (DSVAE) as the backbone for information decomposition. We have …
-
Preference Alignment Improves Language Model-Based TTS
2025
Recent advancements in text-to-speech (TTS) have shown that language model (LM)-based systems offer competitive performance to their counterparts. Further optimization can be achieved through preference alignment algorithms, which adjust LMs to align with the preferences …