Zhengqi Wen
5 papers in the PaperMetrix corpus
Papers by this author
-
Self-Attention Transducers for End-to-End Speech Recognition
2019
Recurrent neural network transducers (RNN-T) have been successfully applied in end-to-end speech recognition. However, the recurrent structure makes it difficult for parallelization . In this paper, we propose a self-attention transducer (SA-T) for speech recognition. …
-
One In A Hundred: Select The Best Predicted Sequence from Numerous Candidates for Streaming Speech Recognition
2020 · arXiv (Cornell University)
The RNN-Transducers and improved attention-based encoder-decoder models are widely applied to streaming speech recognition. Compared with these two end-to-end models, the CTC model is more efficient in training and inference. However, it cannot capture the …
-
Decoupling Pronunciation and Language for End-to-end Code-switching Automatic Speech Recognition
2020 · arXiv (Cornell University)
Despite the recent significant advances witnessed in end-to-end (E2E) ASR system for code-switching, hunger for audio-text paired data limits the further improvement of the models' performance. In this paper, we propose a decoupled transformer model …
-
Transferring Personality Knowledge to Multimodal Sentiment Analysis
2024
Multimodal sentiment analysis systems have achieved remarkable success. However, significant challenges persist in tailoring sentiment analysis to individualized needs. Recognizing the pivotal role of personality traits in shaping emotional expression-where distinct personalities manifest emotions with …
-
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
2025 · arXiv (Cornell University)
Large language models (LLMs) have achieved significant progress across various domains, but their increasing scale results in high computational and memory costs. Recent studies have revealed that LLMs exhibit sparsity, providing the potential to reduce …