ملف الباحث
Ruizhi Li
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
M-vectors: Sub-band Based Energy Modulation Features for Multi-stream Automatic Speech Recognition
2019
In this paper, we propose a novel method to capture energy modulations from different frequency bands in speech into frame-level feature vectors, called Modulation-vectors or M-vectors, for use in Automatic Speech Recognition (ASR) systems. We …
-
Multilingual sequence-to-sequence speech recognition: architecture, transfer learning, and language modeling
2018 · arXiv (Cornell University)
Sequence-to-sequence (seq2seq) approach for low-resource ASR is a relatively new direction in speech research. The approach benefits by performing model training without using lexicon and alignments. However, this poses a new problem of requiring more …