ملف الباحث

Ruizhi Li

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. M-vectors: Sub-band Based Energy Modulation Features for Multi-stream Automatic Speech Recognition

    2019

    In this paper, we propose a novel method to capture energy modulations from different frequency bands in speech into frame-level feature vectors, called Modulation-vectors or M-vectors, for use in Automatic Speech Recognition (ASR) systems. We …

  2. Multilingual sequence-to-sequence speech recognition: architecture, transfer learning, and language modeling

    2018 · arXiv (Cornell University)

    Sequence-to-sequence (seq2seq) approach for low-resource ASR is a relatively new direction in speech research. The approach benefits by performing model training without using lexicon and alignments. However, this poses a new problem of requiring more …