Hainan Xu
4 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
An Asynchronous WFST-Based Decoder For Automatic Speech Recognition
2021 · arXiv (Cornell University)
We introduce asynchronous dynamic decoder, which adopts an efficient A* algorithm to incorporate big language models in the one-pass decoding for large vocabulary continuous speech recognition. Unlike standard one-pass decoding with on-the-fly composition decoder which …
-
Pronunciation and silence probability modeling for ASR
2015
In this paper we evaluate the WER improvement from modeling pronunciation probabilities and word-specific silence probabilities in speech recognition. We do this in the context of Finite State Transducer (FST)-based decoding, where pronunciation and silence …
-
A Pruned Rnnlm Lattice-Rescoring Algorithm for Automatic Speech Recognition
2018
Lattice-rescoring is a common approach to take advantage of recurrent neural language models in ASR, where a word-lattice is generated from 1st-pass decoding and the lattice is then rescored with a neural model, and ann-gram …
-
Saliency-driven Word Alignment Interpretation for Neural Machine Translation
2019
Despite their original goal to jointly learn to align and translate, Neural Machine Translation (NMT) models, especially Transformer, are often perceived as not learning interpretable word alignments. In this paper, we show that NMT models …