Rohit Prabhavalkar
3 papers in the PaperMetrix corpus
Papers by this author
-
From Audio to Semantics: Approaches to End-to-End Spoken Language Understanding
2018
Conventional spoken language understanding systems consist of two main components: an automatic speech recognition module that converts audio to a transcript, and a natural language understanding module that transforms the resulting text (or top N …
-
Lingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling
2019 · arXiv (Cornell University)
Lingvo is a Tensorflow framework offering a complete solution for collaborative deep learning research, with a particular focus towards sequence-to-sequence models. Lingvo models are composed of modular building blocks that are flexible and easily extensible, …
-
Exploring architectures, data and units for streaming end-to-end speech recognition with RNN-transducer
2017 · 2017 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU)
We investigate training end-to-end speech recognition models with the recurrent neural network transducer (RNN-T): a streaming, all-neural, sequence-to-sequence architecture which jointly learns acoustic and language model components from transcribed acoustic data. We explore various model …