Sabato Marco Siniscalchi
3 papers in the PaperMetrix corpus
Papers by this author
-
It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition
2024 · arXiv (Cornell University)
Recent studies have successfully shown that large language models (LLMs) can be successfully used for generative error correction (GER) on top of the automatic speech recognition (ASR) output. Specifically, an LLM is utilized to carry …
-
Boosting End-to-End Multilingual Phoneme Recognition Through Exploiting Universal Speech Attributes Constraints
2024
We propose a first step toward multilingual end-to-end automatic speech recognition (ASR) by integrating knowledge about speech articulators. The key idea is to leverage a rich set of fundamental units that can be defined "universally" …
-
Improving non-native mispronunciation detection and enriching diagnostic feedback with DNN-based speech attribute modeling
2016
We propose the use of speech attributes, such as voicing and aspiration, to address two key research issues in computer assisted pronunciation training (CAPT) for L2 learners, namely detecting mispronunciation and providing diagnostic feedback. To …