ملف الباحث

Sabato Marco Siniscalchi

3 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition

    2024 · arXiv (Cornell University)

    Recent studies have successfully shown that large language models (LLMs) can be successfully used for generative error correction (GER) on top of the automatic speech recognition (ASR) output. Specifically, an LLM is utilized to carry …

  2. Boosting End-to-End Multilingual Phoneme Recognition Through Exploiting Universal Speech Attributes Constraints

    2024

    We propose a first step toward multilingual end-to-end automatic speech recognition (ASR) by integrating knowledge about speech articulators. The key idea is to leverage a rich set of fundamental units that can be defined "universally" …

  3. Improving non-native mispronunciation detection and enriching diagnostic feedback with DNN-based speech attribute modeling

    2016

    We propose the use of speech attributes, such as voicing and aspiration, to address two key research issues in computer assisted pronunciation training (CAPT) for L2 learners, namely detecting mispronunciation and providing diagnostic feedback. To …