article
Constructing a speech audio–video corpus by aligning long segments of speech and text
Research footprint
At a glance
- الاستشهادات
- 2
- المراجع
- 12
- Comments
- 0
Paper overview
Abstract
A new algorithm for aligning text with speech audio signals having lengths of up to several hours is proposed. The algorithm allows its quality to be effectively evaluated. The requirements on the acoustic model are not very demanding. The algorithm can be used to design an audio–video course for learning the Russian language.
Record transparency
Publication details
- DOI
- 10.3103/s0278641917020030
- OpenAlex
- W2620660273
- Document type
- article
- Language
- EN
- Source
- Moscow University Computational Mathematics and Cybernetics
- Last metadata update
Comments
تسجيل الدخول للانضمام إلى النقاش.