preprint وصول مفتوح

THCHS-30 : A Free Chinese Speech Corpus

  • arXiv (Cornell University)
  • Cornell University
Research footprint

At a glance

الاستشهادات
192
المراجع
12
Comments
0
Paper overview

Abstract

Speech data is crucially important for speech recognition research. There are quite some speech databases that can be purchased at prices that are reasonable for most research institutes. However, for young people who just start research activities or those who just gain initial interest in this direction, the cost for data is still an annoying barrier. We support the `free data' movement in speech recognition: research institutes (particularly supported by public funds) publish their data freely so that new researchers can obtain sufficient data to kick of their career. In this paper, we follow this trend and release a free Chinese speech database THCHS-30 that can be used to build a full- edged Chinese speech recognition system. We report the baseline system established with this database, including the performance under highly noisy conditions.

Record transparency

Publication details

DOI
10.48550/arxiv.1512.01882
OpenAlex
W2284628133
Document type
preprint
Language
EN
Source
arXiv (Cornell University)
Last metadata update
المجتمع

Comments

تسجيل الدخول للانضمام إلى النقاش.

  1. لا توجد تعليقات بعد. ابدأ النقاش.