Title | ||
---|---|---|
On The Development Of Matched And Mismatched Italian Children'S Speech Recognition Systems |
Abstract | ||
---|---|---|
While at least read speech corpora are available for Italian children's speech research, there exist many languages which completely lack children's speech corpora. We propose that learning statistical mappings between the adult and child acoustic space using existing adult/children corpora may provide a future direction for generating children's models for such data deficient languages. In this work the recent advances in the development of the SONIC Italian children's speech recognition system will be described. This work, completing a previous one developed in the past, was conducted with the specific goals of integrating the newly trained children's speech recognition models into the Italian version of the Colorado Literacy Tutor platform. Specifically, children's speech recognition research for Italian was conducted using the complete training and test set of the FBK (ex ITC-irst) Italian Children's Speech Corpus (Child It). Using the University of Colorado SONIC LVSR system, we demonstrate a phonetic recognition error rate of 12,0% for a system which incorporates Vocal Tract Length Normalization (VTLN), Speaker-Adaptive Trained phonetic models, as well as unsupervised Structural MAP Linear Regression (SMAPLR). |
Year | Venue | Keywords |
---|---|---|
2009 | INTERSPEECH 2009: 10TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION 2009, VOLS 1-5 | children, ASR, Italian, adaptation |
Field | DocType | Citations |
Literacy,Speech corpus,TUTOR,Normalization (statistics),Computer science,Word error rate,Speech recognition,Acoustic space,Vocal tract,Test set | Conference | 3 |
PageRank | References | Authors |
0.40 | 10 | 1 |
Name | Order | Citations | PageRank |
---|---|---|---|
Piero Cosi | 1 | 224 | 43.27 |