A new method for mispronunciation detection using Support Vector Machine   based on Pronunciation Space Models

Wei Si<sup>*</sup>; Hu Guoping; Hu Yu; Wang Ren Hua

doi:10.1016/j.specom.2009.03.004

摘要

This paper presents two new ideas for text dependent mispronunciation detection. Firstly, mispronunciation detection is formulated as a classification problem to integrate various predictive features. A Support Vector Machine (SVM) is used as the classifier and the log-likelihood ratios between all the acoustic models and the model corresponding to the given text are employed as features for the classifier. Secondly, Pronunciation Space Models (PSMs) are proposed to enhance the discriminative capability of the acoustic models for pronunciation variations. In PSMs, each phone is modeled with several parallel acoustic models to represent pronunciation variations of that phone at different proficiency levels, and an unsupervised method is proposed for the construction of the PSMs. Experiments on a database consisting of more than 500,000 Mandarin syllables collected from 1335 Chinese speakers show that the proposed methods can significantly outperform the traditional posterior probability based method. The overall recall rates for the 13 most frequently mispronounced phones increase from 17.2%, 7.6% and 0% to 58.3%, 44.3% and 29.5% at three precision levels of 60%, 70%, and 80%, respectively. The improvement is also demonstrated by a subjective experiment with 30 subjects, in which 53.3% of the subjects think the proposed method is better than the traditional one and 23.3% of them think that the two methods are comparable.

出版日期2009-10
单位中国科学技术大学

全文

访问全文

收藏分享被引(56) 浏览

更新时间：2024-03-29 04:48

A new method for mispronunciation detection using Support Vector Machine based on Pronunciation Space Models

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友