Data-Driven and Feedback Based Spectro-Temporal Features for Speech Recognition

Sivaram G S V S<sup>*</sup>; Nemala Sridhar Krishna; Mesgarani Nima; Hermansky Hynek

doi:10.1109/LSP.2010.2079930

摘要

This paper proposes novel data-driven and feedback based discriminative spectro-temporal filters for feature extraction in automatic speech recognition (ASR). Initially a first set of spectro-temporal filters are designed to separate each phoneme from the rest of the phonemes. A hybrid Hidden Markov Model/Multilayer Perceptron (HMM/MLP) phoneme recognition system is trained on the features derived using these filters. As a feedback to the feature extraction stage, top confusions of this system are identified, and a second set of filters are designed specifically to address these confusions. Phoneme recognition experiments on TIMIT show that the features derived from the combined set of discriminative filters outperform conventional speech recognition features, and also contain significant complementary information.

出版日期2010-11

全文

访问全文

收藏分享被引(5) 浏览

更新时间：2018-01-19 14:39

Data-Driven and Feedback Based Spectro-Temporal Features for Speech Recognition

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友