Speech enhancement models suited for speech recognition using composite source and wavelet decomposition model
作者
P.S. Rajakumar,S. Ravi,R. M. Suresh
标识
DOI:10.1109/icsip.2010.5697529
摘要
To compare the performance of two speech coders, it is necessary to have some indicator of the intelligibility and quality of the speech produced by each coder. The term intelligibility usually refers to whether the output speech is easily understandable, while the term quality is an indicator of how natural the speech sounds. It is possible for a coder to produce highly intelligible speech from low quality, in that the speech may sound very machine-like and the speaker is not identifiable. On the other hand, it is unlikely that unintelligible speech would be called high quality, but there are situations in which perceptually pleasing speech does not have high intelligibility. In this paper the most common measures of speech enhancement models suited for speech recognition with specific emphasis on wavelet based approach are presented.