Higher order cepstral moment normalization (HOCMN) for robust speech recognition
Resource
Acoustics, Speech, and Signal Processing, 2004. Proceedings. (ICASSP '04). IEEE International Conference on
Journal
IEEE International Conference on Acoustics, Speech, and Signal Processing
Pages
197-200
Date Issued
2004-05
Date
2004-05
Author(s)
Hsu, Chang-Wen
DOI
1520-6149
Abstract
Cepstral mean subtraction (CMS) and cepstral normalization (CN) have been popularly used to normalize the first and the second moments of cepstral coefficients, and proved to be very helpful for robust speech recognition (Furui, S. 1981; Viikki, O. and Laurila, K., 1998). A unified formulation for higher order cepstral moment normalization (HOCMN) is developed by extending the concept of CMS and CN to orders much higher than three. A whole family of normalization techniques for different orders is thus proposed. Preliminary experimental results based on Aurora 2.0 showed that the recognition accuracy can be significantly improved with this approach under all noisy conditions. For example, HOCMN/sub (1,5,100)/ (normalization of the first, fifth and 100th order cepstral moments) is shown to offer an error rate reduction of 32.83% as compared to the conventional CN with a full-utterance processing interval, or an error rate reduction of 20.78% as compared to CN with a segmental processing interval.
Type
journal article
File(s)![Thumbnail Image]()
Loading...
Name
01325956.pdf
Size
278.14 KB
Format
Adobe PDF
Checksum
(MD5):0a4b116e738f3b9bfd34042267a447f2
