Complete recognition of continuous Mandarin speech for Chinese language with very large vocabulary but limited training data
Resource
Acoustics, Speech, and Signal Processing, 1995. ICASSP-95., 1995 International Conference on
Journal
International Conference on Acoustics, Speech, and Signal Processing, 1995. ICASSP-95
Pages
-
Date Issued
1995-05
Date
1995-05
Author(s)
DOI
N/A
Abstract
This paper presents the first known results for complete recognition of continuous Mandarin speech for Chinese language with very large vocabulary but very limited training data. Although some isolated-syllable-based or isolated-word-based large-vocabulary Mandarin speech recognition systems have been successfully developed, a continuous-speech-based system of this kind has never been reported before. For successful development of this system, several important techniques have been used, including acoustic modeling of a set of sub-syllabic models for base syllable recognition and another set of context-dependent models for tone recognition, a multiple candidate searching technique based on a concatenated syllable matching algorithm to synchronize base syllable and tone recognition, and a word-class-based Chinese language model for linguistic decoding. The best recognition accuracy achieved is 88.69% for finally decoded Chinese characters, with 88.69%, 91.57%, and 81.37% accuracy for base syllables, tones, and tonal syllables respectively.
SDGs
Type
journal article
File(s)![Thumbnail Image]()
Loading...
Name
00479273.pdf
Size
453.89 KB
Format
Adobe PDF
Checksum
(MD5):4d96e76762e31aabced86655b8e09f4a
