SELECTION OF WAVEFORM UNITS FOR CORPUS-BASED MANDARIN SPEECH SYNTHESIS BASED ON DECISION TREES AND PROSODIC MODIFICATION COSTS

Chou, Fu Chiang; Tseng, Chiu Yu; LIN-SHAN LEE

SELECTION OF WAVEFORM UNITS FOR CORPUS-BASED MANDARIN SPEECH SYNTHESIS BASED ON DECISION TREES AND PROSODIC MODIFICATION COSTS

Journal

6th European Conference on Speech Communication and Technology, EUROSPEECH 1999

Date Issued

1999-01-01

Author(s)

Chou, Fu Chiang

Tseng, Chiu Yu

LIN-SHAN LEE

URI

https://scholars.lib.ntu.edu.tw/handle/123456789/631219

URL

https://api.elsevier.com/content/abstract/scopus_id/85135272125

Abstract

A lazy decision tree approach is described in this paper for the selection of concatenative units for Mandarin speech synthesis. The concept is not to induce a concise hypothesis from a given training data; the selection is delayed until a test instance is given. Thus we can construct the “best” decision tree for each selection. The selection of waveform units is guided with other simultaneously selected prosodic parameters. The selected waveform units can be directly concatenated into output speech or modified with the selected prosodic parameters. This method can produce synthetic speech that sounds very natural and resembles the acoustic and prosodic characteristics of the original speaker. A Mandarin speech synthesizer is described in this paper. However, most of the technologies are language independent and can be extended to a multilingual system.

Type

conference paper

SELECTION OF WAVEFORM UNITS FOR CORPUS-BASED MANDARIN SPEECH SYNTHESIS BASED ON DECISION TREES AND PROSODIC MODIFICATION COSTS

關於 (About)

聯絡資訊 (Contact Us)

相關網站 (Useful Links)

關於開放取用 (Open Access, OA)

出版社期刊論文授權政策 (Copyright)

使用說明 (Instructions)

登入說明 (Sign-in)

匯入著作 (Submission)