Korean articulatory speech synthesis using physical vocal tract model.
Huynh Van Luong, Jong-Myon Kim, Cheol Hong Kim
Abstract
Huynh Van Luong, Jong-Myon Kim, Cheol Hong Kim
Abstract
Artificial vocal tract models provide the support of learning a second language and the therapy of speech disorders. Moreover, phonetic education and research can benefit from articulatory speech synthesis. Articulatory speech synthesis models are constructed by the source-filter model of the human vocal tract. In this study, we generated a Korean articulatory speech synthesis model using Artisynth [Fels et al., ISSP, 419–426 (2006)], which is a 3-D biomechanical open-source simulation platform. As the origin of the Korean language, it has 10 basic vowel phonemes and 11 complicated vowels in which some vowels can be rounded and unrounded such as /eu/, /yeo/, /wae/, etc. To synthesize these specific vowels, we created a new physical vocal tract model, which interconnects to form a complete integrated biomechanical system. The created model efficiently supports recording the Korean vowel sounds and linguistic analysis based on the linear prediction model. As a result, parameters of the glottis and controllable vocal tract filter are automatically evaluated. The acoustic quality of the synthesizer for Korean vowels is comparable with that of the existing commercial speech synthesis systems such as concatenation synthesizers [Donovan (1996)] and [Hamza (2000)]. [Work supported by the MKE, Korea, under the ITRC supervised by IITA (IITA-2008-(C1090-0801-0039)).]
A significance statement is not available in the OpenAlex record.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Artificial vocal tract models provide the support of learning a second language and the therapy of speech disorders. Moreover, phonetic education and research can benefit from articulatory speech synthesis. Articulatory speech synthesis models are constructed by the source-filter model of the human vocal tract. In this study, we generated a Korean articulatory speech synthesis model using Artisynth [Fels et al., ISSP, 419–426 (2006)], which is a 3-D biomechanical open-source simulation platform. As the origin of the Korean language, it has 10 basic vowel phonemes and 11 complicated vowels in which some vowels can be rounded and unrounded such as /eu/, /yeo/, /wae/, etc. To synthesize these specific vowels, we created a new physical vocal tract model, which interconnects to form a complete integrated biomechanical system. The created model efficiently supports recording the Korean vowel sounds and linguistic analysis based on the linear prediction model. As a result, parameters of the glottis and controllable vocal tract filter are automatically evaluated. The acoustic quality of the synthesizer for Korean vowels is comparable with that of the existing commercial speech synthesis systems such as concatenation synthesizers [Donovan (1996)] and [Hamza (2000)]. [Work supported by the MKE, Korea, under the ITRC supervised by IITA (IITA-2008-(C1090-0801-0039)).]
Key concepts: Vocal tract, Speech synthesis, Computer science, Speech recognition, Concatenation (mathematics), Vowel, Speech production, Glottis