Phoneme based text-to-speech synthesis system
Ichiro Mikuni, Kazutoku Ohta
Abstract
Ichiro Mikuni, Kazutoku Ohta
Abstract
A system for speech synthesis by rule is described employing pseudo-phonemes as phonetic units. After scanning the system construction, the problems of selection and concatenation of phonetic units are discussed. In Japanese speech synthesis systems, the naturalness of synthesized speech depends mainly on the generated pitch contour and spectrum transitions between phonemes. To avoid the problem of coarticulation and to concentrate efforts on developing a set of prosodic rules, we propose a system that generates the spectrum parameter sequences by concatenating the spectrum parameters of pseudo-phonemes which suffer coarticulation effects from preceding and succeeding phonemes. Our system consists of a syntactic analyzer, a prosodic and phonemic symbol generator, a pseudo-phoneme concatenator, a prosodic parameter generator, and an LPC synthesizer.
OpenAlex reports 5 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
A system for speech synthesis by rule is described employing pseudo-phonemes as phonetic units. After scanning the system construction, the problems of selection and concatenation of phonetic units are discussed. In Japanese speech synthesis systems, the naturalness of synthesized speech depends mainly on the generated pitch contour and spectrum transitions between phonemes. To avoid the problem of coarticulation and to concentrate efforts on developing a set of prosodic rules, we propose a system that generates the spectrum parameter sequences by concatenating the spectrum parameters of pseudo-phonemes which suffer coarticulation effects from preceding and succeeding phonemes. Our system consists of a syntactic analyzer, a prosodic and phonemic symbol generator, a pseudo-phoneme concatenator, a prosodic parameter generator, and an LPC synthesizer.
Key concepts: Coarticulation, Speech synthesis, Concatenation (mathematics), Computer science, Speech recognition, Naturalness, Generator (circuit theory), Pitch contour