2003•Unpublished venueRequires access

Analysis-by-synthesis linear predictive speech coding at 2.4 kbit/s

F.F. Tzeng

Open publisher page 11 citations

Abstract

Novel 2.4-kb/s linear predictive speech coders based on the analysis-by-syntheses method are proposed. The introduction of a perceptually weighted distortion measure between the original speech and the reconstructed speech implicitly optimizes both the voiced/unvoiced decision and the pitch estimation/tracking. The coders are also shown to be more robust to background acoustic noises. The resultant speech quality is significantly enhanced by judicious parameter coding. Three excitation models are proposed and investigated. It is found that the model which selects the excitation signal from either a random sequence codebook or a pitch synthesizer produces the best perceived quality speech.>

About this research paper

What this paper is about

Novel 2.4-kb/s linear predictive speech coders based on the analysis-by-syntheses method are proposed. The introduction of a perceptually weighted distortion measure between the original speech and the reconstructed speech implicitly optimizes both the voiced/unvoiced decision and the pitch estimation/tracking. The coders are also shown to be more robust to background acoustic noises. The resultant speech quality is significantly enhanced by judicious parameter coding. Three excitation models are proposed and investigated. It is found that the model which selects the excitation signal from either a random sequence codebook or a pitch synthesizer produces the best perceived quality speech.>

Why it matters

OpenAlex reports 11 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Novel 2.4-kb/s linear predictive speech coders based on the analysis-by-syntheses method are proposed. The introduction of a perceptually weighted distortion measure between the original speech and the reconstructed speech implicitly optimizes both the voiced/unvoiced decision and the pitch estimation/tracking. The coders are also shown to be more robust to background acoustic noises. The resultant speech quality is significantly enhanced by judicious parameter coding. Three excitation models are proposed and investigated. It is found that the model which selects the excitation signal from either a random sequence codebook or a pitch synthesizer produces the best perceived quality speech.>

Key concepts: Linear predictive coding, Codebook, Speech recognition, Speech coding, Codec2, Linear prediction, Computer science, Speech synthesis

Related papers

Back to paper searchBrowse research topicsOriginal source
Analysis-by-synthesis linear predictive speech coding at 2.4 kbit/s — Research Paper | ScholarLens