Source-dependent variable rate speech coding below 3 KBPS
M. Stefanovic, A.M. Kondoz
Abstract
M. Stefanovic, A.M. Kondoz
Abstract
This work addresses the need for high quality, very low bit rate speech compression algorithms that can be utilised in many forthcoming multimedia applications. The speech coding algorithm proposed in this paper is a variable rate system based on the adaptive source-driven frame length scheme. Maximum speech compression is achieved for long-term steady-state speech and nonspeech (silence and unvoiced) conditions. In addition, shorter frame sizes are used to code those difficult-tomodel speech transitions, thus improving the overall perceptual quality when compared with traditional fixed rate schemes. This codec may be used for implementing various voice communication systems, such as Voice Store and Forward systems and digital answering machines, or to augment low bit rate integrated digital packet networks.
A significance statement is not available in the OpenAlex record.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
This work addresses the need for high quality, very low bit rate speech compression algorithms that can be utilised in many forthcoming multimedia applications. The speech coding algorithm proposed in this paper is a variable rate system based on the adaptive source-driven frame length scheme. Maximum speech compression is achieved for long-term steady-state speech and nonspeech (silence and unvoiced) conditions. In addition, shorter frame sizes are used to code those difficult-tomodel speech transitions, thus improving the overall perceptual quality when compared with traditional fixed rate schemes. This codec may be used for implementing various voice communication systems, such as Voice Store and Forward systems and digital answering machines, or to augment low bit rate integrated digital packet networks.
Key concepts: Codec2, Computer science, Speech coding, Adaptive Multi-Rate audio codec, Full Rate, Voice activity detection, Codec, Linear predictive coding