2010Unpublished venueRequires access

Perceptual evaluation of speech enhancement

Mousa Al-Akhras, Khaled Daqrouq, Abdul Rahman Al-Qawasmi

Open publisher page 11 citations

Abstract

Speech enhancement is the process of de-noising a speech signal for improved quality and better intelligibility. Several speech enhancement methods have been proposed including: DWFM filtering, Donoho, Massart, and Kalman. To measure the performance of these filters, a speech evaluation method is needed. SNR is one of the most common methods for speech evaluation. The problem of SNR as a waveform speech evaluation is it is too general and can fit any type of signal, even non-speech signals. In this paper the performance of several speech enhancement methods is compared using both SNR and PESQ which is an evaluation method that has been proposed by the ITU-T for speech-specific quality evaluation. The speech sources used during the experiments are artificial voices produced in ITU-T recommendation P.50 Appendix I. These artificial voices have the same spectral and temporal characteristics as the human speech signals.

About this research paper

What this paper is about

Speech enhancement is the process of de-noising a speech signal for improved quality and better intelligibility. Several speech enhancement methods have been proposed including: DWFM filtering, Donoho, Massart, and Kalman. To measure the performance of these filters, a speech evaluation method is needed. SNR is one of the most common methods for speech evaluation. The problem of SNR as a waveform speech evaluation is it is too general and can fit any type of signal, even non-speech signals. In this paper the performance of several speech enhancement methods is compared using both SNR and PESQ which is an evaluation method that has been proposed by the ITU-T for speech-specific quality evaluation. The speech sources used during the experiments are artificial voices produced in ITU-T recommendation P.50 Appendix I. These artificial voices have the same spectral and temporal characteristics as the human speech signals.

Why it matters

OpenAlex reports 11 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Speech enhancement is the process of de-noising a speech signal for improved quality and better intelligibility. Several speech enhancement methods have been proposed including: DWFM filtering, Donoho, Massart, and Kalman. To measure the performance of these filters, a speech evaluation method is needed. SNR is one of the most common methods for speech evaluation. The problem of SNR as a waveform speech evaluation is it is too general and can fit any type of signal, even non-speech signals. In this paper the performance of several speech enhancement methods is compared using both SNR and PESQ which is an evaluation method that has been proposed by the ITU-T for speech-specific quality evaluation. The speech sources used during the experiments are artificial voices produced in ITU-T recommendation P.50 Appendix I. These artificial voices have the same spectral and temporal characteristics as the human speech signals.

Key concepts: PESQ, Speech enhancement, Intelligibility (philosophy), Speech recognition, Computer science, Speech processing, PSQM, Voice activity detection

Related papers

Back to paper searchBrowse research topicsOriginal source
Perceptual evaluation of speech enhancement — Research Paper | ScholarLens