2002•e-Publications@Marquette (Marquette University)Requires access

Speech Signal Enhancement Using a Microphone Array

Heather Elaine Ewalt

Open publisher page 1 citations

Abstract

This thesis describes the design and implementation of a speech enhancement system that uses microphone array beamforming and speech enhancement algorithms applied to a speech signal in a multiple source environment. The goal of the system is to improve the quality of the primary speech signal. Beamfonners work by means of steering an array of microphones towards a desired look direction through utilizing signal information rather than physically moving the array. They accomplish this through minimizing the energy of interference sources and noise in non-look directions while increasing the energy of the signal in the look direction. In this research, two beamfoming methods are examined: the delay and sum (DS) beamformer and the minimum variance distortionless response (MVDR) beamformer. The input signals are first split into frequency bands so that narrowband beamfonning techniques can be used. Multiple source Wiener filtering and multiple source spectral subtraction enhancement algorithms are incorporated into the two methods of beamforming. The algorithms utilize signal estimates of each source obtained from the initial beamfonning algorithms as inputs. These multiple source enhancement algorithms result in iterative techniques to improve those estimates while improving the signal to noise ratio of the primary source. The experimental setup presented here consists of both two and three speech sources using a linear microphone input system. The algorithms are performed on both simulated experimental setups and on data obtained from a data acquisition system in an acoustically treated sound room. To measure the improvement in quality of the enhanced signal, overall SNR and segmental SNR improvement is determined for the original, beamformed, and enhanced signal. In addition to these quality improvement metrics, listener opinion testing is performed.

About this research paper

What this paper is about

This thesis describes the design and implementation of a speech enhancement system that uses microphone array beamforming and speech enhancement algorithms applied to a speech signal in a multiple source environment. The goal of the system is to improve the quality of the primary speech signal. Beamfonners work by means of steering an array of microphones towards a desired look direction through utilizing signal information rather than physically moving the array. They accomplish this through minimizing the energy of interference sources and noise in non-look directions while increasing the energy of the signal in the look direction. In this research, two beamfoming methods are examined: the delay and sum (DS) beamformer and the minimum variance distortionless response (MVDR) beamformer. The input signals are first split into frequency bands so that narrowband beamfonning techniques can be used. Multiple source Wiener filtering and multiple source spectral subtraction enhancement algorithms are incorporated into the two methods of beamforming. The algorithms utilize signal estimates of each source obtained from the initial beamfonning algorithms as inputs. These multiple source enhancement algorithms result in iterative techniques to improve those estimates while improving the signal to noise ratio of the primary source. The experimental setup presented here consists of both two and three speech sources using a linear microphone input system. The algorithms are performed on both simulated experimental setups and on data obtained from a data acquisition system in an acoustically treated sound room. To measure the improvement in quality of the enhanced signal, overall SNR and segmental SNR improvement is determined for the original, beamformed, and enhanced signal. In addition to these quality improvement metrics, listener opinion testing is performed.

Why it matters

OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

This thesis describes the design and implementation of a speech enhancement system that uses microphone array beamforming and speech enhancement algorithms applied to a speech signal in a multiple source environment. The goal of the system is to improve the quality of the primary speech signal. Beamfonners work by means of steering an array of microphones towards a desired look direction through utilizing signal information rather than physically moving the array. They accomplish this through minimizing the energy of interference sources and noise in non-look directions while increasing the energy of the signal in the look direction. In this research, two beamfoming methods are examined: the delay and sum (DS) beamformer and the minimum variance distortionless response (MVDR) beamformer. The input signals are first split into frequency bands so that narrowband beamfonning techniques can be used. Multiple source Wiener filtering and multiple source spectral subtraction enhancement algorithms are incorporated into the two methods of beamforming. The algorithms utilize signal estimates of each source obtained from the initial beamfonning algorithms as inputs. These multiple source enhancement algorithms result in iterative techniques to improve those estimates while improving the signal to noise ratio of the primary source. The experimental setup presented here consists of both two and three speech sources using a linear microphone input system. The algorithms are performed on both simulated experimental setups and on data obtained from a data acquisition system in an acoustically treated sound room. To measure the improvement in quality of the enhanced signal, overall SNR and segmental SNR improvement is determined for the original, beamformed, and enhanced signal. In addition to these quality improvement metrics, listener opinion testing is performed.

Key concepts: Speech enhancement, Speech recognition, Microphone, SIGNAL (programming language), Computer science, Acoustics, Microphone array, Background noise

Related papers

Back to paper searchBrowse research topicsOriginal source
Speech Signal Enhancement Using a Microphone Array — Research Paper | ScholarLens