Multi-stage speech enhancement for automatic speech recognition
Seungyeol Lee, Youngwoo Lee, Namgook Cho
Abstract
Seungyeol Lee, Youngwoo Lee, Namgook Cho
Abstract
In this paper, we propose a multi-stage speech enhancement technique for speech recognition. At first, a multi-channel speech enhancement method takes advantage of the spatial information of speech source. Then, in the second stage, single-channel speech enhancement based on data-driven approach is adopted to improve performance of speech recognition at server side. This method can improve the quality of speech signal which maximizes the advantage of each speech enhancement technique. The experimental result shows that the proposed technique is superior to conventional multi-stage speech enhancement algorithms.
OpenAlex reports 4 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
In this paper, we propose a multi-stage speech enhancement technique for speech recognition. At first, a multi-channel speech enhancement method takes advantage of the spatial information of speech source. Then, in the second stage, single-channel speech enhancement based on data-driven approach is adopted to improve performance of speech recognition at server side. This method can improve the quality of speech signal which maximizes the advantage of each speech enhancement technique. The experimental result shows that the proposed technique is superior to conventional multi-stage speech enhancement algorithms.
Key concepts: Speech enhancement, Computer science, Speech recognition, Voice activity detection, Speech processing, Speech coding, Acoustic model, PSQM