Speech visualization based on Auditory Model for hearing impaired
Xu Wang, Lifang Xue, Dan di Yang, Zhiyan Han
Abstract
Xu Wang, Lifang Xue, Dan di Yang, Zhiyan Han
Abstract
This paper describes a novel speech visualization method that creates a readable pattern based on auditory model. The auditory model extracts the critical information of the speech signal and presents more robust than a system based on conventional acoustic processing techniques (MFCC). Firstly, speech signal undergoes a series of preprocessing course. Secondly, we make use of auditory model for speech feature extracting. Finally, we utilize plot display algorithm to generate a speech plot. The auditory feature of speech signal is displayed on the CRT by plot patterns and the deaf can utilize their own brain to identify different speech for training their oral ability effectively. The aim of this paper is to reduce speech multidimensionality in order to obtain a simple and accurate image for speech visualization.
OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
This paper describes a novel speech visualization method that creates a readable pattern based on auditory model. The auditory model extracts the critical information of the speech signal and presents more robust than a system based on conventional acoustic processing techniques (MFCC). Firstly, speech signal undergoes a series of preprocessing course. Secondly, we make use of auditory model for speech feature extracting. Finally, we utilize plot display algorithm to generate a speech plot. The auditory feature of speech signal is displayed on the CRT by plot patterns and the deaf can utilize their own brain to identify different speech for training their oral ability effectively. The aim of this paper is to reduce speech multidimensionality in order to obtain a simple and accurate image for speech visualization.
Key concepts: Computer science, Visualization, Speech recognition, Preprocessor, Plot (graphics), Speech processing, Feature (linguistics), Mel-frequency cepstrum