Dissertation > Excellent graduate degree dissertation topics show
Research on Improved Zcpa Speech Recognition Feature Extraction Algorithm
Author: JiaoZhiPing
Tutor: ZhangXueYing
School: Taiyuan University of Technology
Course: Signal and Information Processing
Keywords: speech recognition feature extraction wavelet transform auditory model
CLC: TN912.3
Type: Master's thesis
Year: 2005
Downloads: 241
Quote: 13
Read: Download Dissertation
Abstract
|
Recently, most of speech recognition systems have much more recognition rates in the clean environment, however, the performances of these systems are severely degraded when there exists noisy environment. For the practicability of the speech recognition technology, it has the important significance to study the robustness of the speech recognition.The recognition ability of the human ear is very well, even in the noisy environment. So some researchers have devoted to the study of the auditory model to extract speech feature parameters that will improve the robustness of the system.This paper focuses on the robust noise speech recognition, completing the following research works.Firstly, this paper accomplished the speech recognition withthe Zero-crossings with Peak Amplitudes features, which is based on the auditory model of the human ear. This model reflects the speech signals’ frequency information by analyzing and computing the adjacent upward-going zero-crossing intervals, and allocates them to the corresponding frequency bins. Then it is detected that the peak amplitudes of the successive upward-going intervals. Finally, it goes along the compressive nonlinearity to weight the frequency bins’ peaks. This paper has also analyzed the robust noise performance. The results of many experiments showed that the robust noise performance of this system outperforms the other recognition system that uses the LPCC, MFCC as the recognition features.Secondly, this paper presented the improved ZCPA features based on the above system, namely, combining difference ZCPA features. These features use the characteristic of the difference signals, adding the difference information to the ZCPA features. The new features can extract the high frequency information mixed in the low frequency information, so the deficiency of the ZCPA features are made up for, and the improved recognition results are obtained.And this paper studied the front-end filters of this recognition system, introducing to use the Bark wavelet filters instead of the FIR filters. However, most of wavelet transforms, whether they aredyadic wavelets, wavelet packets or M-band wavelet transform, their frequency allocations all are based on octave relation, however, these frequency allocations are much more different with critical frequency band allocations of the human. So, if there is a wavelet that can allocate frequency according to the critical band, this wavelet will much more accord with the perception of the human to the speech and will improve the performance of the system. The basal thought of construction Bark wavelet is that the selected wavelet mother function satisfies the minimum of the time and bandwidth product, namely, the gauss function of the Bark fields, the mother wavelet has the equal bandwidth in the Bark fields. This paper analyzed the decomposition and reconstruction of this wavelet, presented the characteristic of time and frequency fields about this wavelet, and introduced the theory of this wavelet used in the front-end preprocessing.Finally, this paper simulated the speech recognition based on the Bark wavelet filters and ZCPA features, obtained the improved results and increased the recognition rates of the system.
|
Related Dissertations
- Research on Automatic Detection Algorithm for Substructure Distress of Highway Pavement Based on SVM,U418.6
- Multiple ANN/HMM Hybrid Used in Speech Recognition,TN912.34
- The Design of a DSP-Based Robot Speech Command Recgnition System,TN912.34
- ISAR Imaging Simulation of Space Targets and Target Recognition Based on ISAR Images,TN957.52
- Research on Feature Extraction and Classification of Pulse Waveform for Cholecystitis and Nephrotic Syndrome Diagnosis,TP391.41
- Application of Q-Learning in the Content-Based Image Retrieval Technology,TP391.41
- Research on Transductive Support Vector Machine and Its Application in Image Retrieval,TP391.41
- Research on Feature Extraction and Classification of Tongue Shape and Tooth-Marked Tongue in TCM Tongue Diagnosis,TP391.41
- Research on Visual Measurement for Spacecraft Rendezvous and Approach,TP391.41
- Research on the Image Real-Time Acquisition, Storage and Image Processing System,TP391.41
- Feature Extraction, Selection and Combination in Lipreading,TP391.41
- The Research on Paper Currency Classification Method Based on Harr-Like Feature and Minimal Ball Including Samples,TP391.41
- Research on Fusion Algorithm of Hyper Spectral and High Spatial Resolution Remote Sensing Image,TP751
- Tobacco Diseases Auto-Recognition Research Based on Image Processing Technology,S435.72
- Research on Identification System of Cashmere and Wool Fiber,TS101.921
- Research for Infrared Image Target Identification and Tracking Technology,TP391.41
- Research of Diagnosing Cucumber Diseases Based on Hyperspectral Imaging,S436.421
- Characteristics of sensory stimulation evoked,R318.0
- Network transmission ROI image coding algorithm,TN919.81
- The Research of Image Matching Method Based on Feature Descriptor,TP391.41
- Research of Fault Diagnosis Method of Analog Circuit Based on Improved Support Vector Machines,TN710
CLC: > Industrial Technology > Radio electronics, telecommunications technology > Communicate > Electro-acoustic technology and speech signal processing > Speech Signal Processing
© 2012 www.DissertationTopic.Net Mobile
|