Dissertation > Excellent graduate degree dissertation topics show
Research on Preprocessing of Robust Speech Recognition
Author: WangYue
Tutor: QuBaiDa
School: Jiangnan University
Course: Control Theory and Control Engineering
Keywords: endpoint detection speech enhancement wavelet transform threshold de-noising feature extraction
CLC: TN912.34
Type: Master's thesis
Year: 2008
Downloads: 118
Quote: 2
Read: Download Dissertation
Abstract
|
Robust Speech Recognition extracts the essential features of speech signal to recognize and confirm the noisy speech. The preprocessing of robust speech recognition is studied here, the purpose of which is to eliminate the noise interference and extract“clean”signal parameters. This paper mainly include following three parts: endpoint detection, speech enhancement and feature extraction.Firstly, endpoint detection is studied, the purpose of which is to pick the“meaningful”speech parts out and avoid the interference of noise from silence parts. Some classic endpoint detection methods are discussed here, such as: short-time energy, average zero-crossing rate, double-threshold detection, spectral entropy, power spectral entropy and spectrum variance. The related results all show the characteristics of their own. By analyzing the faults of spectrum variance, a modified endpoint detection method is proposed, namely sub-band spectrum variance method. Finally, the experimental results prove its superiority.Secondly, speech enhancement is studied, the purpose of which is to improve the SNR and intelligibility of speech. It is a key step to realize the robustness of speech recognition system. This part mainly studied some different de-noising methods like wavelet soft-threshold de-noising and wavelet hard-threshold de-noising. The Bionic Wavelet Transform which consider the auditory perceptual is deeply studied, then, the idea of threshold de-noising is applied to Bionic Wavelet Transform, so, a new speech enhancement method based on bionic wavelet transform is presented. The results indicate that the proposed method outperforms some classic methods including spectral subtraction, wiener filtering and threshold de-noising based on Discrete Wavelet Transform in four kinds of realistic noise environments, and has a better enhancement performance.Finally, feature extraction is studied, the purpose of which is to remove the redundant parts of speech and extract the essential features of speech for recognition. It mainly studied some common characteristic parameters of speech like LPC, LPCC and MFCC. MFCC and LPCC are compared here. It’s proved that MFCC is better than LPCC as characteristic parameters in representing the speech by constructing an isolated word recognition platform.
|
Related Dissertations
- Research on Automatic Detection Algorithm for Substructure Distress of Highway Pavement Based on SVM,U418.6
- ISAR Imaging Simulation of Space Targets and Target Recognition Based on ISAR Images,TN957.52
- Research on Feature Extraction and Classification of Pulse Waveform for Cholecystitis and Nephrotic Syndrome Diagnosis,TP391.41
- Application of Q-Learning in the Content-Based Image Retrieval Technology,TP391.41
- Research on Transductive Support Vector Machine and Its Application in Image Retrieval,TP391.41
- Research on Feature Extraction and Classification of Tongue Shape and Tooth-Marked Tongue in TCM Tongue Diagnosis,TP391.41
- Research on Visual Measurement for Spacecraft Rendezvous and Approach,TP391.41
- Feature Extraction, Selection and Combination in Lipreading,TP391.41
- Research on Visual Detection and Tracking of Mobile Robots,TP242.62
- Research on Identification System of Cashmere and Wool Fiber,TS101.921
- Research and System Implementation of Image Retrieval Method Based on Fuzzy Clustering,TP391.41
- The Research of Image Matching Method Based on Feature Descriptor,TP391.41
- The Speech Enhancement System Based on Binary Mask and Perceptual Wavelet Packet Transform,TN912.35
- The Discrimination of Rice Weeds Based on DSP,TP391.41
- SOPC Design of Transient Power Quality Disturbances Detecting Based on Nios Ⅱ,TN47
- Research and Implementation of Liver Cancer Identification Based on Improved SVM Model,TP391.41
- Based on Lifting Scheme Wavelet Packet Transform and Artificial Neural Network Wing-box Multi-damage Detection,V224
- Research of Pressure Fingerprint Identification System Key Technology,TP391.41
- Research on the Building Foundation Settlement Prediction by Wavelet Neural Network,TP183
- Recognition and Implementation of Multi-sintering Condition Based on Complete Binary Tree Supporting Vector Machine,TP391.41
- The Study of Coal Calorific Capacity Based on Texture Feature,TP391.41
CLC: > Industrial Technology > Radio electronics, telecommunications technology > Communicate > Electro-acoustic technology and speech signal processing > Speech Signal Processing > Speech Recognition and equipment
© 2012 www.DissertationTopic.Net Mobile
|