Dissertation > Excellent graduate degree dissertation topics show

Research on Emotional Recognition in Multilingual Speech Signal

Author: LiBo
Tutor: WangYuTai
School: Jinan University
Course: Signal and Information Processing
Keywords: speech emotional recognition multimode recognition Principal Component Analysis Gaussian Mixture Model Dynamic Time Warping
CLC: TN912.34
Type: Master's thesis
Year: 2010
Downloads: 145
Quote: 2
Read: Download Dissertation

Abstract


Speech is an important human-specific means of emotional expression, which includes specific emotional psychological characteristics as well as semantic information. The traditional speech processing system usually focuses on the accuracy of the content of speech, and neglects the research of psychological characteristics. In recent years, with the increasing application requirement of natural human-machine interaction, psychological testing, intelligent robots and many other fields, the emotional analysis and recognition in speech signal get more and more attention, and become the new research hotspot. However, the research of emotional recognition still needs further study. The building of emotional speech database, the selection and extraction of emotional characteristic parameters, and the emotional recognition have not formed systematic theory. The research of emotional recognition is usually based on English, but Chinese get less research. Further, the emotional parameters are mainly focus on prosodic characteristics, while the research of multimode recognition, which intergrates semantic information, facial expression and physiology signal, is also paid little attention. Therefore, it can be said that speech emotional recognition is still in the preliminary stage, and more deep research is needed.In order to research the emotional recognition based on multilingual speech signals, this thesis focuses on the building of emotional speech database on the basis of multi-languages, which includes Chinese, English, Japanese, Korean and Russian, analysis of prosodic characteristic parameters, extraction of emotional characteristic parameters, speech emotional recognition and emotional recognition combined with semantic information. The main contents of this thesis are as follows:First, the emotion is devided into five categories, i.e. quiet, happiness, anger, surprise and sadness. Then, we record the emotional speech in laboratory conditions, and the multilingual emotional speech database is built up for further research.Second, the speech signals of five emotions spoken in different languages are acoustically analyzed, and prosodic characteristic parameters are extracted. After analysis of emotional speech signals and comparison of acoustic features between different emotions, the general rule of speech emotional features is concluded, i.e. the changes of different languages’parameters in the same emotion exists commonness.Third, emotional recognition experiment is carried out using two algorithm, namely Principal Component Analysis and Gaussian Mixture Model, based on multilingual emotional speech database. The two algorithms achieve 74.2% and 78.1% average recognition accuracy respectively.Fourth, on the basis of acoustic features, emotional recognition intergrated semantic information experimentized. Words with different emotional color are annotated; the semantic information of a sentence is extracted using Dynamic Time Warping which is used to recognize emotional key words. Then, the prosodic characteristics are combined with semantic information for recognize the emotion using Gaussian Mixture Model. The experimental results demonstrate that the recognition accuracy of combined characteristics makes 3 percent improvement compared with the prosodic characteristics.The main innovations of this thesis are: first, built emotional speech database based on multi-languages, extracted prosodic features and concluded the general rule of speech emotional features; second, carried out emotional recognition experiment intergrated semantic information on the basis of acoustic features, and achieved better recognition accuracy than the prosodic characteristics.

Related Dissertations

  1. Application of Improved Principal Component Analysis Algorithm in Course Construction,G642.4
  2. Research of Diagnosing Cucumber Diseases Based on Hyperspectral Imaging,S436.421
  3. The Impact of Tourism on Typical Vegetation in Luya Mountain Nature Reserve, Shanxi Province,S759.9
  4. Research on Cultural Industrial Competitiveness of Chong Qing,F224
  5. Detection and Tracking of Moving Object in Complex Background,TP391.41
  6. The Study of Competitiveness of Coastal Ports in China Based on Principal Component Analysis,F552
  7. The Research of Prairie Road Light Environment Effects on Physiological Indicators of Drivers,U491.254
  8. The Research and Empirical Analysis on the Relation between Financial Ecology and Economic Growth in Jiangsu Province,F127;F224
  9. Key Algorithm in High Quality Voice Conversion System,TN912.3
  10. Research of Moving Object Detection and Tracking Technology,TP391.41
  11. Image Spam Detecting Based on Combinatorial and Statistical Classifier,TP391.41
  12. The Research of E-government System Performance Evaluation Index System in Linyi City Based on the Principal Component Analysis,G206
  13. Multi-feature fusion of visual tracking algorithm,TP391.41
  14. Video surveillance system moving target detection algorithm,TP391.41
  15. Waste Road density points seismic signal automatic identification and picking first method,P631.4
  16. The Research on Microscopic Evaluation System of Road Traffic Safety,U491
  17. Signature Verification Based on Video,TP391.41
  18. Improved Principal Component Analysis in Chinese Universities Ranked in the Application of Mathematics,O212.4
  19. Effect of Roasting on Camellia Oil Aqueous Extraction Technique and Its Quality and Aroma,TS224
  20. The Prediction Research of Furnace Status As to Tendency to Tcold and Hot Based on Support Vector Machine,TF57
  21. Mobile robot voice recognition control simulation system design and implementation,TN912.34

CLC: > Industrial Technology > Radio electronics, telecommunications technology > Communicate > Electro-acoustic technology and speech signal processing > Speech Signal Processing > Speech Recognition and equipment
© 2012 www.DissertationTopic.Net  Mobile