Dissertation > Excellent graduate degree dissertation topics show

Notation of Speaking Face Based on Video and Text Infomation

Author: LiuGuangZheng
Tutor: DingYuXin
School: Harbin Institute of Technology
Course: Computer Science and Technology
Keywords: face detection and tracking mouth location lip-moving dynamic time warping speaking face detectio
CLC: TP391.41
Type: Master's thesis
Year: 2010
Downloads: 32
Quote: 0
Read: Download Dissertation

Abstract


The Speaking face detection based on video information, which means through lip-moving to judge who is speaking without audio information. The correlation technique is: the shot division of the video, the face detection and tracking, the lip location as well as the decision of lip-moving. Regarding name labeling, simultaneously needs the text information, and considered the characters of subtitles and the script, need to fuse the text information.About he fusion of the subtitles and script, the paper introduces a dynamic time warping algorithm, the algorithm uses the idea of fusion and is characterized by the word, and achieved good results. For the human face detection and tracking, taking into account the accuracy of detection, efficiency, and code reuse, this paper uses AdaBoost algorithm and MeanShift algorithm in OpenCV vision library, this combination of methods through the verification of the experiment, and achieved good results. And in this paper, we use this method in face sequence extraction.The detection of mouth has always been the content of lip-reading research field, this article will introduce it to the lips extraction in speaker detection process. Consider the methods proposed in the literature, we use lip color to extract the lip region, and do some improve in this paper, then extract more accuracy lip area. After tested and achieved good results.In previous studies, the use of the speaker’s lip detection method, less complex than the lip-reading field, mostly just compute the difference between two image on the mouth region, and set a threshold to determine whether the lip is moving. This paper introduces a machine learning method, by extracting a variety of features in lip region, and training classifier to determine whether the lip is moving. The experiments prove the accuracy and robustness of this method.The method used in the literatures to detect speaking face is based on single frame, that is, people judge whether a face in a frame is speaking. But for the situation, in a face sequence just some of image in the sequence is lip-moving and this sequence isn’t speaking, this method can not distinguish the difference. Therefore we propose this method which is based on the selected image sequence to determine if the sequence is speaking in a period of time. The proposed method is more realistic, and also achieved good results.

Related Dissertations

  1. Signature Verification Based on Video,TP391.41
  2. Study on Human Face Detection and Tracking Algorithm Based on Skin Color,TP391.41
  3. Mobile robot voice recognition control simulation system design and implementation,TN912.34
  4. Fast Time Series Similarity Matching and Its Application Research in Molten Iron Silicon Content Modeling,TF513
  5. Study on Object Tracking and Recognition Algorithm in Human-Robot Interaction,TP391.41
  6. DSP Network Video Monitoring System Research Based on Human Face Detection and Position,TP391.41
  7. Research on Stroke Distance Based Handwriting Document Retrieval Algorithm,TP391.43
  8. The Application of Similarity Query Based on DTW in Well-completion Depth Calculation,TE257
  9. DaVinci - based intelligent monitoring system,TP277
  10. Design & Implementation of Medolic-Based Music Retrieval System,TP391.3
  11. Research of Emotion Recognition Based on Combined Speech Feature,TN912.3
  12. Research and Implementation of Speech Recognition System for Mobile Robot,TN912.34
  13. The Chinese vowel length adjustment based speech recognition,TN912.34
  14. Gait acceleration signal based authentication method,TN911.7
  15. The Study of Gait Recognition Based on Image Sequence and Pressure,TP391.41
  16. Aksu River Runoff Time Series Analysis,TV121
  17. Personal Identification Based on Two Dimensional Gait,TP391.41
  18. ASIC Design of Specific Speaker Isolated Speech Recognition System,TN912.34
  19. Study on Engine Monitoring and Fault Diagnosis Methods Based on Exhaust Pressure Wave Analysis,U472.43
  20. Music Feature Analysis and Its Application in Content-based Retrieval,TP391.3

CLC: > Industrial Technology > Automation technology,computer technology > Computing technology,computer technology > Computer applications > Information processing (information processing) > Pattern Recognition and devices > Image recognition device
© 2012 www.DissertationTopic.Net  Mobile