Dissertation > Excellent graduate degree dissertation topics show

Research and Implementation of Characteristics of Complex Sentence Analyzer in Chinese Information Processing

Author: XuZuo
Tutor: HuJinZhu
School: Central China Normal University
Course: Computer Software and Theory
Keywords: Chinese information processing Complex sentence processing Sentence similarity algorithm Clause marker Dependency Parser Sentence element analysis
CLC: TP391.1
Type: Master's thesis
Year: 2011
Downloads: 39
Quote: 2
Read: Download Dissertation

Abstract


Chinese information processing as a cross disciplines combined with the discipline of computer science , linguistics , mathematics , informatics and acoustic , with the popularity of the Internet and the development of information processing technology in recent years, the rapid development . Chinese information processing , Chinese information processing , including word processing , word processing , sentence processing and chapter handling . However , due to the special nature and complexity of the Chinese language , so far , most of the studies only new words and word processing \This paper studies the complex sentence signature analyzer is a core part of the relationship between word auto-tagging system engineering of complex sentences , mainly responsible for the decimation of the basic characteristics of the Chinese complex sentences . Complex sentence signature analyzer seven functional modules : a sentence structure similar calculation , the two syntactic component analysis , 3 - string matching , 4 part-of-speech tagging , clause marker span calculated 6 semantic association calculation , 7 the relationship between word processing . Research and explore key technologies of of complex sentences characteristics analyzer : 1 , put forward a new kind of Chinese sentences similar to the algorithm . This structure is similar to a string based on the Chinese sentences parts of speech algorithm , which combines the degree of association between the parts of speech to find the two sentences to the corresponding parts of speech of the longest string matching string . 2, proposed a clause labeled algorithm . The basic idea of the algorithm : a practical and efficient merger principle associated with the word alone into sentences , sentence elements separate into sentences independent clauses normalized to the adjacent clause , the clause so as to realize the reasonable mark . 3 , the proposed sentence component analysis algorithm based on dependency syntax . The algorithm uses the rules of syntax component analysis : that nuclear identification mechanism , the trunk identification mechanism , the modification component identification mechanism and the conjuncts recognition mechanism , semantic clause division of the Chinese complex sentences , each semantic clause SVO division , as well as core words, modifiers , and tied for the division of components .

Related Dissertations

  1. Research of Copy Detection for Chinese Text,TP391.1
  2. Research on Keeping Consistency of Chinese Corpus of Complete Parsing,TP391.1
  3. The Study of Modern Chinese New Word Extraction,H08
  4. Study on Compound Predicate-complement Structure for Chinese Information Processing,H146
  5. Study on Syntactic and Semantic Relation in "V+N" Structure for Chinese Information Processing,H146
  6. Research on Detection of Approximate Mirror Web Pages,TP393.092
  7. Research on Statistical Model and Treebank Conversion for Dependency Parsing,TP391.1
  8. Dependency Parsing Based Semantic Role Labeling,TP391.1
  9. Chinese information processing research on key issues,TP391.1
  10. Chinese Words Segmentation Based on Context and Stopwords,TP391.1
  11. Research on Named Entity Recognition Based on Rules,TP391.1
  12. Case Studies of the "Answering Online" System Based on Chinese Words Segmentation Techniques,TP311.52
  13. Research on Chinesese Segmentation Method Based on Optimization Maximum Matching,TP391.1
  14. Research on Semantic Orientation Classification of Chinese Textsbased Evaluation Objects and Affective Characteristics,TP391.1
  15. Chinese semantic information data mining technology research,TP391.1
  16. Hot Topic Mining and Opinion Analysis on BBS,TP393.094
  17. Design and Implementation of Automatical Recognition of Chinese Personal Name,TP391.1
  18. Chinese Information Processing on the Link Writing for Chinese Word,TP391.1
  19. The Realization of Statistic Software of Typed Error Rate in Chinese Text,TP391.1
  20. General Purpose Designing Of Modern Chinese Text Automate Segmentation and Dealing the Ambiguity in Segmentation,TP391.1

CLC: > Industrial Technology > Automation technology,computer technology > Computing technology,computer technology > Computer applications > Information processing (information processing) > Text Processing
© 2012 www.DissertationTopic.Net  Mobile