Dissertation > Excellent graduate degree dissertation topics show
Research and Implementation of Characteristics of Complex Sentence Analyzer in Chinese Information Processing
Author: XuZuo
Tutor: HuJinZhu
School: Central China Normal University
Course: Computer Software and Theory
Keywords: Chinese information processing Complex sentence processing Sentence similarity algorithm Clause marker Dependency Parser Sentence element analysis
CLC: TP391.1
Type: Master's thesis
Year: 2011
Downloads: 39
Quote: 2
Read: Download Dissertation
Abstract
|
Chinese information processing as a cross disciplines combined with the discipline of computer science , linguistics , mathematics , informatics and acoustic , with the popularity of the Internet and the development of information processing technology in recent years, the rapid development . Chinese information processing , Chinese information processing , including word processing , word processing , sentence processing and chapter handling . However , due to the special nature and complexity of the Chinese language , so far , most of the studies only new words and word processing \This paper studies the complex sentence signature analyzer is a core part of the relationship between word auto-tagging system engineering of complex sentences , mainly responsible for the decimation of the basic characteristics of the Chinese complex sentences . Complex sentence signature analyzer seven functional modules : a sentence structure similar calculation , the two syntactic component analysis , 3 - string matching , 4 part-of-speech tagging , clause marker span calculated 6 semantic association calculation , 7 the relationship between word processing . Research and explore key technologies of of complex sentences characteristics analyzer : 1 , put forward a new kind of Chinese sentences similar to the algorithm . This structure is similar to a string based on the Chinese sentences parts of speech algorithm , which combines the degree of association between the parts of speech to find the two sentences to the corresponding parts of speech of the longest string matching string . 2, proposed a clause labeled algorithm . The basic idea of the algorithm : a practical and efficient merger principle associated with the word alone into sentences , sentence elements separate into sentences independent clauses normalized to the adjacent clause , the clause so as to realize the reasonable mark . 3 , the proposed sentence component analysis algorithm based on dependency syntax . The algorithm uses the rules of syntax component analysis : that nuclear identification mechanism , the trunk identification mechanism , the modification component identification mechanism and the conjuncts recognition mechanism , semantic clause division of the Chinese complex sentences , each semantic clause SVO division , as well as core words, modifiers , and tied for the division of components .
|
Related Dissertations
- Research of Copy Detection for Chinese Text,TP391.1
- Research on Keeping Consistency of Chinese Corpus of Complete Parsing,TP391.1
- The Study of Modern Chinese New Word Extraction,H08
- Study on Compound Predicate-complement Structure for Chinese Information Processing,H146
- Study on Syntactic and Semantic Relation in "V+N" Structure for Chinese Information Processing,H146
- Research on Detection of Approximate Mirror Web Pages,TP393.092
- Research on Statistical Model and Treebank Conversion for Dependency Parsing,TP391.1
- Dependency Parsing Based Semantic Role Labeling,TP391.1
- Chinese information processing research on key issues,TP391.1
- Chinese Words Segmentation Based on Context and Stopwords,TP391.1
- Research on Named Entity Recognition Based on Rules,TP391.1
- Case Studies of the "Answering Online" System Based on Chinese Words Segmentation Techniques,TP311.52
- Research on Chinesese Segmentation Method Based on Optimization Maximum Matching,TP391.1
- Research on Semantic Orientation Classification of Chinese Textsbased Evaluation Objects and Affective Characteristics,TP391.1
- Chinese semantic information data mining technology research,TP391.1
- Hot Topic Mining and Opinion Analysis on BBS,TP393.094
- Design and Implementation of Automatical Recognition of Chinese Personal Name,TP391.1
- Chinese Information Processing on the Link Writing for Chinese Word,TP391.1
- The Realization of Statistic Software of Typed Error Rate in Chinese Text,TP391.1
- General Purpose Designing Of Modern Chinese Text Automate Segmentation and Dealing the Ambiguity in Segmentation,TP391.1
CLC: > Industrial Technology > Automation technology,computer technology > Computing technology,computer technology > Computer applications > Information processing (information processing) > Text Processing
© 2012 www.DissertationTopic.Net Mobile
|