Dissertation > Excellent graduate degree dissertation topics show
Study on Information Extraction for Multimedia Program Catalogue Based on Internet
Author: LiMingYue
Tutor: ZhouJun
School: Shanghai Jiaotong University
Course: Signal and Information Processing
Keywords: Multimedia programs cataloging information HTML automatic classification Information extraction Feature extraction
CLC: TP391.41
Type: Master's thesis
Year: 2008
Downloads: 34
Quote: 0
Read: Download Dissertation
Abstract
|
In recent years, with the continuous expansion of business of digital multimedia, digital multimedia business audience cataloging information requirements of multimedia programs will follow. There is currently no research in this area to meet the urgent needs of the audience, which is the background of this study selection and power. The rapid development of Internet of Web data is growing, leading to the generation of a large number of semi-structured (semi-structured) data, a large number of multimedia description on the Internet allows us to derive the multimedia programs cataloging information may. In order to obtain multimedia programs cataloging information, the paper looks at the Internet resources, the cataloging information network of the multimedia programs automatically extracting technology as the objectives and tasks. Firstly, the general implementation of Web information extraction technology general classification and web information extraction system. On this basis, combined with the characteristics of multimedia programs cataloging information, the automatic extraction of a multimedia program cataloging information network system NMPIES, the system design is relatively simple, clear structure, in theory, can accomplish the goal of the proposed policy. WEB pretreatment and Web classification the catalog information extracted thesis focus. Traditional WEB pretreatment technology generally relates to noise filtering of HTML, text extraction technology, the use of these simple techniques is difficult to prepare for of multimedia programs cataloging information extraction. Therefore, the paper studied the characteristics of HTML pages, a set of multimedia programs cataloging information extraction WEB pretreatment technology, including HTML-Tree center content key to determine, Web Feature extraction based on HTML-Tree method technology, through the implementation of these key technologies to achieve the purpose of preprocessing of Web information to improve the Web page automatic classification precision and recall. Then the thesis, the main implementation technology for the multimedia programs cataloging information extraction, topic-based information extraction method, the method by multimedia programs cataloging information template, theme similarity judgment and pattern matching to finally get a more complete multimedia programs cataloging information, the method can be used to complete the target for some simple common cataloging information. Finally, on the Java platform paper multimedia program catalog information automatically extracted system NMPIES, and a large number of experiments, achieved good results.
|
Related Dissertations
- Research on Automatic Detection Algorithm for Substructure Distress of Highway Pavement Based on SVM,U418.6
- ISAR Imaging Simulation of Space Targets and Target Recognition Based on ISAR Images,TN957.52
- Research on Feature Extraction and Classification of Pulse Waveform for Cholecystitis and Nephrotic Syndrome Diagnosis,TP391.41
- Application of Q-Learning in the Content-Based Image Retrieval Technology,TP391.41
- Research on Domain Entity Attribute and Event Extraction Technology,TP391.1
- Research on Transductive Support Vector Machine and Its Application in Image Retrieval,TP391.41
- Research on Feature Extraction and Classification of Tongue Shape and Tooth-Marked Tongue in TCM Tongue Diagnosis,TP391.41
- Research on the Image Real-Time Acquisition, Storage and Image Processing System,TP391.41
- Feature Extraction, Selection and Combination in Lipreading,TP391.41
- Multi-currency Notes Technology Research and Implementation,TP391.41
- The Research on Paper Currency Classification Method Based on Harr-Like Feature and Minimal Ball Including Samples,TP391.41
- Pavement Distress Recognition Based on Image,TP391.41
- Research on Visual Detection and Tracking of Mobile Robots,TP242.62
- Research on Fusion Algorithm of Hyper Spectral and High Spatial Resolution Remote Sensing Image,TP751
- Tobacco Diseases Auto-Recognition Research Based on Image Processing Technology,S435.72
- Research on Nondestructive Detection Technology for External Qualities of Papayas Based-on Vision,S667.9
- Research on Identification System of Cashmere and Wool Fiber,TS101.921
- Study on Growth Monitoring Technique Based on Pixel Un-Mixing Method and HJ Remote Sensing Images in Paddy Rice,S511
- Research for Infrared Image Target Identification and Tracking Technology,TP391.41
- The Compression and Fusion Technique Research of Underwater Target Feature,TN911.7
- Research of Diagnosing Cucumber Diseases Based on Hyperspectral Imaging,S436.421
CLC: > Industrial Technology > Automation technology,computer technology > Computing technology,computer technology > Computer applications > Information processing (information processing) > Pattern Recognition and devices > Image recognition device
© 2012 www.DissertationTopic.Net Mobile
|