Dissertation > Excellent graduate degree dissertation topics show
Research of Web Database Approximate Query Based on Semantic Similarity Computing
Author: PanQiaoZhi
Tutor: MaZongMin
School: Northeastern University
Course: Applied Computer Technology
Keywords: Web databases Approximate query Ranking Preference Similarity
CLC: TP311.13
Type: Master's thesis
Year: 2009
Downloads: 30
Quote: 0
Read: Download Dissertation
Abstract
|
With the rapid expansion of the Internet, Web database has obtained the widespread application. More and more common users access the Web database by using the query interface to gain more information. Database query processing models have always assumed that the user knows what he wants and is able to formulate query that accurately express his query needs. However, most of the Web database users are common users. They often have insufficient knowledge about database contents and structure.As a result, their query intentions are usually vague or imprecise as well.And Therefore, the query results based on the user’s query intentions often lead to empty or umsatisfactory answers. In order to reduce the user’s tentative query times, or retrieve more satisfactory query results, the query submitted should act as approximate constraints for the query results. Therefore, the query should be an approximate one,in which the user’s initial query can be extended through the use of related technologies. Also the approximate query can result in too many answers over large Web database, while the users are only interested in the results which meet their intention most closely. So it is important to rank the query results of the approximate queries.This thesis presents a novel approach of the approximate query and ranking (AQR) which depends on the original data from database and the query workload statistics to find the most similar values based on semantic similarity. The user’s initial query criteria can be extended and the user’s expectation information are then obtained as far as possible. AQR extends the categorical query criteria with the most similar values by evaluating the similarity of different pairs of values in the query workloads and expands the numerical query range to nearby values by using the kernel density estimation technology. AQR speculates the importance of each specified attribute based on the user’s query and assigns the score of each attribute according to its "desirability" of the user, and then the relevant tuples are ranked according to their satisfaction degree to the user’s needs and preferences.The thesis presents a approach to evaluate the user’s preference to the unspecified attributes and solves the tuple ranking question of the same similarity. This approach further improves the ranking quality of the results of approximate query.The quality of the proposed approach is evaluated with experiments on two real databases. Here no domain knowledge or user’s feedback is required in the whole process. The experimental results demonstrate that the proposed approach can capture the user’s preferences effectively and have a high ranking quality as well.
|
Related Dissertations
- Study on Host Selection of Encarsia Sophia between Bemisia Tabaci B-Biotype and Q-Biotype Nymph,S476.3
- The Impact of Tourism on Typical Vegetation in Luya Mountain Nature Reserve, Shanxi Province,S759.9
- Ontology -based Semantic Web service matching and composition method,TP393.09
- WordNet and the \,G254
- Research of Text Clustering on Food Complaint Documents Based on Ontology,TP391.1
- The Study of 3d Face Recognition System,TP391.41
- Multi-Objective Particle Swarm Optimization Based on Fuzzy Preferences and Its Application in Inventory Control,F253.4
- Agent-Based Relevant Reseatch of Auto Negotiation,TP18
- Research on Pricing Strategy of Manufacturing/Remanufacturing System in Different Market Structure,F224
- Research on Image Super-resolution Reconstruction Based on Non-local Similarity,TP391.41
- Study on the Brand Preference for Private Cars Based on City Consumers,F426.471;F224
- Research and Application of Similarity-based Mining of Financial Data Analysis System,TP311.13
- Design and Develop of the Distance Education Automatic Question-answering System Based on Distributed Technology,TP391.6
- Research of Semantic Retrieval on Instructional Resources Based on Ontology,TP391.3
- Ontology-based Knowledge Retrieval System Smart Grid,TM76
- Head and shoulders image automatic segmentation of video,TP391.41
- Based on the structural characteristics of human visual similarity image quality evaluation,TP391.41
- Based on structural similarity with the MTF image quality evaluation method,TP391.41
- Research on Protein Complex Identification and Visualization in Protein Interaction Networks,TP391.41
- Image Matching Based on Local Invariant Features,TP391.41
- Study and Application of Integrated Evaluation Methods of Distribution Systems,TM715
CLC: > Industrial Technology > Automation technology,computer technology > Computing technology,computer technology > Computer software > Program design,software engineering > Programming > Database theory and systems
© 2012 www.DissertationTopic.Net Mobile
|