|
Recent years, with the development of the network , the information rapidly increased in the vast ocean of information quickly and effectively to obtain the required information , is troubled by the problem of online users . Users use a search engine to browse Web pages , can be part of the solution to resource discovery function , but its accuracy is not high , you can not provide users with structured information , can not provide the document classification , filtering , information resources a main forms - text , there is an urgent need to be able to from a large number of Web text collection quickly and effectively find the resources and knowledge of tools . This article by clustering data mining technology to do in-depth research , a web -based intelligent clustering system , a clustering algorithm as the core , automatically aggregate similar content pages , and eventually submitted to the user interface display . Clustering algorithm using the vector space model page documents using fuzzy clustering algorithm to dig out a high similarity set of documents , initially divided the document category , while the evaluation of the \will have a rough similarity in results document set is divided into a number of clusters , expanding a cluster document content similarity , and the similarity between different clusters shrinking , eventually achieve reasonably feather flock together \Hierarchical clustering by using the basic mining tools , the basic realization of the online , interactive , semantic level search engine search results clustering , which basically solved the users to retrieve information complicated problems.
|