Abstract
Clustering the web documents is one of the most important approaches for mining and extracting knowledge from the web. Recently, one of the most attractive trends in clustering the high dimensional web pages has been tilt toward the learning and optimization approaches. In this paper, we propose novel hybrid harmony search (HS) based algorithms for clustering the web documents that finds a globally optimal partition of them into a specified number of clusters. By modeling clustering as an optimization problem, first, we propose a pure harmony search-based clustering algorithm that finds near global optimal clusters within a reasonable time. Then, we hybridize K-means and harmony clustering in two ways to achieve better clustering. Experimental results reveal that the proposed algorithms can find better clusters when compared to similar methods and also illustrate the robustness of the hybrid clustering algorithms.
| Original language | English (US) |
|---|---|
| Pages (from-to) | 441-451 |
| Number of pages | 11 |
| Journal | Applied Mathematics and Computation |
| Volume | 201 |
| Issue number | 1-2 |
| DOIs | |
| State | Published - Jul 15 2008 |
All Science Journal Classification (ASJC) codes
- Computational Mathematics
- Applied Mathematics
Fingerprint
Dive into the research topics of 'Novel meta-heuristic algorithms for clustering web documents'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver