Two-level Clustering of Web Sites Using Self-Organizing Maps

被引:0
作者
Dimitris Petrilis
Constantin Halatsis
机构
[1] University of Athens,Department of Informatics and Telecommunications
来源
Neural Processing Letters | 2008年 / 27卷
关键词
Access-logs; Clustering; Content mining; Context mining; Data mining; Neural networks; Self-organizing map (SOM); Text mining; Web-logs; Web mining;
D O I
暂无
中图分类号
学科分类号
摘要
Web sites contain an ever increasing amount of information within their pages. As the amount of information increases so does the complexity of the structure of the web site. Consequently it has become difficult for visitors to find the information relevant to their needs. To overcome this problem various clustering methods have been proposed to cluster data in an effort to help visitors find the relevant information. These clustering methods have typically focused either on the content or the context of the web pages. In this paper we are proposing a method based on Kohonen’s self-organizing map (SOM) that utilizes both content and context mining clustering techniques to help visitors identify relevant information quicker. The input of the content mining is the set of web pages of the web site whereas the source of the context mining is the access-logs of the web site. SOM can be used to identify clusters of web sessions with similar context and also clusters of web pages with similar content. It can also provide means of visualizing the outcome of this processing. In this paper we show how this two-level clustering can help visitors identify the relevant information faster. This procedure has been tested to the access-logs and web pages of the Department of Informatics and Telecommunications of the University of Athens.
引用
收藏
页码:85 / 95
页数:10
相关论文
共 11 条
[1]  
Andrade MA(1993)Evaluation of secondary structure of proteins from UV circular dichroism spectra Protein Eng 6 383-390
[2]  
Chacón P(1997)Computationally efficient approximation of a probabilistic model for document representation in the websom full-text analysis method Neural Process Lett 5 69-81
[3]  
Merelo-Guervós J(2004)Mining massive document collections by the WEBSOM method Information Sci 163 135-156
[4]  
Kaski S(2002)Application of artificial aging techniques to samples of rum and comparison with traditionally aged rums by analysis with artificial neural nets J Agric Food chem 50 1470-1477
[5]  
Lagus K(1969)A nonlinear mapping for data structure analysis IEEE Trans Comput 18 401-409
[6]  
Kaski S(undefined)undefined undefined undefined undefined-undefined
[7]  
Kohonen T(undefined)undefined undefined undefined undefined-undefined
[8]  
Quesada J(undefined)undefined undefined undefined undefined-undefined
[9]  
Merelo-Guervós JJ(undefined)undefined undefined undefined undefined-undefined
[10]  
Oliveras MJ(undefined)undefined undefined undefined undefined-undefined