Clustering on Web usage data using Approximations and Set Similarities
Ms K. Santhisree, Dr A. Damodaram · International Journal of Computer Applications · 2010
Web usage mining is the application of data mining techniques to web log data repositories. It is used in finding the user access patterns from web access log. User page visits are sequential in nature. In this paper we presented clustering web transactions based on the set similarity measures from web log data which identifies the behavior of the users page visits, order of occurrence of visits. Web data Clusters are formed using the Similarity Upper Approximations. We present the experimental results on MSNBC web navigation dataset which are sequential in nature. clustering in web usage mining is finding the groups which share common interests and behavior by analyzing the data collected in the web servers. This study contributes the topic clustering of web usage data and shows the interests and behaviors of the various user visits. KEYWORDS:webusagemining,sequences,set similarity,sequence similarity,