Classification via k-means clustering and distance-based outlier detection
Surasit Songma, Witcha Chimphlee, Kiattisak Maichalernnukul, Parinya Sanguansat · 2012
We propose a two-phase classification method. Specifically, in the first phase, a set of patterns (data) are clustered by the k-means algorithm. In the second phase, outliers are constructed by a distance-based technique and a class label is assigned to each pattern. The Knowledge Discovery Databases (KDD) Cup 1999 data set, which has been utilized extensively for development of intrusion detection systems, is used in our experiment. The results show that the proposed method is effective in intrusion detection.