Data Summarisation by Typicality-based Clustering for Vectorial and Non Vectorial Data
Marie‐Jeanne Lesot, Rudolf Kruse · 2006
In this paper, a typicality-based clustering algorithm is proposed: it exploits typicality degrees defined in a prototype construction framework to identify a decomposition of the dataset into homogeneous and distinct clusters and to provide characteristic representatives of the obtained clusters, so as to summarise the initial dataset. The proposed algorithm can be applied both to vectorial and non vectorial data, such as trees for instance. Tests performed on artificial and real data illustrate the interest of the proposed approach.