I-diversity anonymity method based on clustering
Yuchen Xu · Journal of Yanshan University · 2012
The traditional data generalization strategy of I-diversity model is based on the concept-hierarchy structure,but this kind of data generalization strategy may cause some unnecessary loss of information as taking measure of anonymous protection to sensitive attributes.To solve this problem,the technique of cluster for data anonymity is adopted and a corresponding anonymous protection method is proposed.Under the constraint condition of-diversity model,the new method makes partition of tuple according to the hierarchical clustering algorithm based on distance,take different generalization strategy for different kinds of identifiers,and describe the loss of information caused by data generalization according to the change of uncertainty degree of attributes.By contrast with the original model,the new model proposed in this paper performs better than in the protection of customer's sensitive attribute and can reduce the information loss caused by the generalization to some degree.