A 2/sup d/-tree-based blocking method for microaggregating very large data sets
Agustí Solanas, Antoni Martínez-Ballesté, J. Domingo-Ferrer, Josep M. Mateo‐Sanz · 2006
Blocking is a well-known technique used to partition a set of records into several subsets of manageable size. The standard approach to blocking is to split the records according to the values of one or several attributes (called blocking attributes). This paper presents a new blocking method based on 2/sup d/-trees for intelligently partitioning very large data sets for micro aggregation. A number of experiments has been carried out in order to compare our method with the most typical univariate one.