Matching Data Records Among Multi Data Sources Based on Clustering Techniques

Tang Yi-fang · Mini-micro Systems · 2005

This paper put forward an algorithm, by using the canopy clustering technique which focuses on large data set, to match data records among multi data sources. The algorithm is a kind of domain-independent method, and compare to other model, when it promises the algorithm's accuracy, this method increases the effectiveness.

Read the paper · More papers on PaperTik