Chunking Algorithm for Data deduplication

Dharmashankar Subramanian · 2014

This paper presents an structure and algorithm for a deduplication method which can be expeditiously used for removing similar data between files existing on different machines with high rate and performing it within rapid time. The algorithm anticipate similar parts between source and destination files very fast, and then check the identical parts and transfers only those parts of blocks that proved to be in unique region. The primal aspect is reaching faster high scalability and determining duplicate result is that data are carried as fixed-size block chunks which are distributed by chunk's to Index-table both side boundary values. Index- is a fixed sized table structure; chunk's boundary byte values are used as their cell row and column numbers. Tested result shows that the given solution minimizes data storage capacity and increases data deduplication performance extensively

Read the paper · More papers on PaperTik