A survey on DE – Duplication schemes in cloud servers for secured data analysis in various applications
K. Pragash, J. Jayabharathy · Measurement Sensors · 2022
Data deduplication in a system perspective is termed to be a single or multiple copy of an original data which could increase the computational complexity while accessing such a data. In clinical terms the deduplication is considerably the multiple copy of a same sample. The data deduplication is also vulnerable towards security issues which in turn could affect the performance of a cloud server. This type of several copies had to be handled not only to analyze the data but also to identify the underlying patterns which could support in prediction and classification of the data. There are several deduplication handling schemes in multiple domains had been proposed to handle the repetition of the data from which the most effective and recently proposed schemes had been considered for survey. The deduplication schemes are compared in terms of their performance and their pros and cons are discussed in this article which could pave a path for researchers to propose and perform a better analysis with the deduplication issues.