Cloud Storage Performance Improvement Using Deduplication and Compression Techniques
Pradeep Mohan Kumar, E. Pugazhendhi, Rudra Kalyan Nayak · 2022 4th International Conference on Smart Systems and Inventive Technology (ICSSIT) · 2022
Cloud Computing, an internet based computing is composed of various applications and hardware deployed to the end-user. Existing system deals with block level de-duplication. In this deduplication process, unique chunks of data or byte patterns are identified and stored in the cloud file system. Then the other chunks are compared to the stored copy and whenever a match occurs, the redundant chunk is replaced with a small reference that points to the stored chunk. But in case of block-level deduplication maintenance of large number of blocks is highly difficult. It also requires high processing power when compared to other deduplication techniques. This is the reason why File-level Deduplication comes into picture. Replication is the presence of same file in all servers. This technique increases the efficiency by making the data available at all time. Replicating data avoids single point failure. At the same time it leads to wastage of space. For increasing the cloud storage efficiency, deduplication and compression techniques are adapted. Our approach to deduplication and compression aims at enhancing the efficiency of cloud storage. Deduplication, specialized data compression technique stores only single copy of the redundant data. Detecting the presence of duplicate files and storing it in compressed format in all other servers other than the original server improves the efficiency of cloud storage. In this approach, Deduplication is done at file level. Files present in the compressed format can be uncompressed (unzipped) and downloaded by the user, hence maximizing the storage efficiency and increasing the availability.