Data Deduplication Techniques for Big Data Storage Systems

DOI: https://doi.org/10.35940/ijitee.j9129.0881019
2019-08-10
VOLUME 8 ISSUE 10, AUGUST 2019, REGULAR ISSUE
Abstract:The enormous growth of digital data, especially the data in unstructured format has brought a tremendous challenge on data analysis as well as the data storage systems which are essentially increasing the cost and performance of the backup systems. The traditional systems do not provide any optimization techniques to keep the duplicated data from being backed up. Deduplication of data has become an essential and financial way of the capacity optimization technique which replaces the redundant data. The following paper reviews the deduplication process, types of deduplication and techniques available for data deduplication. Also, many approaches proposed by various researchers on deduplication in Big data storage systems are studied and compared.
What problem does this paper attempt to address?