Data deduplication techniques for big data storage systems

Citations of this article
Mendeley users who have this article in their library.
Get full text


The enormous growth of digital data, especially the data in unstructured format has brought a tremendous challenge on data analysis as well as the data storage systems which are essentially increasing the cost and performance of the backup systems. The traditional systems do not provide any optimization techniques to keep the duplicated data from being backed up. Deduplication of data has become an essential and financial way of the capacity optimization technique which replaces the redundant data. The following paper reviews the deduplication process, types of deduplication and techniques available for data deduplication. Also, many approaches proposed by various researchers on deduplication in Big data storage systems are studied and compared.




Sharma, N., Krishna Prasad, A. V., & Kakulapati, V. (2019). Data deduplication techniques for big data storage systems. International Journal of Innovative Technology and Exploring Engineering, 8(10), 1145–1150.

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free