Abstract
The enormous growth of digital data, especially the data in unstructured format has brought a tremendous challenge on data analysis as well as the data storage systems which are essentially increasing the cost and performance of the backup systems. The traditional systems do not provide any optimization techniques to keep the duplicated data from being backed up. Deduplication of data has become an essential and financial way of the capacity optimization technique which replaces the redundant data. The following paper reviews the deduplication process, types of deduplication and techniques available for data deduplication. Also, many approaches proposed by various researchers on deduplication in Big data storage systems are studied and compared.
Publisher
Blue Eyes Intelligence Engineering and Sciences Engineering and Sciences Publication - BEIESP
Subject
Electrical and Electronic Engineering,Mechanics of Materials,Civil and Structural Engineering,General Computer Science
Cited by
6 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献
1. Fog-assisted de-duplicated data exchange in distributed edge computing networks;Scientific Reports;2024-09-04
2. A Hash Based De duplication to optimize Medical Image Storage;2023 International Conference on Research Methodologies in Knowledge Management, Artificial Intelligence and Telecommunication Engineering (RMKMATE);2023-11-01
3. Data Deduplication Using Python;2023 7th International Conference On Computing, Communication, Control And Automation (ICCUBEA);2023-08-18
4. Big Data De-duplication Using Classification Scheme based on Histogram of File Stream;2022 International Conference on Intelligent Technology, System and Service for Internet of Everything (ITSS-IoE);2022-12-03
5. Big Data Backup Deduplication : A Survey;International Journal of Scientific Research in Science, Engineering and Technology;2022-07-05