Data deduplication is a developing innovation that presents a decrease of storage use and an efficient method for dealing with data replication in the reinforcement condition. In cloud data storage, deduplication strategy improves the efficiency of storage by doing a noteworthy process in virtual machine setup, data sharing network, organized and unorganized data handling. The exponential volume of digital data in cloud storage is a critical issue, and duplication is inevitable that demands additional storage space. This issue can be handled efficiently by implementing de-duplication techniques which eliminates redundant data and increases storage utilization. Existing systems generate 128-bit or 160-bit hash values, respectively, using MD5 or SHA algorithms to implement the deduplication approach that can be used to identify and discard redundant data. Therefore, additional memory space is required to store this hash value. In this work, proposed another design utilizing the Distributed Storage Hash Algorithm (DSHA) for file-based data deduplication. As the hash value is changed by using the proposed method, the amount of memory space occupied by the hash value has reduced and also it improves data read/write performance. The additional use of the new mechanism, which hides file information in text cover message, further strengthens security. Upon comparing different deduplication algorithms, the proposed model displays significant improvement in deduplication, Cover file generation, Secret Key Generation, Key Conversion and Combine, File Deduplication, Mixing Operation, Secret Message Extraction.
Volume 12 | 07-Special Issue
Pages: 652-663
DOI: 10.5373/JARDCS/V12SP7/20202155