Matches in SemOpenAlex for { <https://semopenalex.org/work/W2785562034> ?p ?o ?g. }
- W2785562034 abstract "Content-Defined Chunking (CDC) detect maximum redundancy in data deduplication systems in the past years. In this research work, we focus on optimizing the deduplication system by adjusting the pertinent factors in content defined chunking (CDC) to identify as the key ingredients by declaring chunk cut-points and efficient fingerprint lookup using bucket based index partitioning. For efficient chunking, we propose Genetic Evolution (GE) algorithm based approach which is optimized Two Thresholds Two Divisors (TTTD-P) CDC algorithm where we significantly reduce the number of computing operations by using single dynamic optimal parameter divisor D with optimal threshold value exploiting the multi-operations nature of TTTD. To reduce the chunk-size variance, TTTD algorithm introduces an additional backup divisor D' that has a higher probability of finding cut-points. However, adding an additional divisor decreases the chunking throughput, meaning that TTTD algorithm aggravates Rabin's CDC performance bottleneck. To this end, Asymmetric Extremum (AE) significantly improves chunking throughput while providing comparable deduplication efficiency by using the local extreme value in a variable-sized asymmetric window to overcome the Rabin, MAXP and TTTD boundaries-shift problem. FAST CDC in the year 2016 is about 10 times faster than unimodal Rabin CDC and about 3 times faster than Gear and Asymmetric Extremum (AE) CDC, while achieving nearby the same deduplication ratio (DR). Therefore, we propose GE based TTTD-P optimized chunking to maximize chunking throughput with increased DR; and bucket indexing approach reduces hash values judgement time to identify and declare redundant chunk about 16 times than unimodal baseline Rabin CDC, 5 times than AE CDC, 1.6 times than FAST CDC. Our experimental results comparative analysis reveals that TTTD-P using fast BUZ rolling hash function with bucket indexing on Hadoop Distributed File System (HDFS) provide a comparatively maximum redundancy detection with higher throughput, higher deduplication ratio, lesser computation time and very low hash values comparison time as being best data deduplication for distributed big data storage systems." @default.
- W2785562034 created "2018-02-23" @default.
- W2785562034 creator A5052522764 @default.
- W2785562034 creator A5052527915 @default.
- W2785562034 creator A5073506604 @default.
- W2785562034 creator A5085053994 @default.
- W2785562034 date "2017-09-01" @default.
- W2785562034 modified "2023-09-22" @default.
- W2785562034 title "Genetic optimized data deduplication for distributed big data storage systems" @default.
- W2785562034 cites W1521996498 @default.
- W2785562034 cites W1542686980 @default.
- W2785562034 cites W1545493325 @default.
- W2785562034 cites W1553098517 @default.
- W2785562034 cites W1976024527 @default.
- W2785562034 cites W2008185810 @default.
- W2785562034 cites W2045202139 @default.
- W2785562034 cites W2056980397 @default.
- W2785562034 cites W2063547430 @default.
- W2785562034 cites W2066529295 @default.
- W2785562034 cites W2073370301 @default.
- W2785562034 cites W2107551255 @default.
- W2785562034 cites W2110824055 @default.
- W2785562034 cites W2112939204 @default.
- W2785562034 cites W2124632914 @default.
- W2785562034 cites W2169875292 @default.
- W2785562034 cites W2560823575 @default.
- W2785562034 cites W4230077428 @default.
- W2785562034 doi "https://doi.org/10.1109/ispcc.2017.8269581" @default.
- W2785562034 hasPublicationYear "2017" @default.
- W2785562034 type Work @default.
- W2785562034 sameAs 2785562034 @default.
- W2785562034 citedByCount "4" @default.
- W2785562034 countsByYear W27855620342020 @default.
- W2785562034 countsByYear W27855620342021 @default.
- W2785562034 countsByYear W27855620342022 @default.
- W2785562034 crossrefType "proceedings-article" @default.
- W2785562034 hasAuthorship W2785562034A5052522764 @default.
- W2785562034 hasAuthorship W2785562034A5052527915 @default.
- W2785562034 hasAuthorship W2785562034A5073506604 @default.
- W2785562034 hasAuthorship W2785562034A5085053994 @default.
- W2785562034 hasConcept C111919701 @default.
- W2785562034 hasConcept C11413529 @default.
- W2785562034 hasConcept C124101348 @default.
- W2785562034 hasConcept C152124472 @default.
- W2785562034 hasConcept C154945302 @default.
- W2785562034 hasConcept C157764524 @default.
- W2785562034 hasConcept C162319229 @default.
- W2785562034 hasConcept C173608175 @default.
- W2785562034 hasConcept C182407805 @default.
- W2785562034 hasConcept C199360897 @default.
- W2785562034 hasConcept C203357204 @default.
- W2785562034 hasConcept C2780945871 @default.
- W2785562034 hasConcept C32587265 @default.
- W2785562034 hasConcept C41008148 @default.
- W2785562034 hasConcept C555944384 @default.
- W2785562034 hasConcept C76155785 @default.
- W2785562034 hasConcept C77088390 @default.
- W2785562034 hasConceptScore W2785562034C111919701 @default.
- W2785562034 hasConceptScore W2785562034C11413529 @default.
- W2785562034 hasConceptScore W2785562034C124101348 @default.
- W2785562034 hasConceptScore W2785562034C152124472 @default.
- W2785562034 hasConceptScore W2785562034C154945302 @default.
- W2785562034 hasConceptScore W2785562034C157764524 @default.
- W2785562034 hasConceptScore W2785562034C162319229 @default.
- W2785562034 hasConceptScore W2785562034C173608175 @default.
- W2785562034 hasConceptScore W2785562034C182407805 @default.
- W2785562034 hasConceptScore W2785562034C199360897 @default.
- W2785562034 hasConceptScore W2785562034C203357204 @default.
- W2785562034 hasConceptScore W2785562034C2780945871 @default.
- W2785562034 hasConceptScore W2785562034C32587265 @default.
- W2785562034 hasConceptScore W2785562034C41008148 @default.
- W2785562034 hasConceptScore W2785562034C555944384 @default.
- W2785562034 hasConceptScore W2785562034C76155785 @default.
- W2785562034 hasConceptScore W2785562034C77088390 @default.
- W2785562034 hasLocation W27855620341 @default.
- W2785562034 hasOpenAccess W2785562034 @default.
- W2785562034 hasPrimaryLocation W27855620341 @default.
- W2785562034 hasRelatedWork W1553098517 @default.
- W2785562034 hasRelatedWork W2107551255 @default.
- W2785562034 hasRelatedWork W2110824055 @default.
- W2785562034 hasRelatedWork W2188279733 @default.
- W2785562034 hasRelatedWork W2351279544 @default.
- W2785562034 hasRelatedWork W2356209611 @default.
- W2785562034 hasRelatedWork W2481696877 @default.
- W2785562034 hasRelatedWork W2486142004 @default.
- W2785562034 hasRelatedWork W2549430943 @default.
- W2785562034 hasRelatedWork W2577472534 @default.
- W2785562034 hasRelatedWork W2766698357 @default.
- W2785562034 hasRelatedWork W2787300667 @default.
- W2785562034 hasRelatedWork W2805733941 @default.
- W2785562034 hasRelatedWork W2835458487 @default.
- W2785562034 hasRelatedWork W2890603862 @default.
- W2785562034 hasRelatedWork W2982562745 @default.
- W2785562034 hasRelatedWork W3014193728 @default.
- W2785562034 hasRelatedWork W3094771490 @default.
- W2785562034 hasRelatedWork W3157678865 @default.
- W2785562034 hasRelatedWork W3159666886 @default.
- W2785562034 isParatext "false" @default.
- W2785562034 isRetracted "false" @default.
- W2785562034 magId "2785562034" @default.