UNSOLVED

rbocchinfuso

updated

12 years ago

0

3699

November 17th, 2014 12:00

Data Domain sfs_dump question

I have a question about the data produced by sfs_dump.  I want some clarification to make sure I am calculating a few things correctly.

I have a DD installation where the dedup ratio seems to be less than impressive so I set out to classify workloads and attempt to determine at a more granular level are are suffering from poor deduplication.

Note:  Because of the the directory structure on the DD the workloads were easily parsed and aggregated in a pivot table.

  1. I captured sfs_dump data from the DD in question
  2. I parsed the sfs_dump data and produced the following output (sample):

2014-11-17 14_57_40-DD6402_Details_Dedup_Analysis_11-13-2014.xlsx - Excel.png

From the above raw data I produced the following pivot table to to aggregate the statistics:

2014-11-17 14_49_02-DD6402_Details_Dedup_Analysis_11-13-2014.xlsx - Excel.png

My questions are as follow:

  1. Am I calculating the comp_% properly?
    1. =(([@[pre_lc_size]]-[@[post_lc_size]])/[@[pre_lc_size]])
  2. Am I calculating the dedupe ratio properly?
    1. =[@[pre_lc_size]]/[@[post_lc_size]]

Note:  I reference the folowing link when creating the formulas:  Understanding DataDomain Compression

The documentation link references the following two formulas:

  • Total-Comp Factor = Pre-Comp / Post-Comp             
  • Reduction % = ((Pre-Comp - Post-Comp) / Pre-Comp) * 100

Note:  I did not * by 100 because I formatted the workbook column to %


Thanks in advance for any insights.