
UNSOLVED
R
rbocchinfuso
3 Posts
0
3699
November 17th, 2014 12:00
Data Domain sfs_dump question
I have a question about the data produced by sfs_dump. I want some clarification to make sure I am calculating a few things correctly.
I have a DD installation where the dedup ratio seems to be less than impressive so I set out to classify workloads and attempt to determine at a more granular level are are suffering from poor deduplication.
Note: Because of the the directory structure on the DD the workloads were easily parsed and aggregated in a pivot table.
- I captured sfs_dump data from the DD in question
- I parsed the sfs_dump data and produced the following output (sample):
From the above raw data I produced the following pivot table to to aggregate the statistics:
My questions are as follow:
- Am I calculating the comp_% properly?
- =(([@[pre_lc_size]]-[@[post_lc_size]])/[@[pre_lc_size]])
- Am I calculating the dedupe ratio properly?
- =[@[pre_lc_size]]/[@[post_lc_size]]
Note: I reference the folowing link when creating the formulas: Understanding DataDomain Compression
The documentation link references the following two formulas:
- Total-Comp Factor = Pre-Comp / Post-Comp
- Reduction % = ((Pre-Comp - Post-Comp) / Pre-Comp) * 100
Note: I did not * by 100 because I formatted the workbook column to %
Thanks in advance for any insights.
Responses (1)
Solutions (0)

mikuszed
90 Posts
2418
0
Posted November 24th, 2014 09:00
Rich,
Are you trying to average out the data you got from the sfs_dump by compression and deduplication? Here's something I've used in the past to understand compression: https://support.emc.com/kb/180554.
I just found this one, but haven't used it yet: https://support.emc.com/kb/181054