UNSOLVED

mmarzotto

updated

16 years ago

M

mmarzotto

14 Posts

0

1119

December 14th, 2010 13:00

Celerra replicator performance

Good Afternoon,

I have been troubleshooting performance issues with EMC support around my Celerra Replicator V2.

I have a 2TB filesystem that houses nightly database backups and setup a replication job to my remote site Celerra.

However, the inital copy cannot even finish within 36 hours and fails because the checkpoint volume cannot expand anymore because of the high level of change from the nightly backups.

I am currently getting 3000-5000KB/sec throughput on the replication pipe.

I have an NLAN 100 meg connection from source to target data centers (no WAN routers and no WAN acceleration - very simple setup). I have 3750g switches that the data movers are plugged into at each site (same exact code installed in all of them). The ports the data movers are plugged into are set to 1000FULL as well as the data mover cge ports set to 1000FULL. There are zero errors on all the ports from a switch level and the ports are not even being pushed hard throughout the entire switch.

I also have Recoverpoint replicating over this same line but it is using only about 30% of the 100 meg pipe, so there is plenty of room for the Celerra to claim throughput but it simply is not pushing the data out of the cge ports fast enough. (recoverpoint is able to take up the entire pipe when needed if it ever has to do a full rescan - which is very very rare for me)

I have tried failing over the data movers and even changing the ports the data movers are plugged into but I get no performance increase. I have changed the fastRTO setting to '1' and I still have not seen any performance boost.

Is 3-5MB/sec a normal for the replication throughput? It seems VERY low. The throughput I am getting on the other cge ports that house CIFS shares are getting 40-50MB/sec throughput...so to me it seems that the replication piece of the Celerra is not getting the data out of the cge ports fast enough.

For reasons beyond me, the support tech I continue to troubleshoot with continues to think that it is the network as the underlying problem even though I have proven it time and time again that it isn't.

Is there anything else that can be done to try to increase performance for the replication?

Thanks