I'm having some issues with a particular VDI datastore. The current data store is:
1. 5 x 300GB FC Drives RAID5
2. Hosts are 2Gb FC multipath RR to my EMC Clariion CX4-240
3. View Desktops are XP SP3 running basic office and terminal emulation software
4. Average 60 Desktops on this datastore all users have between 2GB-3GB of memory.
We are having sporadic issues where in the middle of the day, the datastore Latency spikes up to around 50ms, hit's a very high queue, etc. This does not happen every day, and based on load/usage, it's sporadic. Monday is our heaviest day, with Friday being our lightest, but it could happen on Friday very easily or any day during the week.
It's happened twice over the last two weeks. Nothing has changed, the amount of users has been the same..etc VDI desktop IOPs haven't changed from the average daily I was seeing over the last couple of months or so.
The alarms show External i/o workload. We are not pusing DAT's or WSUS updates during this window etc. This is also beyond what I see during our small bootstorms, like first thing in the morning where I see it get higher amount of latency then normalize. This actually freezes up the VM's where they all have to manually reset and then it goes away and normalizes again. This is typically what I see in the morning:

When the issue occurrs, it stays at upwards of 70ms of latency..etc and just sits there until it freezes the VM's.
I have a ticket in with EMC at the moment to see if we have disk issues but they haven't found anything. VMware is also not seeing anything, of course, the problem is getting historical data so i'm going to setup a syslog server so that they can see if first hand.
Any input or thoughts on my next course of action?
1. 5 x 300GB FC Drives RAID5
2. Hosts are 2Gb FC multipath RR to my EMC Clariion CX4-240
3. View Desktops are XP SP3 running basic office and terminal emulation software
4. Average 60 Desktops on this datastore all users have between 2GB-3GB of memory.
We are having sporadic issues where in the middle of the day, the datastore Latency spikes up to around 50ms, hit's a very high queue, etc. This does not happen every day, and based on load/usage, it's sporadic. Monday is our heaviest day, with Friday being our lightest, but it could happen on Friday very easily or any day during the week.
It's happened twice over the last two weeks. Nothing has changed, the amount of users has been the same..etc VDI desktop IOPs haven't changed from the average daily I was seeing over the last couple of months or so.
The alarms show External i/o workload. We are not pusing DAT's or WSUS updates during this window etc. This is also beyond what I see during our small bootstorms, like first thing in the morning where I see it get higher amount of latency then normalize. This actually freezes up the VM's where they all have to manually reset and then it goes away and normalizes again. This is typically what I see in the morning:

When the issue occurrs, it stays at upwards of 70ms of latency..etc and just sits there until it freezes the VM's.
I have a ticket in with EMC at the moment to see if we have disk issues but they haven't found anything. VMware is also not seeing anything, of course, the problem is getting historical data so i'm going to setup a syslog server so that they can see if first hand.
Any input or thoughts on my next course of action?
Last edited: