EMC on XtremIO SSD brickup ballsup: Its LIFETIME downtime is under 3 minutes

Replace a failed X-Brick SSD once every 5 years

By Chris Mellor

4th December 2013

No DOAs here: EMC’s XTremIO arrays are expected to have less than three minutes downtime in their rated life, with X-Brick component SSDs failing once every five years or so.

We have reported that the DOA rate was too high and the co-founder and general manager of EMC's acquired XtremIO business, Ehud Rokach, has written blog which is partly a counter to that, writing that our “article presented speculations that may lead to false conclusions.”

He says that, using data obtained by monitoring “hundreds of XtremIO X-Bricks at customer sites globally and across all major verticals”:

Pretty darn convincing. He hammers away: “Our actual measured field performance demonstrates exceptional SSD and array-level reliability. Since initiating XtremIO’s Directed Availability program we have seen a grand total of single-digit SSD failures out of thousands of deployed SSDs.”

To refresh your memory, in the comment we reproduced from Xtremio chief techie Robin Ren, he said: “I am not too happy about our field hardware failure rate for many reasons. However, the vast majority of failures – we have seen over 150 X-Bricks so far – [pauses] in real customer environments … [and] another 200 systems internally. I think we have seen a lot of DOAs in terms of drives.”

Ren spoke in October by the way, several months after Directed Availability started.

Rokach says our story, based on Ren’s comments, “referenced (unknowingly) a couple of early DOA SSD failure events during Beta, prior to product being released for Directed Availability. The two failures during pre-release Beta were analysed, and corrective action applied (firmware update). Not surprisingly, ever since we started Directed Availability and to this very day, we have seen no excess SSD failures of any kind (in the field or DOA). This is indeed a non-issue.”

