Feeds

EMC on XtremIO SSD brickup ballsup: Its LIFETIME downtime is under 3 minutes

Replace a failed X-Brick SSD once every 5 years

Gartner critical capabilities for enterprise endpoint backup

No DOAs here: EMC’s XTremIO arrays are expected to have less than three minutes downtime in their rated life, with X-Brick component SSDs failing once every five years or so.

We have reported that the DOA rate was too high and the co-founder and general manager of EMC's acquired XtremIO business, Ehud Rokach, has written blog which is partly a counter to that, writing that our “article presented speculations that may lead to false conclusions.”

He says that, using data obtained by monitoring “hundreds of XtremIO X-Bricks at customer sites globally and across all major verticals”:

  • XtremIO delivers world class 99.9999 per cent (six nines) field-proven availability (less than 32 Seconds of unavailability in a year, and less than 3 minutes of unavailability over the lifetime of the product.)
  • Our SSD Mean Time Between Part Replacement (MTBPR) was field-measured to be 922,240 hours, or 105 years.
  • Our Annual Replacement Rate (ARR) for SSDs was field-measured to be 0.009. For an entire X-Brick (holding 25 SSDs), the probability of encountering SSD failure at any time during a 1-year period equals (1-0.991^25), or 0.2.
  • A 0.2 ARR means that on average, based on our actual field data, you’ll need to replace a failed SSD (Due to a non-endurance related failure. …) in an X-Brick roughly once every 5 years.

Pretty darn convincing. He hammers away: “Our actual measured field performance demonstrates exceptional SSD and array-level reliability. Since initiating XtremIO’s Directed Availability program we have seen a grand total of single-digit SSD failures out of thousands of deployed SSDs.”

To refresh your memory, in the comment we reproduced from Xtremio chief techie Robin Ren, he said: “I am not too happy about our field hardware failure rate for many reasons. However, the vast majority of failures – we have seen over 150 X-Bricks so far – [pauses] in real customer environments … [and] another 200 systems internally. I think we have seen a lot of DOAs in terms of drives.”

Ren spoke in October by the way, several months after Directed Availability started.

Rokach says our story, based on Ren’s comments, “referenced (unknowingly) a couple of early DOA SSD failure events during Beta, prior to product being released for Directed Availability. The two failures during pre-release Beta were analysed, and corrective action applied (firmware update). Not surprisingly, ever since we started Directed Availability and to this very day, we have seen no excess SSD failures of any kind (in the field or DOA). This is indeed a non-issue.”

Happy to hear it. ®

Secure remote control for conventional and virtual desktops

More from The Register

next story
The Return of BSOD: Does ANYONE trust Microsoft patches?
Sysadmins, you're either fighting fires or seen as incompetents now
Microsoft: Azure isn't ready for biz-critical apps … yet
Microsoft will move its own IT to the cloud to avoid $200m server bill
US regulators OK sale of IBM's x86 server biz to Lenovo
Now all that remains is for gov't offices to ban the boxes
Flash could be CHEAPER than SAS DISK? Come off it, NetApp
Stats analysis reckons we'll hit that point in just three years
Oracle reveals 32-core, 10 BEEELLION-transistor SPARC M7
New chip scales to 1024 cores, 8192 threads 64 TB RAM, at speeds over 3.6GHz
Object storage bods Exablox: RAID is dead, baby. RAID is dead
Bring your own disks to its object appliances
Nimble's latest mutants GORGE themselves on unlucky forerunners
Crossing Sandy Bridges without stopping for breath
prev story

Whitepapers

Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Top 10 endpoint backup mistakes
Avoid the ten endpoint backup mistakes to ensure that your critical corporate data is protected and end user productivity is improved.
Top 8 considerations to enable and simplify mobility
In this whitepaper learn how to successfully add mobile capabilities simply and cost effectively.
Rethinking backup and recovery in the modern data center
Combining intelligence, operational analytics, and automation to enable efficient, data-driven IT organizations using the HP ABR approach.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.