Hi we have a FS5020 that had 4x 6TB drive failures within a few hours of each other, surprisingly the mdisk didn't go offline and i dont understand why, in my mind that mdisk should be a total loss because a 6TB drive would take at least 15 hours to rebuild so the stripe would have been incomplete and draid6 can only tolerate 2 failures before a 3rd one would result in parity failure.
The mdisk has a 60 members, 2 rebuild areas, redundancy of 2 and a stripe width of 12.
A colleague thought that maybe each 12 width stripe was tolerant of 2 failures in its own right? I don't think this is the case.
The drives have been replaced and copyback's are in progress.
If a drive is starting to pre-fail and firmware picks this up before it fails does it build to the spare then fail the drive?
Any comments gratefully received.
------------------------------
stuart wade
------------------------------