Capt.Frito
n00b
- Joined
- Mar 18, 2013
- Messages
- 30
I need help desperately!. The situation is not a common one so hopefully I will be able to convey the situation clearly enough to figure it out. In the end I need some help trying to restore these RAID sets since it involves both the primary data RAID set and the backup set!
I have two Areca 1882ix SAS adapters connecting to an external drive enclosure by Astek. The enclosure can hold 16 drives in total, but it is divided into two groups of 8 drives each. Each group is on its own dedicated SAS bus and its own dedicated Areca 1882ix, and the these two independent cards are located in the same server managed by a single (non-VM), which is Linux (Gentoo) kernel 3.6.11 using the standard Areca "arcmsr" kernel driver. Things have been fine for a year or so.
Yesterday, I plugged the RAIDs into the Areca cards (I use the external 8088 interfaces) as I normally do. One of the RAIDs registered as "failed". I am 100% convinced that all the drives are fully functional. (Note: These RAIDs are not attached permanently -- I hotplug them when I need them... I cannot leave them attached through boot because they are not system boot drives -- and no matter what I do -- either Grub or the kernel prefers to enumerate hd0 from SAS instead of SATA first so it will find the wrong one when the SAS drives are present.. this is another story for another time).
Here's the thing: The two drives missing from the original RAID set were marked as "failed" but appear in the RAID set heirarchy as their own volume set with the same RAID set name. It seems that for some reason these two drives got split out from the array but that the data on them should be still good. I've read about these things happening with other Areca systems.
Here is a link to a screen capture of the failed RAID set heirarchy. Don't let SansDigital's web gui re-branding fool you, it's for sure an Areca 1882:
http://www.enlightenment.org/ss/e-51473dcb4aa081.86408241.jpg
Here is what the other (good) RAID set attached to the second 1882 looks like -- it was created identically to the first set, and it is how the one above should look:
http://www.enlightenment.org/ss/e-514744de4cf025.30854788.jpg
The question is: How can I just reassociate the drives to see if the data is still good? I can't see how to do this wiithout breaking things completely. It's really important to have the data. I have read in several places to just recreate/rescue the array but just don't "init" it, but I can see how to do this -- the gui just keeps reporting that the RAID set already exists!
Before anyone starts thinking that I should have backed up the data, well I did... I backed it up to a similar storage box (but made by SansDigital, not Astek). The SansDigital box is divided into two groups of 12 ea 600GB WD RE drives, whereas the Astek one uses two groups of 8 ea 300GB Seagate's.
When I plugged in the SansDigital drive box, it too did the same thing! It marked two drives as bad and "failed" the array. This is why I think the drives themselves have not failed physically -- way too coincidental: two different mfgrs (Seagate vs WD), two different box makers with different SAS expanders (SansDigital vs Astek).
I realize that this may be Areca the card, I am not too sure about that either, I was not paying attention to which 8088 port I plugged the first array nto -- so there's a 50/50 chance it was a different card each time. I'd try it again but I obviously don't want to make anything any worse at this point.
I do have a Wiindows XP 32bit box with a single 1882ix card in it that I could use to fix things if anyone believes that the OS is to blame (the RAIDs are formatted NTFS because I do use the data on Windows systems too... I much prefer ReiserFS, but this is another topic for another day too...)
TIA,
I have two Areca 1882ix SAS adapters connecting to an external drive enclosure by Astek. The enclosure can hold 16 drives in total, but it is divided into two groups of 8 drives each. Each group is on its own dedicated SAS bus and its own dedicated Areca 1882ix, and the these two independent cards are located in the same server managed by a single (non-VM), which is Linux (Gentoo) kernel 3.6.11 using the standard Areca "arcmsr" kernel driver. Things have been fine for a year or so.
Yesterday, I plugged the RAIDs into the Areca cards (I use the external 8088 interfaces) as I normally do. One of the RAIDs registered as "failed". I am 100% convinced that all the drives are fully functional. (Note: These RAIDs are not attached permanently -- I hotplug them when I need them... I cannot leave them attached through boot because they are not system boot drives -- and no matter what I do -- either Grub or the kernel prefers to enumerate hd0 from SAS instead of SATA first so it will find the wrong one when the SAS drives are present.. this is another story for another time).
Here's the thing: The two drives missing from the original RAID set were marked as "failed" but appear in the RAID set heirarchy as their own volume set with the same RAID set name. It seems that for some reason these two drives got split out from the array but that the data on them should be still good. I've read about these things happening with other Areca systems.
Here is a link to a screen capture of the failed RAID set heirarchy. Don't let SansDigital's web gui re-branding fool you, it's for sure an Areca 1882:
http://www.enlightenment.org/ss/e-51473dcb4aa081.86408241.jpg
Here is what the other (good) RAID set attached to the second 1882 looks like -- it was created identically to the first set, and it is how the one above should look:
http://www.enlightenment.org/ss/e-514744de4cf025.30854788.jpg
The question is: How can I just reassociate the drives to see if the data is still good? I can't see how to do this wiithout breaking things completely. It's really important to have the data. I have read in several places to just recreate/rescue the array but just don't "init" it, but I can see how to do this -- the gui just keeps reporting that the RAID set already exists!
Before anyone starts thinking that I should have backed up the data, well I did... I backed it up to a similar storage box (but made by SansDigital, not Astek). The SansDigital box is divided into two groups of 12 ea 600GB WD RE drives, whereas the Astek one uses two groups of 8 ea 300GB Seagate's.
When I plugged in the SansDigital drive box, it too did the same thing! It marked two drives as bad and "failed" the array. This is why I think the drives themselves have not failed physically -- way too coincidental: two different mfgrs (Seagate vs WD), two different box makers with different SAS expanders (SansDigital vs Astek).
I realize that this may be Areca the card, I am not too sure about that either, I was not paying attention to which 8088 port I plugged the first array nto -- so there's a 50/50 chance it was a different card each time. I'd try it again but I obviously don't want to make anything any worse at this point.
I do have a Wiindows XP 32bit box with a single 1882ix card in it that I could use to fix things if anyone believes that the OS is to blame (the RAIDs are formatted NTFS because I do use the data on Windows systems too... I much prefer ReiserFS, but this is another topic for another day too...)
TIA,