Taking work now — the first look is freeWhole sets posted in from anywhere in the UK, or handed in at ten drop-off pointsQuicker still, give us a ring:0800 6890668
RARRAID Array Data Recovery 0800 6890668 Price my job
RAR / Controllers and hosts / IBM and Lenovo storage

ServeRAID M5015 · M5210 · DS3524 · DS4700 · DS5100 · Storwize V3700 · V5000 · V7000 · FlashSystem · mdisk · pool

IBM and Lenovo storage array recovery. ServeRAID in the servers, and the arrays behind them, in outline.

IBM's storage comes in three layers that arrive here. ServeRAID cards in xSeries, System x and now Lenovo ThinkSystem servers, which are MegaRAID underneath and covered on that page. The DS3000, DS4000 and DS5000 arrays, LSI-built, whose Storage Manager reports arrays as degraded and drives as DDM failed and, on the DS4000, logical drives with unreadable sectors after a rebuild met a read error, which is a puncture by another name. And Storwize V3700, V5000 and V7000 and FlashSystem, where drives are grouped into mdisks by RAID level, mdisks pooled, and volumes carved from the pool, with the DRAID layouts of later firmware distributing spares across the members. The arrays and Storwize are enterprise storage, in outline here: the drives are imaged, most at 520 or 528 bytes a sector, and the mdisks and pools reassembled from the images. From £1,250 + VAT after the free look, fixed in writing.

Free first lookOne fixed figure in writingNo data, no bill on most jobsReturn postage paid

Rather talk it through? An engineer answers the bench line
0800 6890668

Power down. Label every slot before a drive comes out. Do not rebuild, do not force a drive online, do not import or clear a foreign configuration, do not initialise, and do not run chkdsk, fsck, zpool import -F or mdadm --create on the members. A rebuild reads every sector of every survivor and writes to the replacement; each of the commands writes to the drives. On images, all of them are reversible. On the originals, none of them is.

Models and versions we see.

Which level is it →
Family Models or versions Metadata, defaults and notes
ServeRAIDM1015, M5015, M5110, M5210, 930-8iMegaRAID underneath; that page applies
DS3000 to DS5000DS3524, DS4700, DS5100LSI-built; Storage Manager; unreadable sectors
Storwize and FlashSystemV3700, V5000, V7000; FS5000 to 9000mdisks, DRAID, pools; 520-byte drives

What tends to go wrong on IBM storage.

Unreadable sectors on a DS logical driveA rebuild that met a read error on a survivor and marked the stripes unreadable rather than failing the drive. The same thing Dell calls a puncture; the first-dropped drive holds the stripes.
Storwize mdisks and poolsDrives into mdisks, mdisks into pools, volumes out of pools. A reconstruction rebuilds each mdisk from its members and then the pool's extent map, and the drive site's 520-byte page applies to the members.

IBM and Lenovo storage symptoms, and how long each gives you.

Not listed? Describe it on the form →
Packing it and posting it: power down, photograph the controller's screen or export its log, and write the slot number on each drive with a marker before it comes out of its carrier. Send every member of the set, including the one that failed first; it holds the stripes the rebuild never reached. Each drive travels in an anti-static bag inside its own padding, in a box with nothing able to move. Send the controller if it was replaced or holds an encryption key; otherwise it stays. Insure the parcel for what the data is worth, use a tracked service, and know that the posting address is not printed anywhere on this site; it arrives by email in reply to the form, with a booking sheet to print; the sheet inside the parcel is what matches it to your enquiry when it is opened. Or hand the sealed parcel in at the nearest of ten drop-off points, your name on the outside and the sheet inside; say where you are on the form and it comes by email. We pay the postage home either way. The whole of it is written up on the guide to packing and posting.

What the message means on IBM storage.

Describe yours to us →
What you see The usual reason Where that leaves you
Virtual or logical disk degradedA member out; still readableImage the weak drives before rebuilding
Virtual or logical disk failed or offlineToo many members outEvery member imaged; set reassembled from images
Foreign configurationMetadata the controller does not claimDo not clear; import only with every member present
Predictive failure on a survivorThresholds crossedStop; image before any rebuild

From the parcel arriving to your files going back.

Work we have closed →
01

Logged the day it lands, and the first look costs nothing Free

A case number goes on the parcel and a number on every member the day it is opened, matched to the slot you wrote on it. Each member goes on the imager its interface needs, never on a controller, and its metadata is read before a sector is: the level, the order, the stripe size, the event counts that say which member is current and which dropped out first. An engineer settles what has happened to the set and how much of it can honestly be read back. Back to you come two things together: a straight note of what is liftable and what is not, plus one figure, fixed and written down. Accept it, or decline and owe us nothing.

Nothing to pay for lookingA single figure, put in writingNo rebuilds, no imports, no initialise
02

Every member imaged, including the one that failed first

Every drive in the set is imaged sector by sector, weak areas last, on hardware that controls every retry, with a map of what could not be read kept for each. Members with failed heads go to the clean bench first; that is the drive site's work and the same bench does it. The drive that dropped out first is imaged too, because it still holds every stripe written before it dropped, and a rebuild that stalled part-way never reached them.

Sector by sector, weak areas lastThe first-dropped member included
03

The geometry, from the metadata or from the parity

Where the controller's metadata survives on the members, the order, stripe size, parity rotation and data offset are read from it. Where it was cleared or overwritten, they are recovered from the data: parity across the members at the same offset should XOR to zero on a consistent stripe, which confirms the level and finds the stale member; where parity lands says the rotation; entropy at stripe edges and the file system's own anchors give the stripe size, the order and the start.

Metadata first, parity secondOrder, stripe, rotation, offset
04

Assembled in software, and repaired on the virtual volume

The set is put together from the images in software, with nothing written to any of them: the current members in, the stale member used only to fill holes a survivor could not give. The file system is checked and repaired on a copy of the virtual volume, VMFS and CSV volumes opened and the virtual machines' disks extracted, databases repaired where they need it. The originals are not touched again.

Nothing written to the imagesVirtual machines and databases opened
05

You see the file list before you pay

What was recovered is listed for you first, and only then does a bill exist. Approve the list and it is invoiced; turn it down and it is not — and where nothing has come back, most jobs carry no charge at all. Recovered data travels home on fresh media bought in for your job, with the postage at our end. Your case is not closed until you have opened the files on a machine of your own.

No charge until you accept the figureFresh media, supplied with the job5–10 days at the bench

From the bench

  • Export the controller log first. It is the best record of what failed when, and the bench reads it.
  • The drives are the job; the controller is context. Send it if it was replaced or holds an encryption key.
  • Keep the first-dropped drive. It holds the stripes the rebuild could not compute.

What helps, and what harms.

Do this much first

  • Export the log and photograph the screen
  • Power down and label every slot
  • Send every member
  • Tell us the controller, the level, the stripe size if known, and what was tried

What sets us back

  • Rebuilding with a second drive flagged
  • Clearing a foreign configuration
  • Forcing failed drives online
  • Deleting and recreating the virtual disk

Questions answered before you commit.

My Storwize pool is offline after two drives failed in one mdisk. Can the volumes be recovered?

Usually. The mdisk's members are imaged at native sector size, the mdisk reassembled from the images, and the pool's extent map read to rebuild the volumes. It is enterprise storage and priced as such.

Do I send the controller or the unit?

The controller only if it was replaced and the set is now foreign, or if it holds an encryption key. A NAS or DAS unit, only if its bridge holds the layout; the pages say which.

What does it cost?

Enterprise storage: from £1,250 + VAT after the free look, fixed in writing.

How long does it take?

5–10 days at the bench for a set of two to four; 10–15 days at the bench for larger sets.

Nothing gets worse while it is powered down.

Looking at it is free. Back comes a list of what opened and what did not, together with a single price to finish, set down in writing while you are still free to say no. On most jobs an invoice only follows the data. Until that list reaches you, leave the server off and the drives in their slots.

0800 6890668