A drive has failed, or a server will not start. This is one of the few genuinely bad days in small business IT, and it is worth knowing in advance how it is handled, because knowing the shape of it makes the day much less frightening.
First, the calm part. Hardware fails. It is expected, it is planned for, and in most cases the data is recoverable. Many servers and NAS devices are built so that one drive can fail without losing anything at all — the device keeps running on the remaining drives while the failed one is replaced. If that is your situation, you may not even notice a difference in your day.
Where the storage is readable at all, our strong preference is to take a copy of it first and then do the recovery work on the copy.
The reason is simple. A drive that is partly failing may only give you one good read before it deteriorates further. If we spend that one good read making a complete image of it, we can then attempt recovery on that image as many times as we like, and a failed attempt costs nothing. If instead we experiment directly on the original and the drive dies mid-attempt, there is no second try.
Taking the copy takes time, and it can feel like nothing is happening. It is the most valuable part of the process.
You will want a time. Everyone does, and it is a reasonable thing to want. The honest answer in the first hour is usually that we do not know yet, and anyone who gives you a confident figure at that point is guessing.
It depends on what failed, whether the data is readable, whether a replacement part is on the shelf or has to be ordered, and how much data has to be copied back — and copying takes as long as the volume of data dictates, no matter how urgent it is.
What we will do instead of guessing is give you a first assessment as soon as we have one, tell you what we know and what we do not, and update you as each unknown is resolved. Once the picture is clear, you will get a real estimate. We would rather give you an honest "not yet known" than a number that quietly falls apart at four o'clock.
Once things are running, we will write up what failed, what was recovered, anything that was lost, and what would reduce the impact next time. Sometimes that is a hardware change, sometimes it is a change to the backup, sometimes it is nothing — the arrangement worked as designed and it simply took the time it took.
Log a ticket at portal.yougrowit.com.au or email support@yougrowit.com.au. If it is stopping you working right now, call 03 9028 4358.
Support hours are Monday to Friday, 8:30am to 5:30pm Melbourne time, excluding Victorian public holidays.