Lesson 4.2 · Redundancy Is Not Backup
This lesson teaches RAID — combining disks so your server survives a drive failure — and then spends equal effort making sure you never confuse it with backup. That confusion is one of the most expensive mistakes in the industry: organizations have lost everything because they thought “we have RAID” meant “our data is safe.” It doesn’t. RAID and backup solve different problems, and a serious setup has both.
The problem RAID solves
Section titled “The problem RAID solves”Disks fail. In Lesson 4.1, SMART warns you one might be dying — but sometimes a disk just dies, now, with no warning. If your data lives on that one disk, it’s gone (until you restore from backup, with downtime). RAID (Redundant Array of Independent Disks) combines multiple physical disks so that the failure of a disk doesn’t take your data offline. The system keeps running; you replace the dead disk; life continues.
The key word is running — RAID is about availability, staying up through a hardware failure. Hold that thought; it’s the whole distinction with backup.
The RAID levels you need to know
Section titled “The RAID levels you need to know”There are several RAID “levels”; you need to understand the trade-off each makes between capacity, performance, and redundancy. The three that matter:
graph LR
subgraph RAID0["RAID 0 · striping — NO redundancy"]
direction TB
A0[Disk 1: half the data]
B0[Disk 2: other half]
end
subgraph RAID1["RAID 1 · mirroring"]
direction TB
A1[Disk 1: all data]
B1[Disk 2: identical copy]
end
subgraph RAID5["RAID 5 · striping + parity"]
direction TB
A5[Disk 1: data]
B5[Disk 2: data]
C5[Disk 3: parity]
end
| Level | How it works | Survives | Capacity cost | Use when |
|---|---|---|---|---|
| RAID 0 | Data split (“striped”) across disks for speed | Nothing — one disk dies, all data lost | None (all usable) | You want speed and have backups; never for data you care about |
| RAID 1 | Every disk holds an identical copy (mirror) | 1 disk failure (in a 2-disk mirror) | Half (2 disks = 1 disk of space) | Simple, robust redundancy — great for a homelab |
| RAID 5 | Data + parity striped across 3+ disks | 1 disk failure | One disk’s worth | More usable space across many disks |
| RAID 6 | Like 5 but double parity | 2 disk failures | Two disks’ worth | Larger arrays where rebuild risk matters |
Parity, in RAID 5/6, is a clever bit of math: extra information that lets the array reconstruct a failed disk’s contents from the surviving disks. It’s how RAID 5 gives you single-disk redundancy while only “spending” one disk of capacity across the whole array.
How you’ll actually build it
Section titled “How you’ll actually build it”Two common routes on Linux, both fine for the homelab:
mdadm— Linux’s software RAID. It works with any disks and sits below the filesystem: you create an array (e.g./dev/md0) from your disks, then put a filesystem (or LVM) on it. Mature and universal.- ZFS — a combined filesystem and volume manager that does RAID (it calls it RAIDZ / mirror) itself, with major extra benefits (below). Increasingly the homelab favorite.
You’ll build a two-disk mirror and deliberately fail a disk in Lab 2 — watching an array degrade and rebuild is the moment RAID stops being abstract.
ZFS: redundancy plus integrity
Section titled “ZFS: redundancy plus integrity”ZFS deserves special mention because it solves problems plain RAID doesn’t, and homelabbers love it. Beyond mirroring/RAIDZ, ZFS gives you:
- Checksums on every block. ZFS detects silent corruption (“bit rot” — data quietly degrading on disk) that ordinary RAID can’t even see, and with redundancy, repairs it automatically. This is a genuinely bigger deal than it sounds: plain RAID will happily mirror corrupted data.
- Scrubs. A periodic
zpool scrubreads everything and fixes detected errors against the redundant copies — proactive health-checking of your data. - Snapshots. Near-instant, space-efficient point-in-time copies of a filesystem (more on why that matters below and in Lesson 4.4).
The trade-off is that ZFS wants RAM and a bit more learning. For a 16 GB micro PC it’s very
usable, and the data-integrity guarantees are why it’s worth the effort. If you’d rather keep it
simple first, an mdadm mirror + ext4 is perfectly respectable; you can graduate to ZFS later.
The distinction that defines this module
Section titled “The distinction that defines this module”Now the point everything has been building toward. RAID and backup feel similar — both involve “extra copies” — but they defend against completely different disasters:
| RAID / redundancy | Backup | |
|---|---|---|
| Protects against | A disk hardware failure | Deletion, corruption, ransomware, mistakes, fire/theft |
| Keeps you | Running (no downtime) | Recoverable (restore, with downtime) |
| When a file is deleted | Instantly deleted on all mirrored disks | Still safe in yesterday’s backup |
| When ransomware hits | Encrypts the data on all disks equally | Clean copy survives, if it’s offline/immutable |
| Copies are | Live, synchronized, same place | Separate, point-in-time, ideally elsewhere |
The killer insight: RAID copies happen instantly and everywhere. If you rm an important
file, or ransomware encrypts it, or a bad script corrupts it — RAID faithfully replicates the
deletion/corruption to every disk in the array, immediately. Mirroring a mistake doesn’t undo
it. RAID has zero memory of “yesterday”; a backup is precisely a memory of yesterday.
RAID is not optional and not sufficient. It handles the disk-dies case elegantly so you don’t take downtime for a hardware fault. But the moment the threat is anything other than hardware — a human, a bug, malware, a fire — only a real backup saves you. That’s the next lesson, and it’s the one with the graded “delete something and restore it” drill.
Quick self-check
Section titled “Quick self-check”- What single problem does RAID solve, and what’s the key word for what it gives you?
- Why is RAID 0 not redundancy? What is it for?
- What’s the difference between RAID 1 and RAID 5 in what they cost and what they survive?
- What does ZFS give you beyond RAID, and why does “checksums on every block” matter?
- You accidentally
rm -rfan important directory on a RAID 1 mirror. Is the data recoverable from the mirror? Why or why not? - State the RAID-vs-backup distinction in one sentence.