Learning Hub / Kubernetes & Platform
Kubernetes Storage & Data Protection
Practitioner → Advanced7 lessonsAvailable
Stateful workloads on Kubernetes without fear: how CSI works, running Longhorn and Ceph, performance tuning, and backup/restore you have tested.
You'll meetCSILonghornRook-CephsnapshotsVeleroaccess modesStatefulSets
Start lesson 01 →
What you'll be able to do
- Choose between local, replicated and shared storage with clear trade-offs
- Operate Longhorn and Rook-Ceph, including replica rebuilds
- Run Velero backups and prove restores
Before you start
Kubernetes Administration Level 2.
How it works
Each lesson: plain-language idea → how it really works → hands-on. Each section ends with a cheat sheet & self-check.
Curriculum
Lessons marked “Read” are ready; the rest are on the way.
Modules
- 01CSI architectureController, node plugin, provisioning flowRead →
- 02Longhorn in practiceInstall, replication, self-healingRead →
- 03Rook-CephRBD, CephFS, failure domainsRead →
- 04Snapshots & backupsVolumeSnapshots, Velero, object store targetsRead →
- 05Storage performancefio, policies, noisy neighboursRead →
- 06Stateful patternsDatabases on Kubernetes: when and howRead →
- 07Production patternsCapacity, upgrades, monitoringRead →
- 📋Cheat sheet & self-checkEvery command from this section on one page, then 21 questions to check yourself.Open →
Real-world scenarios
Work through each one: symptom → misleading signal → evidence → root cause → prevention.
Volume stuck in Multi-Attach after a node failure
Why the pod can't start elsewhere, and how to recover safely.
Backups ran every night; the restore failed
The restore drill nobody ran.
This site is a public version of my personal engineering knowledge hub. It intentionally excludes confidential company information and internal operational details.