Keyboard shortcuts

Press or to navigate between chapters

Press S or / to search in the book

Press ? to show this help

Press Esc to hide this help

Chapter 18: Troubleshooting Playbook

⏱️ Total chapter time: ~60 min (35 min reading + 25 min exercises)

After this chapter, you will be able to: Systematically debug any broken pod, service, or storage issue in Kubernetes using a repeatable mental model and a concrete set of kubectl commands.


What’s Inside

SectionTopicTime
18.1The Debugging Mental Model~8 min
18.2Pod Failures — CrashLoopBackOff, ImagePullBackOff, OOMKilled~10 min
18.3Networking Failures — DNS, Services, Connectivity~10 min
18.4Storage and Permission Issues~7 min
18.5The Troubleshooting Cheat Sheet~5 min

Prerequisites

  • Completed Chapters 1–17 (especially Chapter 5 — Services and Chapter 8 — Storage)
  • Minikube cluster running (minikube status)

Why a “Playbook” and Not Just a List of Fixes

Kubernetes errors have patterns. The same CrashLoopBackOff might be caused by a bad environment variable, a missing Secret, a wrong command, or an OOM kill.

A playbook gives you a decision tree, not a lookup table. You learn to narrow down the cause systematically rather than guessing and applying random fixes.

💡 Tip: The most important debugging skill in Kubernetes is reading kubectl describe. Most problems announce themselves in the Events section — you just have to know to look there.