Advanced Kubernetes / Cluster Architecture is a practical engineering guide to operating production clusters that survive real operational pressure. It starts from a simple observation: Kubernetes is easy to install and hard to run, and the difference is architectural understanding rather than operational familiarity.
The book walks through the full stack — control plane internals and etcd health, node architecture and CRI runtimes, scheduler constraints and placement patterns, CNI and service mesh and Gateway API, CSI storage and StatefulSets, multi-tenancy and namespace isolation, admission control and policy engines, autoscaling from HPA to Karpenter, cost optimization and FinOps, operators and CRDs, upgrades and disaster recovery, and observability at scale.
It covers the failure modes that break production clusters: an etcd compaction that stalls the API server during a control plane surge, a NetworkPolicy that silently drops traffic between namespaces, a cluster autoscaler that thrashes all night, an admission webhook that fails closed during an outage and blocks every deploy, a spot node pool that gets reclaimed during a traffic spike, a backup that has never been restored and turns out to be incomplete. Each is presented with the failure, the countermeasure, and the operational tradeoff.
Fourteen chapters. Real YAML, kubectl, and operator code. Written for platform engineers and SREs running clusters at scale.