Operations
Node upgrade strategies
48 / 69

Blue and green node upgrades keep a known-good fleet until the new one is proven.

In place

Rolling replace

n1n2n3
n1'n2n3

The managed node group replaces nodes one by one (maxUnavailable). Simple, no extra cost, but a bad new node reduces the only capacity you have.

Blue and green

New group, then shift

blueblueblue
greengreengreen

Create a node group on the new version, cordon and drain the old one gradually, and delete it when verified. Fast rollback, temporary extra cost.

Drift (Karpenter)

Automatic replacement

AMIdriftnew

When the AMI or NodePool spec changes, Karpenter replaces nodes under disruption budgets. Pin AMI versions; a floating @latest alias rolls nodes unannounced.

PDBs decide the pace

A budget that forbids any disruption blocks drains and stalls every strategy. Review PDBs before each upgrade.

Stateful workloads

EBS-backed Pods are pinned to a zone. Upgrade nodes zone by zone, one node group per zone, and watch volume re-attachment.