Blue and green node upgrades keep a known-good fleet until the new one is proven.
In place
Rolling replace
n1n2n3
n1'n2n3
The managed node group replaces nodes one by one (maxUnavailable). Simple, no extra cost, but a bad new node reduces the only capacity you have.
Blue and green
New group, then shift
blueblueblue
greengreengreen
Create a node group on the new version, cordon and drain the old one gradually, and delete it when verified. Fast rollback, temporary extra cost.
Drift (Karpenter)
Automatic replacement
AMIdriftnew
When the AMI or NodePool spec changes, Karpenter replaces nodes under disruption budgets. Pin AMI versions; a floating @latest alias rolls nodes unannounced.
PDBs decide the pace
A budget that forbids any disruption blocks drains and stalls every strategy. Review PDBs before each upgrade.
Stateful workloads
EBS-backed Pods are pinned to a zone. Upgrade nodes zone by zone, one node group per zone, and watch volume re-attachment.