A proposed four-day workshop on validated maintenance
We use the team's current platform and proposed version change to choose representative examples. The workshop develops an operating explanation rather than assuming every platform offers the same rollback mechanism.
Day 1
Workload health during change
Explain probes, replicas, and release behavior before defining maintenance checks. Distinguish workload availability from a successful infrastructure operation.
Hands-on exercises
- Establish application observations before a controlled change.
- Identify an upgrade-sensitive API or workload dependency.
Day 2
Templates and component compatibility
Connect rendered configuration, control-plane components, add-ons, and node versions. Review a supported sequence and separate an in-place change from a replacement approach.
Hands-on exercises
- Build a compatibility inventory for the sample environment.
- Identify stop conditions and a platform-supported recovery option.
Day 3
Draining, traffic, and placement
Examine eviction, disruption budgets, networking continuity, and placement capacity. Investigate why maintenance can block even when remaining nodes look healthy.
Hands-on exercises
- Rehearse a bounded node drain and inspect blocking evidence.
- Verify application traffic while replacement Pods obtain capacity.
Day 4
Stateful dependencies and operating records
Review storage, credentials, autoscaling, and permissions affected by the change. Compare post-change behavior with the baseline before accepting the maintenance step.
Hands-on exercises
- Validate a stateful or identity-dependent application after replacement.
- Produce a staged maintenance record with continue, stop, and recovery criteria.
Your versions, upgrade questions, and component responsibilities can shape the agenda. Get in touch to tailor the workshop to your team's work.