OpenShift training for customer managed-services engineers

We offer private, instructor-led training for managed-services engineers who operate production OpenShift environments for customers.

Your team performs routine platform actions. Customer operations also require evidence, agreed change boundaries, and clear communication, especially when a maintenance task or incident affects running applications.

LearnKube connects workload and cluster mechanisms to a representative customer service, a controlled change, and an incident handoff.

Hands-on learning and the skills engineers take back to work.

Preview: course-wide figures are not yet available.

  • Hands-on learning
    Of instruction time spent on labs and challenges.
  • Troubleshooting confidence
    Of respondents report greater confidence diagnosing Kubernetes problems.
  • Relevant to your work
    Of respondents say the course addressed their engineering responsibilities.
  • Skills put into practice
    Of respondents applied their new skills at work within 90 days.
  • Prepare a controlled change by reviewing workload dependencies and maintenance conditions, so the team can explain its expected effect before acting.
  • Preserve required availability by examining replicas, readiness, and disruption constraints, so maintenance has explicit stop and recovery conditions.
  • Diagnose customer incidents by gathering authorized workload and platform evidence, so restoration or escalation follows a supported hypothesis.
  • Maintain the service record by documenting decisions, results, and remaining owners, so the customer and next operator understand the current state.

Customer platform operations combine technical mechanisms with agreed service boundaries. Disruption budgets constrain eligible evictions, while OpenShift diagnostic collection provides another kind of operational evidence. Neither determines the customer's maintenance decision alone. Engineers must interpret workload health, capacity, permissions, and expected service behavior before changing the environment or requesting further access.

I want an operator to explain why maintenance can proceed, not only which command performs it. A blocked drain can reveal a workload constraint that needs customer agreement, rather than a reason to bypass protection.

— Daniele Polencic, LearnKube founder and Kubernetes instructor

Customer task or requirementKubernetes decisionPractice
Prepare platform maintenanceReplicas and disruption constraintsReview a blocked drain
Restore a customer serviceWorkload and dependency healthInvestigate a service interruption
Gather platform evidenceDiagnostic scope and permissionsSelect an authorized collection
Continue support across shiftsCurrent state and ownershipProduce an incident handoff

AHEAD operates customer cloud and production OpenShift environments. Its work includes approved changes, incident restoration, service reviews, backup validation, runbooks, and platform lifecycle activities.

For a team with those responsibilities, we recommend a four-day workshop connecting Kubernetes behavior to maintenance decisions and useful customer incident records.

This proposed agenda uses an agreed OpenShift lab and a representative customer service. It adapts the Kubernetes baseline to the team's maintenance and escalation duties.

Day 1

We examine controllers, Services, probes, and rollout decisions. You will compare platform status with the application behavior the customer expects.

  • Investigate a service whose replicas are running but not ready.
  • Explain the effect of a failed application update and the available recovery action.

Day 2

You will use Helm and compare Kustomize for repeatable configuration. We trace control-plane and node responsibilities through a planned maintenance scenario.

  • Identify which components a selected change affects.
  • Rehearse a blocked drain and define the evidence needed before proceeding.

Day 3

We trace DNS, Services, Routes, network policies, service mesh use cases, and scheduling constraints. The investigation distinguishes application changes from platform-owned controls.

  • Diagnose an unavailable customer route.
  • Inspect why replacement workloads cannot use the available capacity.

Day 4

You will examine storage, secrets, metrics, authentication, and RBAC. We relate technical recovery checks to the evidence that supports the customer service record.

  • Verify a workload's data access after replacement.
  • Prepare an incident handoff with the observed state, authorized actions, and unresolved ownership.

Your customer service boundaries, maintenance duties, and escalation paths can shape the agenda. Get in touch to tailor the workshop to your managed OpenShift operations.

When an engineer proposes bypassing a blocked eviction, the instructor can examine readiness, replicas, and capacity. The group can explain the underlying constraint and the customer decision it requires.

The interactive challenges and the vast amount of material. Challenges challenges challenges. They were like little puzzles we had to solve.

— Nathaniel, Software Developer at Sabel Systems.

Tell us which customer platforms you operate, which changes and incidents your team owns, and how escalation works. We will recommend technical investigations and operating exercises for those duties.