Cloud & DevOps
Advanced

Platform Engineering and Site Reliability Engineering

Build internal platforms and run reliable systems with Kubernetes, IaC and observability.

A professional course in platform engineering and SRE: service level objectives, error budgets, Kubernetes platforms, infrastructure as code, golden paths, observability, incident response and capacity planning.

130h 2 certificates available
SREPlatform EngineeringKubernetesObservabilityTerraform

Curriculum

Two complete tracks. Study either or both — each has its own exam and certificate.

Eight modules covering reliability engineering practice, Kubernetes platform building, automation, observability and incident management.

What reliability means and how it is measured.

  • Availability, Latency and User-Centred SLIs Lab40 min
  • SLOs, Error Budgets and Policy Lab40 min
  • Toil, Automation and the SRE Model Lab40 min
  • Risk, Redundancy and Failure Domains Lab40 min

Careers this prepares you for

  • Site Reliability Engineer
  • Platform Engineer
  • Infrastructure Engineer
  • DevOps Lead
  • Cloud Reliability Engineer

How you study

Every lesson pairs an in-depth video with written notes, key takeaways, a hands-on lab, exercises and a quiz. Work at your own pace, then sit the final examination for your Erudex certificate.