Skip to content

Site Reliability Engineer

Duration: 16–24 weeks · Difficulty: advanced · Badge: SRE

Reliability, observability, incident response, and production operations at scale.

Complete roadmap

Target audience

Engineers who want the Site Reliability Engineer skill profile and job outcomes below.

Job roles

  • Site Reliability Engineer
  • Production Engineer
  • Reliability Engineer

Expected salary ranges

Mid to senior SRE roles — treat as directional guidance only.

Prerequisites

  • Comfort with a laptop and a terminal
  • Complete earlier phases before later ones when marked ready
  • Prefer the Getting Started overview if you are new to the academy

Phases

Foundations

Platform

Observability

  • Prometheus — planned
  • Grafana — planned
  • Loki — planned
  • Tempo — planned
  • OpenTelemetry — planned
  • Site Reliability Engineering — planned

Skills gained

  • Ordered mastery of the technologies on this path
  • Hands-on labs and production-oriented habits
  • Interview and certification readiness for mapped exams

Projects

See the Projects catalog and Capstones. Map picks to this path’s technologies.

Capstone

Choose a capstone that exercises the final phases of this path (for example Status API for DevOps / Kubernetes, or the Python automation framework for AI for DevOps).

Interview roadmap

  1. Finish ready technology tracks on this path
  2. Use Interview Guides per technology
  3. Rehearse troubleshooting stories from labs

Certification roadmap

CKA, PCA

Estimated duration

16–24 weeks of focused study (tutorials + labs). Stretch if you are new to Linux or cloud.