Site Reliability Engineer¶
Duration: 16–24 weeks · Difficulty: advanced · Badge: SRE
Reliability, observability, incident response, and production operations at scale.
Complete roadmap¶
Target audience¶
Engineers who want the Site Reliability Engineer skill profile and job outcomes below.
Job roles¶
- Site Reliability Engineer
- Production Engineer
- Reliability Engineer
Expected salary ranges¶
Mid to senior SRE roles — treat as directional guidance only.
Prerequisites¶
- Comfort with a laptop and a terminal
- Complete earlier phases before later ones when marked ready
- Prefer the Getting Started overview if you are new to the academy
Phases¶
Foundations¶
- Linux — ready · 25 tutorials
- Networking — ready · 25 tutorials
- Python for DevOps — ready · 27 tutorials
- Git — ready · 20 tutorials
Platform¶
- Kubernetes — ready · 20 tutorials
- Terraform — ready · 20 tutorials
- GitLab CI/CD — ready · 20 tutorials
Observability¶
- Prometheus — planned
- Grafana — planned
- Loki — planned
- Tempo — planned
- OpenTelemetry — planned
- Site Reliability Engineering — planned
Skills gained¶
- Ordered mastery of the technologies on this path
- Hands-on labs and production-oriented habits
- Interview and certification readiness for mapped exams
Projects¶
See the Projects catalog and Capstones. Map picks to this path’s technologies.
Capstone¶
Choose a capstone that exercises the final phases of this path (for example Status API for DevOps / Kubernetes, or the Python automation framework for AI for DevOps).
Interview roadmap¶
- Finish ready technology tracks on this path
- Use Interview Guides per technology
- Rehearse troubleshooting stories from labs
Certification roadmap¶
CKA, PCA
Related career paths¶
Estimated duration¶
16–24 weeks of focused study (tutorials + labs). Stretch if you are new to Linux or cloud.