DevOps & Platform Engineering Academy
Infrastructure · CI/CD · Containers · IaC
Master the complete DevOps toolchain — from containers and orchestration to CI/CD pipelines, infrastructure as code, and production monitoring. Built for engineers who want to move fast without breaking things.
👤 Who This Is For
📋 Prerequisites
🎯 What You'll Be Able to Do
🗺️ Recommended Learning Path
📚 Explore Topics
OS & Linux
Operating system fundamentals for DevOps engineers
Master Linux — the foundation of every server, container, and cloud instance
TCP/IP, DNS, VPNs, firewalls, and cloud networking fundamentals
Automate everything with Bash — the universal automation language of DevOps
Containers & Orchestration
Docker, Kubernetes, and the container ecosystem
Containerisation, Docker Compose, multi-stage builds, and best practices
Container orchestration — Pods, Deployments, Services, networking, and production ops
Kubernetes package manager — chart development, templates, and values management
Service mesh — traffic management, mTLS, observability, and canary rollouts with Istio
CI/CD & GitOps
Automated pipelines and GitOps delivery
CI/CD fundamentals — pipeline stages, build/test/deploy automation, tool comparison (Jenkins, GitHub Actions, GitLab CI)
Version control, branching strategies, PRs, and collaborative workflows
CI/CD pipelines, declarative Jenkinsfiles, shared libraries, and Kubernetes agents
GitOps continuous delivery — App of Apps, sync waves, multi-cluster management
Infrastructure as Code
Provision and manage infrastructure with code
Monitoring & Observability
Metrics, logs, traces, and alerting
Metrics collection, PromQL, alerting rules, and recording rules
Dashboards, visualization, alerts, and unified observability
Elasticsearch, Logstash, Kibana — centralized log management and analytics
High availability and disaster recovery — RTO/RPO, multi-AZ/region failover, backup strategy
Incident response process — severity levels, on-call, postmortems, communication during outages
Deliberately injecting failure to test resilience — chaos experiments, blast radius, game days
Forecasting resource needs — load testing, scaling thresholds, cost-aware headroom planning
Automating operational toil away — runbooks as code, auto-remediation, self-healing systems
Distributed event streaming — topics, partitions, consumer groups, exactly-once semantics
Scripting & Automation
Python and scripting for DevOps automation
OpenShift & Enterprise K8s
Red Hat OpenShift — enterprise Kubernetes with SCCs, Routes, Operators, and built-in monitoring

