Skip to main content

8 docs tagged with "sre"

View all tags

AWS: The Complete Guide

End-to-end reference for AWS — compute, storage, networking, IAM, databases, high availability, and interview-ready Q&A.

Kubernetes: The Complete Guide

End-to-end reference for Kubernetes — architecture, core objects, networking, scheduling, health checks, rollouts, and interview-ready Q&A.

Observability (Grafana & Prometheus): The Complete Guide

End-to-end reference for Observability (Grafana & Prometheus) — Prometheus architecture and PromQL, SLIs/SLOs and error budgets, alerting and dashboards, long-term storage, ELK/EFK and commercial APM tradeoffs, and interview-ready Q&A.

OpenTelemetry: The Complete Guide

End-to-end reference for OpenTelemetry — traces, metrics, logs, SDK/Collector architecture, instrumentation, sampling, and interview-ready Q&A.

System Performance: The Complete Guide

End-to-end reference for diagnosing Linux system performance with the USE and RED methods — CPU, memory, disk I/O, network, the standard toolkit, and interview-ready Q&A.

Terraform: The Complete Guide

End-to-end reference for Terraform — IaC philosophy, core workflow, HCL syntax, state management, modules, and interview-ready Q&A.