← All Topics
Monitoring
You cannot fix what you cannot see. Learn to set up Prometheus and Grafana, write meaningful alerts, build dashboards, and answer the observability questions that come up in every SRE interview.
What this covers
- Prometheus metrics and PromQL
- Grafana dashboards and panels
- Alertmanager and alert routing
- Log aggregation with Loki or ELK
- Distributed tracing basics
- SLOs, SLAs, and error budgets
Monitoring is coming soon
We are building the notes PDF, scenario interview questions, and real failure stories for Monitoring. Follow @rootcausedaily on Instagram to get notified when it launches.
What we are building for Monitoring
Prometheus metrics and PromQL
Grafana dashboards and panels
Alertmanager and alert routing
Log aggregation with Loki or ELK
Distributed tracing basics
SLOs, SLAs, and error budgets