Prometheus Multi-Window Multi-Burn-Rate SLO Alert Authoring Prompt
Author a complete multi-window, multi-burn-rate SLO alerting ruleset (fast + slow burn pairs with for/severity) from an objective and error-budget window, balancing detection speed against false-page rate.
- Target user
- SREs and reliability engineers owning SLOs
- Difficulty
- Advanced
- Tools
- Claude, ChatGPT
The prompt
You are a senior reliability engineer who authors multi-window, multi-burn-rate alert rules following the Google SRE workbook approach. I will provide: - The SLO target (e.g. 99.9% over 30 days) and the SLI as a PromQL good/total ratio - The metric names for good and total events (or a histogram for latency SLOs) - The notification tiers available (page vs ticket) and on-call tolerance for false pages - Any existing recording rules for the SLI Your job: 1. **Compute the budget** — derive the allowed error ratio and translate target + window into burn-rate thresholds for the standard window pairs (e.g. 1h/5m at 14.4x, 6h/30m at 6x, 24h/2h, 3d/6h). 2. **Define the SLI rules** — write recording rules emitting `slo:sli_error:ratio_rate<window>` for each required window so alerts read cheap precomputed series. 3. **Pair the windows** — for each burn rate, write an alert that fires only when both the long and short window exceed the threshold, so transient blips self-clear. 4. **Set severity and for** — assign page vs ticket per burn rate, set a short `for:` to debounce, and add `labels` (severity, slo) plus `annotations` (budget consumed, runbook). 5. **Avoid double-paging** — ensure faster burn rates inhibit or supersede slower ones via Alertmanager inhibition or labeling. 6. **Sanity-check** — show the math for how fast each tier fires at a given sustained error rate. Output as: (a) recording rules YAML, (b) alerting rules YAML with for/severity, (c) the burn-rate math table, (d) an Alertmanager inhibition note.
Run this prompt with AI
Test it, get an AI-improved version, or compare models — live in the Prompt Workspace. No copy-paste.
Related prompts
-
SLO Error Budget & Multi-Window Burn Rate Alerts Prompt
Design SLO-based alerts — error budgets, multi-burn-rate alerting, SLI selection, burn budget calculation.
-
Error Budget Burn-Rate Alert Design Prompt
Design multi-window, multi-burn-rate SLO alerts that page only when the error budget is actually in danger — fast pages for catastrophic burn, tickets for slow leaks — eliminating both flapping and silent budget exhaustion.
-
Grafana SLO Burn-Rate Dashboard Design Prompt
Design a Grafana SLO dashboard that visualizes error-budget remaining, multi-window burn rate, and time-to-exhaustion so stakeholders see reliability health at a glance.
-
Prometheus Alert Severity & Routing Taxonomy Design Prompt
Design a consistent severity label taxonomy and routing-ready label set across Prometheus alerting rules so Alertmanager can route, group, and escalate deterministically.
More Prometheus & Monitoring prompts & error guides
Browse every Prometheus & Monitoring prompt and troubleshooting guide in one place.
Reading prompts? Get all 500 in one free PDF
500 battle-tested, copy-paste AI prompts engineered by a senior systems engineer — every one with fill-in placeholders and safety/back-out notes. Drop your email and it's yours.
- 500 prompts: Linux · Kubernetes · Terraform · OpenStack · GitLab · Docker · Monitoring · Incident Response
- Instant PDF download — yours free, forever
- Plus one practical AI-workflow email a week (no spam)
Single opt-in · unsubscribe anytime · no spam.