Incident Severity Classification Rubric Prompt
Design a clear, defensible SEV classification rubric that on-call engineers can apply in seconds under pressure — with crisp boundaries, escalation triggers, and downgrade rules.
- Target user
- SRE managers and on-call leads standardizing incident severity
- Difficulty
- Intermediate
- Tools
- Claude, ChatGPT
The prompt
You are a senior SRE who has standardized severity definitions across multiple engineering orgs. You know that ambiguous SEV rubrics cause both over-paging and under-reacting, and that the rubric must be usable at 3am by a tired engineer.
I will provide:
- Our services and their criticality tiers
- Customer-facing vs internal surfaces
- SLOs / SLAs in place
- Current pain points (everything becomes a SEV1, or nothing escalates)
Your job:
1. **Define SEV levels (SEV1–SEV4 or our scale)** — for each, give a one-line definition, customer impact, example scenarios, expected response time, who gets paged, and whether comms/status-page updates are required.
2. **Make boundaries unambiguous** — replace fuzzy words ("major", "significant") with measurable thresholds: % users affected, revenue/min at risk, SLO error-budget burn rate, data-loss risk, security exposure. Give a decision flowchart an engineer can run in under 15 seconds.
3. **Escalation triggers** — define explicit conditions that auto-upgrade a SEV (duration exceeded, second service degraded, exec/customer escalation). Define downgrade rules so incidents don't stay inflated.
4. **Special cases** — data integrity, security/breach, partial-region outages, third-party dependency failures, and "is it even an incident?" gating.
5. **Anti-patterns** — severity inflation, severity by seniority of the reporter, and stale severities. Propose guardrails for each.
6. **Validation** — propose a calibration exercise: 8 sample scenarios with the expected SEV and rationale, so teams can self-test consistency.
Output as: (a) the severity table, (b) the sub-15-second decision flowchart, (c) escalation/downgrade rules, (d) the 8-scenario calibration quiz with answer key.
Bias toward: measurable thresholds over judgment calls, fast classification over precision, and consistency across teams.
Run this prompt with AI
Test it, get an AI-improved version, or compare models — live in the Prompt Workspace. No copy-paste.
Related prompts
-
Firing Alert Severity & Escalation Decision Prompt
Given a firing alert and current impact signals, decide an appropriate severity level and whether to escalate or page additional responders, with explicit reasoning against your severity rubric — leaving the final call to a human.
-
Is-This-Real Page Triage Prompt
Help a freshly paged on-call engineer decide in the first two minutes whether an alert is a real incident worth waking people for, a transient blip, or pure noise — before they over- or under-react.
-
Escalation Policy Gap and Single-Point-of-Failure Analysis Prompt
Audit your existing escalation policies and on-call schedules to find coverage gaps, dead-ends, and single points of failure where a page could go unanswered during a real incident.
-
In-Incident Severity Re-Evaluation Prompt
Mid-incident, decide whether to upgrade or downgrade the severity as new facts arrive — so you neither under-respond to a quietly growing outage nor keep executives paged on a resolved blip.
More Incident Response prompts & error guides
Browse every Incident Response prompt and troubleshooting guide in one place.
Reading prompts? Get all 500 in one free PDF
500 battle-tested, copy-paste AI prompts engineered by a senior systems engineer — every one with fill-in placeholders and safety/back-out notes. Drop your email and it's yours.
- 500 prompts: Linux · Kubernetes · Terraform · OpenStack · GitLab · Docker · Monitoring · Incident Response
- Instant PDF download — yours free, forever
- Plus one practical AI-workflow email a week (no spam)
Single opt-in · unsubscribe anytime · no spam.