Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Production incident response automation. Reads logs, checks recent deploys, identifies root cause, suggests fixes, drafts incident comms, creates post-mortem templates. Severity classification (SEV1-4), escalation paths, status page updates. Generates incident-report.md with timeline, root cause, impact assessment, remediation steps, and prevention measures.
.claude/skills/onewave-ai-incident-responder/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 32% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 30% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 285% | 0% |
| case-08 | ✓→✗ | ▼ Worse | 25% | 0% |
| case-02 | ✓→✓ | = Same ✓ | 60% | 0% |
Act as an expert SRE and production incident responder. Systematically investigate, diagnose, classify, and guide an incident through resolution, then produce actionable reports, audience-specific communications, and a prevention-focused post-mortem.
references/severity-matrix.md -- SEV1-4 classification criteria, response expectations, escalation/de-escalation rules.references/investigation-protocol.md -- log sources, deploy checks, dependency and resource analysis, root cause chain, codebase patterns.references/diagnostic-commands.md -- shell commands for logs, resources, containers, databases, git history.references/communication-templates.md -- status page, internal, executive, and customer-facing templates.references/incident-report-template.md -- full incident-report.md structure.references/escalation-and-status.md -- escalation paths, IC responsibilities, status page cadence and rules.references/checklists.md -- declaration, verification, resolution, and post-mortem checklists.references/severity-matrix.md, taking the highest level matched by any criterion. State the classification, its implications, and the required response cadence.references/investigation-protocol.md: identify log sources, check recent deployments, analyze dependencies and resources, and build an evidence-backed failure chain to a confirmed root cause. Use references/diagnostic-commands.md when shell access is available.references/checklists.md.references/communication-templates.md appropriate to the severity: status page updates for all customer-facing incidents, internal engineering updates, plus executive summary and customer email for SEV1/SEV2. Map impact to component status and follow the cadence in references/escalation-and-status.md.incident-report.md following references/incident-report-template.md. Include the complete timeline with evidence, the root cause chain, and prioritized action items with owners across all prevention categories.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 14,363 | 31,004 | +116% | 1 | 1 | 0% | 2,387 | 3,161 | +32% | 0 | 0 | — |
case-02 | pass→pass | 13,921 | 14,989 | +8% | 1 | 1 | 0% | 1,975 | 3,167 | +60% | 0 | 0 | — |
case-03 | fail→pass | 10,614 | 8,222 | -23% | 1 | 1 | 0% | 1,730 | 2,251 | +30% | 0 | 0 | — |
case-04 | pass→pass | 7,782 | 5,400 | -31% | 1 | 1 | 0% | 1,035 | 1,830 | +77% | 0 | 0 | — |
case-05 | fail→pass | 3,701 | 9,167 | +148% | 1 | 1 | 0% | 572 | 2,204 | +285% | 0 | 0 | — |
case-06 | pass→pass | 13,073 | 13,343 | +2% | 1 | 1 | 0% | 1,937 | 2,741 | +42% | 0 | 0 | — |
case-07 | pass→pass | 10,026 | 6,668 | -33% | 1 | 1 | 0% | 1,452 | 1,948 | +34% | 0 | 0 | — |
case-08 | pass→fail | 14,846 | 11,495 | -23% | 1 | 1 | 0% | 2,106 | 2,628 | +25% | 0 | 0 | — |
case-09 | pass→pass | 10,620 | 9,207 | -13% | 1 | 1 | 0% | 1,510 | 2,261 | +50% | 0 | 0 | — |
case-10 | pass→pass | 23,557 | 62,520 | +165% | 1 | 1 | 0% | 3,726 | 5,190 | +39% | 0 | 0 | — |
case-11 | pass→pass | 13,311 | 10,021 | -25% | 1 | 1 | 0% | 2,189 | 2,512 | +15% | 0 | 0 | — |
case-12 | pass→pass | 11,810 | 11,807 | -0% | 1 | 1 | 0% | 2,385 | 3,220 | +35% | 0 | 0 | — |
case-13 | pass→pass | 5,713 | 6,911 | +21% | 1 | 1 | 0% | 1,102 | 2,132 | +93% | 0 | 0 | — |
case-14 | pass→pass | 7,102 | 5,766 | -19% | 1 | 1 | 0% | 1,229 | 1,898 | +54% | 0 | 0 | — |
case-15 | pass→pass | 7,926 | 4,840 | -39% | 1 | 1 | 0% | 1,701 | 1,950 | +15% | 0 | 0 | — |
case-16 | pass→pass | 6,865 | 5,740 | -16% | 1 | 1 | 0% | 1,292 | 1,900 | +47% | 0 | 0 | — |
case-17 | pass→pass | 13,621 | 13,029 | -4% | 1 | 1 | 0% | 2,567 | 3,574 | +39% | 0 | 0 | — |
case-18 | pass→pass | 4,722 | 8,723 | +85% | 1 | 1 | 0% | 835 | 2,443 | +193% | 0 | 0 | — |
case-19 | pass→pass | 11,960 | 9,255 | -23% | 1 | 1 | 0% | 1,723 | 2,248 | +30% | 0 | 0 | — |
case-20 | pass→pass | 13,231 | 13,616 | +3% | 1 | 1 | 0% | 2,050 | 3,035 | +48% | 0 | 0 | — |
case-21 | pass→pass | 14,973 | 17,519 | +17% | 1 | 1 | 0% | 2,218 | 3,606 | +63% | 0 | 0 | — |
case-22 | pass→pass | 10,959 | 10,547 | -4% | 1 | 1 | 0% | 2,035 | 2,873 | +41% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +9 percentage points is the difference between those two pass rates over the 22 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.