Install any skill in seconds. Free to start, no credit card required.
Get Started Free →/cs:post-mortem <decision> — Honest retrospective on an executed decision, scored against original assumptions and dissent. Closes the strategic sprint loop. Use when a decision hits its 90-day review checkpoint or its kill criteria trigger — e.g. scoring last quarter's pricing change against its pre-committed success metrics.
.claude/skills/alirezarezvani-post-mortem/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-03 | ✗→✓ | ▲ Improved | 0% | 0% |
| case-01 | ✗→✓ | ▲ Improved | -23% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -36% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 10% | 0% |
| case-08 | ✗→✓ | ▲ Improved | -3% | 0% |
Command: /cs:post-mortem <decision-path>
Closes the strategic sprint loop. Scores a decision against the success and kill criteria written before the decision (not retro-fitted) and revisits the preserved dissent. This is the rigor that compounds over time.
/cs:office-hours → /cs:brief → /cs:boardroom → /cs:decide → /cs:execute → /cs:post-mortem
↑ you are here/cs:decide)/cs:decide)/cs:execute)Saved to ~/.claude/postmortems/YYYY-MM-DD-<slug>.md:
markdown# Post-Mortem: <decision title> **Decision date:** YYYY-MM-DD **Post-mortem date:** YYYY-MM-DD **Status:** WIN / PARTIAL / LOSS / MIXED ## Outcome Scoring (against pre-committed criteria) | Success Criterion | Threshold | Actual | Met? | |---|---|---|---| | <metric 1> | <threshold> | <actual> | ✅ / ❌ | | <metric 2> | <threshold> | <actual> | ✅ / ❌ | | Kill Criterion | Threshold | Actual | Triggered? | |---|---|---|---| | <metric> | <threshold> | <actual> | ✅ / ❌ | **Overall:** WIN / PARTIAL / LOSS / MIXED ## What We Got Right - <factor 1> - <factor 2> ## What We Got Wrong - <factor 1> - <factor 2> ## Preserved Dissent — Revisited [Original dissent from the boardroom memo, scored:] - **<dissenter>:** <original concern> - **Did it materialize?** YES / NO / PARTIAL - **Cost if YES:** <quantified impact> - **Lesson:** <one sentence> ## Assumption Audit [Original brief's assumptions, scored:] - **Assumption 1:** <text> - **Held?** YES / NO / PARTIAL - **Why:** <explanation> ## Process Lessons - **Phase 2 isolation worked?** YES / NO - **Devil's advocate concerns played out?** YES / NO / PARTIAL - **Cadence was right?** YES / TOO LOOSE / TOO TIGHT ## Forward Actions - [ ] <change to operating system or routing logic> - [ ] <new decision to make based on this learning> - [ ] <update company-context.md> ## Status - WIN → archive, log lesson - LOSS → schedule follow-up boardroom: `/cs:brief` for the next call
The biggest temptation in post-mortems is retroactive justification: "we always knew X, that's why we did Y." Pre-committed criteria, signed at /cs:decide time, eliminate that move. The numbers either matched or they didn't.
The dissent column from /cs:boardroom is the single most useful piece of organizational memory. Most of the time, the dissenter was directionally right. Revisiting and scoring it builds calibration over years.
/cs:brief — if the post-mortem surfaces a new decision/cs:freeze — if the post-mortem reveals a process gap that needs cooldown enforcementcs-onboarddecision-loggercs-chief-of-staff/em:postmortem — adversarial single-decision post-mortemVersion: 1.0.0
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-17 | pass→pass | 13,399 | 12,473 | -7% | 1 | 1 | 0% | 2,171 | 3,071 | +41% | 0 | 0 | — |
case-04 | pass→pass | 23,540 | 25,107 | +7% | 1 | 1 | 0% | 3,941 | 4,319 | +10% | 0 | 0 | — |
case-03 | fail→pass | 30,267 | 22,574 | -25% | 1 | 1 | 0% | 4,054 | 4,055 | +0% | 0 | 0 | — |
case-05 | pass→pass | 22,948 | 27,230 | +19% | 1 | 1 | 0% | 3,527 | 4,477 | +27% | 0 | 0 | — |
case-01 | fail→pass | 36,653 | 22,293 | -39% | 1 | 1 | 0% | 5,399 | 4,156 | -23% | 0 | 0 | — |
case-02 | fail→pass | 31,466 | 23,303 | -26% | 1 | 1 | 0% | 5,253 | 3,384 | -36% | 0 | 0 | — |
case-06 | pass→pass | 25,136 | 21,419 | -15% | 1 | 1 | 0% | 3,348 | 3,824 | +14% | 0 | 0 | — |
case-07 | fail→pass | 14,948 | 10,482 | -30% | 1 | 1 | 0% | 1,865 | 2,060 | +10% | 0 | 0 | — |
case-08 | fail→pass | 16,538 | 5,075 | -69% | 1 | 1 | 0% | 1,972 | 1,905 | -3% | 0 | 0 | — |
case-09 | fail→pass | 18,844 | 8,403 | -55% | 1 | 1 | 0% | 2,249 | 2,350 | +4% | 0 | 0 | — |
case-10 | fail→pass | 11,541 | 5,601 | -51% | 1 | 1 | 0% | 1,479 | 1,955 | +32% | 0 | 0 | — |
case-16 | fail→pass | 10,406 | 12,506 | +20% | 1 | 1 | 0% | 1,903 | 2,027 | +7% | 0 | 0 | — |
case-11 | fail→fail | 8,243 | 2,855 | -65% | 1 | 1 | 0% | 1,310 | 1,470 | +12% | 0 | 0 | — |
case-12 | fail→pass | 15,756 | 6,791 | -57% | 1 | 1 | 0% | 2,060 | 2,229 | +8% | 0 | 0 | — |
case-13 | fail→pass | 12,779 | 7,872 | -38% | 1 | 1 | 0% | 2,019 | 1,412 | -30% | 0 | 0 | — |
case-14 | fail→pass | 7,964 | 7,814 | -2% | 1 | 1 | 0% | 1,335 | 1,484 | +11% | 0 | 0 | — |
case-15 | fail→fail | 13,385 | 6,958 | -48% | 1 | 1 | 0% | 1,851 | 2,212 | +20% | 0 | 0 | — |
case-18 | fail→pass | 11,078 | 9,582 | -14% | 1 | 1 | 0% | 1,800 | 1,529 | -15% | 0 | 0 | — |
case-19 | fail→pass | 13,254 | 3,178 | -76% | 1 | 1 | 0% | 1,921 | 1,549 | -19% | 0 | 0 | — |
case-20 | pass→pass | 16,462 | 16,103 | -2% | 1 | 1 | 0% | 1,867 | 2,911 | +56% | 0 | 0 | — |
case-21 | pass→pass | 18,014 | 7,347 | -59% | 1 | 1 | 0% | 1,965 | 2,146 | +9% | 0 | 0 | — |
case-22 | fail→pass | 12,761 | 2,775 | -78% | 1 | 1 | 0% | 1,992 | 1,439 | -28% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +64 percentage points is the difference between those two pass rates over the 22 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 8/10/2026 | +45% |
Other measured skills in the registry, with their headline benchmark lift.