Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Retroactively audit an executed AI phase's evaluation coverage — scores each eval dimension as COVERED/PARTIAL/MISSING and produces an actionable EVAL-REVIEW.md with remediation plan
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-07 | ✗→✓ | ▲ Improved | -72% | 0% |
| case-04 | ✗→✓ | ▲ Improved | -52% | 0% |
| case-06 | ✗→✓ | ▲ Improved | -70% | 0% |
| case-08 | ✗→✓ | ▲ Improved | -74% | 0% |
| case-09 | ✗→✓ | ▲ Improved | -56% | 0% |
<objective> Conduct a retroactive evaluation coverage audit of a completed AI phase. Checks whether the evaluation strategy from AI-SPEC.md was implemented. Produces EVAL-REVIEW.md with score, verdict, gaps, and remediation plan. </objective>
<execution_context> @${CLAUDE_PLUGIN_ROOT}/workflows/eval-review.md @${CLAUDE_PLUGIN_ROOT}/references/ai-evals.md </execution_context>
<context> Phase: $ARGUMENTS — optional, defaults to last completed phase. </context>
<process> Execute @${CLAUDE_PLUGIN_ROOT}/workflows/eval-review.md end-to-end. Preserve all workflow gates. </process>
Other measured skills in the registry, with their headline benchmark lift.