Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when completing tasks, implementing major features, or before merging to verify work meets requirements
.claude/skills/requesting-code-review/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-09 | ✗→✓ | ▲ Improved | -25% | 0% |
| case-04 | ✓→✓ | = Same ✓ | -16% | 0% |
| case-05 | ✓→✓ | = Same ✓ | 0% | 0% |
| case-06 | ✓→✓ | = Same ✓ | -20% | 0% |
| case-07 | ✓→✓ | = Same ✓ | -17% | 0% |
Dispatch a code reviewer subagent to catch issues before they cascade. The reviewer gets precisely crafted context for evaluation — never your session's history.
Core principle: Review early, review often.
Mandatory:
Optional but valuable:
1. Get git SHAs:
bashBASE_SHA=$(git rev-parse HEAD~1) # or origin/main HEAD_SHA=$(git rev-parse HEAD)
2. Dispatch code reviewer subagent:
Dispatch a general-purpose subagent, filling the template at code-reviewer.md
Placeholders:
{DESCRIPTION} - Brief summary of what you built{PLAN_OR_REQUIREMENTS} - What it should do{BASE_SHA} - Starting commit{HEAD_SHA} - Ending commit3. Act on feedback:
[Just completed Task 2: Add verification function]
You: Let me request code review before proceeding.
BASE_SHA=$(git log --oneline | grep "Task 1" | head -1 | awk '{print $1}')
HEAD_SHA=$(git rev-parse HEAD)
[Dispatch code reviewer subagent]
DESCRIPTION: Added verifyIndex() and repairIndex() with 4 issue types
PLAN_OR_REQUIREMENTS: Task 2 from docs/superpowers/plans/deployment-plan.md
BASE_SHA: a7981ec
HEAD_SHA: 3df7661
[Subagent returns]:
Strengths: Clean architecture, real tests
Issues:
Important: Missing progress indicators
Minor: Magic number (100) for reporting interval
Assessment: Ready to proceed
You: [Fix progress indicators]
[Continue to Task 3]| Excuse | Reality | |--------|---------| | "I'll just review the diff myself instead of dispatching a reviewer" | You're the coordinator — reviewing the diff inline burns the context window you need to keep driving the work. Dispatch a reviewer subagent: the diff and the evaluation live in its context, and only the findings come back to you. | | "The reviewer needs my whole session history to understand the change" | Hand it precisely crafted context, never your session's history. That keeps the reviewer on the work product, not your thought process. |
Never:
If reviewer wrong:
See template at: code-reviewer.md
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 6,995 | 12,059 | +72% | 1 | 1 | 0% | 1,086 | 922 | -15% | 0 | 0 | — |
case-02 | fail→fail | 11,097 | 3,183 | -71% | 1 | 1 | 0% | 1,727 | 896 | -48% | 0 | 0 | — |
case-03 | fail→fail | 4,436 | 5,344 | +20% | 1 | 1 | 0% | 654 | 956 | +46% | 0 | 0 | — |
case-04 | pass→pass | 10,033 | 3,605 | -64% | 1 | 1 | 0% | 1,581 | 1,324 | -16% | 0 | 0 | — |
case-05 | pass→pass | 8,902 | 4,675 | -47% | 1 | 1 | 0% | 1,395 | 1,392 | -0% | 0 | 0 | — |
case-06 | pass→pass | 9,935 | 3,333 | -66% | 1 | 1 | 0% | 1,567 | 1,255 | -20% | 0 | 0 | — |
case-07 | pass→pass | 9,351 | 3,408 | -64% | 1 | 1 | 0% | 1,469 | 1,215 | -17% | 0 | 0 | — |
case-08 | pass→pass | 12,706 | 4,905 | -61% | 1 | 1 | 0% | 2,026 | 1,452 | -28% | 0 | 0 | — |
case-09 | fail→pass | 9,224 | 3,264 | -65% | 1 | 1 | 0% | 1,647 | 1,230 | -25% | 0 | 0 | — |
case-10 | pass→pass | 4,534 | 2,516 | -45% | 1 | 1 | 0% | 808 | 1,152 | +43% | 0 | 0 | — |
case-11 | pass→pass | 4,381 | 1,871 | -57% | 1 | 1 | 0% | 673 | 998 | +48% | 0 | 0 | — |
case-12 | pass→pass | 7,756 | 2,909 | -62% | 1 | 1 | 0% | 1,329 | 1,210 | -9% | 0 | 0 | — |
case-13 | pass→pass | 8,119 | 2,847 | -65% | 1 | 1 | 0% | 1,311 | 1,149 | -12% | 0 | 0 | — |
case-14 | pass→pass | 10,784 | 7,712 | -28% | 1 | 1 | 0% | 1,827 | 1,880 | +3% | 0 | 0 | — |
case-15 | pass→pass | 8,310 | 3,469 | -58% | 1 | 1 | 0% | 1,295 | 1,267 | -2% | 0 | 0 | — |
case-16 | pass→pass | 10,351 | 5,832 | -44% | 1 | 1 | 0% | 1,641 | 1,343 | -18% | 0 | 0 | — |
case-17 | pass→pass | 8,883 | 4,264 | -52% | 1 | 1 | 0% | 1,483 | 1,392 | -6% | 0 | 0 | — |
case-18 | pass→pass | 9,605 | 4,604 | -52% | 1 | 1 | 0% | 1,485 | 1,490 | +0% | 0 | 0 | — |
case-19 | pass→pass | 12,233 | 5,080 | -58% | 1 | 1 | 0% | 1,897 | 1,458 | -23% | 0 | 0 | — |
case-20 | pass→pass | 3,562 | 2,678 | -25% | 1 | 1 | 0% | 679 | 1,214 | +79% | 0 | 0 | — |
case-21 | pass→pass | 5,777 | 3,697 | -36% | 1 | 1 | 0% | 1,001 | 1,348 | +35% | 0 | 0 | — |
case-22 | fail→fail | 3,670 | 4,427 | +21% | 1 | 1 | 0% | 639 | 1,444 | +126% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 20 counted toward the lift figure. The other 2 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +5 percentage points is the difference between those two pass rates over the 20 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 7/24/2026 | +32% |
Other measured skills in the registry, with their headline benchmark lift.