Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when a written implementation plan is ready to execute, when the user says "fire", "execute the plan", "build it", or hands over a plan file to implement. The step after recipe.
.claude/skills/escoffier-labs-fire/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✗→✓ | ▲ Improved | 343% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 73% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 78% | 0% |
| case-10 | ✗→✓ | ▲ Improved | -31% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 11% | 0% |
The expeditor calls "fire" and the line cooks the order: exactly as the ticket reads, station by station, nothing leaving the pass unchecked and nothing improvised on the line. This skill executes a written implementation plan task by task. The plan is the ticket; its checkboxes are the board.
fire/<plan>-<date>), or isolate in a worktree: prefer the harness's native worktree tool; fall back to git worktree add only without one, and verify the worktree directory is gitignored first.Read the entire plan critically against the actual code before executing anything. Stale line numbers, helpers the plan names that the code lacks, tasks that contradict each other: find them now, while the tree is clean, not mid-task with it half-changed. Surface concerns and get them resolved before task 1. A plan that references structure that does not exist is a planning gap that goes back to its author, not an invitation to improvise.
Two modes, same discipline. Pick brigade mode when subagents are available and the plan has more than a couple of tasks; solo otherwise.
Per task, either mode:
Execute continuously. No "should I continue?" between tasks and no progress check-ins; the user asked for the plan to be executed. The only stops are: a structural divergence, a blocker you cannot resolve, verification that fails twice on focused attempts, or the last task done.
After the last task:
discard; show what gets deleted first)| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-07 | fail→fail | 6,981 | 3,559 | -49% | 1 | 1 | 0% | 1,185 | 1,829 | +54% | 0 | 0 | — |
case-01 | fail→fail | 4,052 | 2,288 | -44% | 1 | 1 | 0% | 235 | 1,521 | +547% | 0 | 0 | — |
case-02 | fail→fail | 3,204 | 4,642 | +45% | 1 | 1 | 0% | 139 | 1,473 | +960% | 0 | 0 | — |
case-03 | fail→fail | 29,590 | 5,081 | -83% | 1 | 1 | 0% | 6,289 | 1,664 | -74% | 0 | 0 | — |
case-04 | fail→pass | 2,347 | 3,172 | +35% | 1 | 1 | 0% | 394 | 1,746 | +343% | 0 | 0 | — |
case-05 | pass→pass | 12,795 | 3,092 | -76% | 1 | 1 | 0% | 1,818 | 1,762 | -3% | 0 | 0 | — |
case-06 | pass→pass | 6,782 | 3,595 | -47% | 1 | 1 | 0% | 1,059 | 1,791 | +69% | 0 | 0 | — |
case-08 | fail→pass | 5,811 | 2,371 | -59% | 1 | 1 | 0% | 976 | 1,688 | +73% | 0 | 0 | — |
case-09 | fail→pass | 6,125 | 4,820 | -21% | 1 | 1 | 0% | 1,024 | 1,826 | +78% | 0 | 0 | — |
case-10 | fail→pass | 15,591 | 3,497 | -78% | 1 | 1 | 0% | 2,706 | 1,856 | -31% | 0 | 0 | — |
case-11 | pass→fail | 5,793 | 2,726 | -53% | 1 | 1 | 0% | 930 | 1,627 | +75% | 0 | 0 | — |
case-12 | fail→pass | 9,234 | 2,576 | -72% | 1 | 1 | 0% | 1,498 | 1,658 | +11% | 0 | 0 | — |
case-13 | fail→fail | 5,442 | 2,562 | -53% | 1 | 1 | 0% | 834 | 1,667 | +100% | 0 | 0 | — |
case-14 | pass→pass | 6,176 | 2,406 | -61% | 1 | 1 | 0% | 886 | 1,664 | +88% | 0 | 0 | — |
case-15 | pass→fail | 2,593 | 1,631 | -37% | 1 | 1 | 0% | 407 | 1,507 | +270% | 0 | 0 | — |
case-16 | fail→pass | 5,531 | 1,846 | -67% | 1 | 1 | 0% | 993 | 1,590 | +60% | 0 | 0 | — |
case-17 | fail→pass | 3,994 | 2,535 | -37% | 1 | 1 | 0% | 630 | 1,661 | +164% | 0 | 0 | — |
case-18 | pass→pass | 8,317 | 2,529 | -70% | 1 | 1 | 0% | 1,434 | 1,686 | +18% | 0 | 0 | — |
case-19 | pass→pass | 8,391 | 4,279 | -49% | 1 | 1 | 0% | 1,355 | 1,937 | +43% | 0 | 0 | — |
case-20 | fail→pass | 11,121 | 3,176 | -71% | 1 | 1 | 0% | 1,790 | 1,740 | -3% | 0 | 0 | — |
case-21 | pass→fail | 17,312 | 2,818 | -84% | 1 | 1 | 0% | 3,303 | 1,677 | -49% | 0 | 0 | — |
case-22 | pass→pass | 12,115 | 7,230 | -40% | 1 | 1 | 0% | 2,394 | 2,581 | +8% | 0 | 0 | — |
case-23 | fail→fail | 5,151 | 4,863 | -6% | 1 | 1 | 0% | 207 | 1,441 | +596% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted, and 19 counted toward the lift figure. The other 4 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +22 percentage points is the difference between those two pass rates over the 19 comparable cases. 3 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.