Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Base template for Stage1 reviewer-comment breakdown. Use when creating or extending conference-specific Stage1 skills (for example ICLR, ICML, NeurIPS, ACL) so shared extraction, splitting, and response-mapping logic stays consistent.
.claude/skills/runtsang-stage1-breakdown-template/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 35% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -13% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 48% | 0% |
| case-04 | ✗→✓ | ▲ Improved | 149% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 64% | 0% |
Use this template as the shared core for all Stage1 conference variants. Conference-specific files must apply this template first, then add only necessary overrides.
6, 4, 3), without scale explanations.Response1 ... ResponseN)title: generated short subtitlesource: weakness or questionsource_id: weaknessN or questionNquoted_issue: exact verbatim text of the mapped atomic issueweaknessN.questionN.weaknessN and questionN independently.Split when any of these applies:
Do not split when sentences elaborate one single concern.
Use this structure, with conference-specific keys provided by the extension file:
markdown# Stage1 <CONFERENCE> Breakdown ## Scores - <score_key_1>: <number> - <score_key_2>: <number> ## Preserved Sections - summary: | <verbatim text> - strength: | <verbatim text> ## Atomic Issues - weakness1: "<verbatim quoted issue>" - weakness2: "<verbatim quoted issue>" - question1: "<verbatim quoted issue>" ## Responses ### Response1 - title: <generated subtitle> - source: weakness - source_id: weakness1 - quoted_issue: "<same verbatim text as weakness1>"
Before finalizing, verify:
Atomic Issues.quoted_issue is verbatim.Conference-specific Stage1 files must define:
# Stage1 <CONFERENCE> Breakdown).Do not rewrite shared logic in conference files unless a true conference-specific exception exists.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 9,552 | 21,456 | +125% | 1 | 1 | 0% | 2,113 | 2,849 | +35% | 0 | 0 | — |
case-02 | fail→pass | 12,613 | 5,052 | -60% | 1 | 1 | 0% | 2,358 | 2,049 | -13% | 0 | 0 | — |
case-03 | fail→pass | 8,263 | 5,793 | -30% | 1 | 1 | 0% | 1,512 | 2,244 | +48% | 0 | 0 | — |
case-04 | fail→pass | 6,213 | 7,580 | +22% | 1 | 1 | 0% | 1,016 | 2,533 | +149% | 0 | 0 | — |
case-09 | pass→pass | 9,463 | 5,401 | -43% | 1 | 1 | 0% | 1,362 | 1,660 | +22% | 0 | 0 | — |
case-05 | fail→pass | 10,675 | 10,792 | +1% | 1 | 1 | 0% | 1,708 | 2,799 | +64% | 0 | 0 | — |
case-06 | fail→pass | 9,256 | 4,492 | -51% | 1 | 1 | 0% | 1,620 | 1,752 | +8% | 0 | 0 | — |
case-07 | fail→pass | 5,319 | 9,984 | +88% | 1 | 1 | 0% | 854 | 2,374 | +178% | 0 | 0 | — |
case-08 | pass→pass | 5,908 | 5,331 | -10% | 1 | 1 | 0% | 1,060 | 1,833 | +73% | 0 | 0 | — |
case-15 | fail→pass | 7,049 | 4,812 | -32% | 1 | 1 | 0% | 1,065 | 1,460 | +37% | 0 | 0 | — |
case-10 | fail→pass | 10,702 | 3,157 | -71% | 1 | 1 | 0% | 1,779 | 1,460 | -18% | 0 | 0 | — |
case-11 | fail→pass | 7,833 | 4,470 | -43% | 1 | 1 | 0% | 1,329 | 1,406 | +6% | 0 | 0 | — |
case-12 | fail→pass | 8,029 | 4,385 | -45% | 1 | 1 | 0% | 1,342 | 1,649 | +23% | 0 | 0 | — |
case-13 | fail→fail | 9,544 | 3,595 | -62% | 1 | 1 | 0% | 1,633 | 1,458 | -11% | 0 | 0 | — |
case-14 | pass→pass | 7,519 | 9,768 | +30% | 1 | 1 | 0% | 1,104 | 2,291 | +108% | 0 | 0 | — |
case-16 | pass→pass | 13,471 | 6,422 | -52% | 1 | 1 | 0% | 2,037 | 1,679 | -18% | 0 | 0 | — |
case-17 | pass→pass | 6,420 | 5,351 | -17% | 1 | 1 | 0% | 923 | 1,575 | +71% | 0 | 0 | — |
case-18 | fail→pass | 5,173 | 5,949 | +15% | 1 | 1 | 0% | 956 | 2,069 | +116% | 0 | 0 | — |
case-19 | pass→pass | 14,997 | 6,546 | -56% | 1 | 1 | 0% | 2,160 | 2,097 | -3% | 0 | 0 | — |
case-20 | pass→pass | 7,795 | 3,623 | -54% | 1 | 1 | 0% | 1,360 | 1,304 | -4% | 0 | 0 | — |
case-21 | pass→fail | 15,905 | 13,647 | -14% | 1 | 1 | 0% | 3,211 | 2,697 | -16% | 0 | 0 | — |
case-22 | pass→pass | 19,703 | 18,802 | -5% | 1 | 1 | 0% | 3,055 | 3,708 | +21% | 0 | 0 | — |
case-23 | fail→fail | 4,647 | 6,139 | +32% | 1 | 1 | 0% | 708 | 1,531 | +116% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted. The headline lift of +48 percentage points is the difference between those two pass rates over the 23 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.