Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Structured clarification workflow for underspecified requirements. Use before planning to resolve ambiguities through coverage-based questioning. Records answers in spec clarifications section.
.claude/skills/foryourhealth111-pixel-speckit-clarify/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✗→✓ | ▲ Improved | 85% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 77% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 70% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 113% | 0% |
| case-15 | ✗→✓ | ▲ Improved | 226% | 0% |
text$ARGUMENTS
You MUST consider the user input before proceeding (if not empty).
Goal: Detect and reduce ambiguity or missing decision points in the active feature specification and record the clarifications directly in the spec file.
Note: This clarification workflow is expected to run (and be completed) BEFORE invoking /speckit.plan. If the user explicitly states they are skipping clarification (e.g., exploratory spike), you may proceed, but must warn that downstream rework risk increases.
Execution steps:
.specify/scripts/powershell/check-prerequisites.ps1 -Json -PathsOnly from repo root once (combined --json --paths-only mode / -Json -PathsOnly). Parse minimal JSON payload fields:FEATURE_DIRFEATURE_SPECIMPL_PLAN, TASKS for future chained flows.)/speckit.specify or verify feature branch environment.Functional Scope & Behavior:
Domain & Data Model:
Interaction & UX Flow:
Non-Functional Quality Attributes:
Integration & External Dependencies:
Edge Cases & Failure Handling:
Constraints & Tradeoffs:
Terminology & Consistency:
Completion Signals:
Misc / Placeholders:
For each category with Partial or Missing status, add a candidate question opportunity unless:
**Recommended:** Option [X] - <reasoning>| Option | Description | |--------|-------------| | A | <Option A description> | | B | <Option B description> | | C | <Option C description> (add D/E as needed up to 5) | | Short | Provide a different short answer (<=5 words) (Include only if free-form alternative is appropriate) |
You can reply with the option letter (e.g., "A"), accept the recommendation by saying "yes" or "recommended", or provide your own short answer.**Suggested:** <your proposed answer> - <brief reasoning>Format: Short answer (<=5 words). You can accept the suggestion by saying "yes" or "suggested", or provide your own answer.## Clarifications section exists (create it just after the highest-level contextual/overview section per the spec template if missing).### Session YYYY-MM-DD subheading for today.- Q: <question> → A: <final answer>.(formerly referred to as "X") once.## Clarifications, ### Session YYYY-MM-DD.FEATURE_SPEC./speckit.plan or run /speckit.clarify again later post-plan.Behavior rules:
/speckit.specify first (do not create a new spec here).Context for prioritization: $ARGUMENTS
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 3,031 | 3,917 | +29% | 1 | 1 | 0% | 505 | 2,550 | +405% | 0 | 0 | — |
case-02 | fail→fail | 7,431 | 4,915 | -34% | 1 | 1 | 0% | 1,220 | 2,599 | +113% | 0 | 0 | — |
case-03 | fail→fail | 4,573 | 4,353 | -5% | 1 | 1 | 0% | 727 | 2,563 | +253% | 0 | 0 | — |
case-04 | fail→pass | 10,967 | 8,539 | -22% | 1 | 1 | 0% | 1,763 | 3,267 | +85% | 0 | 0 | — |
case-05 | fail→fail | 8,055 | 5,412 | -33% | 1 | 1 | 0% | 1,323 | 2,621 | +98% | 0 | 0 | — |
case-06 | fail→fail | 3,818 | 12,013 | +215% | 1 | 1 | 0% | 524 | 3,720 | +610% | 0 | 0 | — |
case-07 | fail→pass | 8,719 | 2,151 | -75% | 1 | 1 | 0% | 1,510 | 2,669 | +77% | 0 | 0 | — |
case-12 | fail→fail | 13,859 | 6,673 | -52% | 1 | 1 | 0% | 2,044 | 2,680 | +31% | 0 | 0 | — |
case-08 | pass→fail | 8,536 | 1,689 | -80% | 1 | 1 | 0% | 1,398 | 2,586 | +85% | 0 | 0 | — |
case-09 | pass→fail | 10,244 | 2,774 | -73% | 1 | 1 | 0% | 1,525 | 2,800 | +84% | 0 | 0 | — |
case-10 | fail→pass | 9,347 | 3,278 | -65% | 1 | 1 | 0% | 1,703 | 2,895 | +70% | 0 | 0 | — |
case-11 | fail→pass | 8,759 | 4,364 | -50% | 1 | 1 | 0% | 1,458 | 3,107 | +113% | 0 | 0 | — |
case-13 | fail→fail | 7,959 | 3,389 | -57% | 1 | 1 | 0% | 1,326 | 2,854 | +115% | 0 | 0 | — |
case-14 | fail→fail | 11,132 | 5,287 | -53% | 1 | 1 | 0% | 1,948 | 2,721 | +40% | 0 | 0 | — |
case-15 | fail→pass | 5,237 | 2,117 | -60% | 1 | 1 | 0% | 827 | 2,697 | +226% | 0 | 0 | — |
case-16 | pass→pass | 11,162 | 7,288 | -35% | 1 | 1 | 0% | 1,790 | 3,007 | +68% | 0 | 0 | — |
case-17 | fail→pass | 8,786 | 2,597 | -70% | 1 | 1 | 0% | 1,428 | 2,866 | +101% | 0 | 0 | — |
case-18 | pass→fail | 8,603 | 6,145 | -29% | 1 | 1 | 0% | 1,423 | 2,811 | +98% | 0 | 0 | — |
case-19 | pass→pass | 10,159 | 4,648 | -54% | 1 | 1 | 0% | 1,712 | 3,258 | +90% | 0 | 0 | — |
case-20 | fail→fail | 10,393 | 8,579 | -17% | 1 | 1 | 0% | 1,725 | 3,227 | +87% | 0 | 0 | — |
case-21 | pass→pass | 10,758 | 13,027 | +21% | 1 | 1 | 0% | 1,764 | 3,934 | +123% | 0 | 0 | — |
case-22 | pass→fail | 12,329 | 6,069 | -51% | 1 | 1 | 0% | 1,995 | 2,643 | +32% | 0 | 0 | — |
case-23 | fail→fail | 22,331 | 5,720 | -74% | 1 | 1 | 0% | 4,607 | 2,629 | -43% | 0 | 0 | — |
case-24 | fail→fail | 17,182 | 5,754 | -67% | 1 | 1 | 0% | 3,594 | 2,640 | -27% | 0 | 0 | — |
case-25 | fail→fail | 3,356 | 7,730 | +130% | 1 | 1 | 0% | 458 | 2,787 | +509% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 25 cases were attempted, and 13 counted toward the lift figure. The other 12 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +8 percentage points is the difference between those two pass rates over the 13 comparable cases. 4 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.