Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Challenge each assumption's validity — shared cross-repo SOP
.claude/skills/yogsoth-ai-assumption-challenging/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✓→✗ | ▼ Worse | 40% | 0% |
| case-05 | ✓→✗ | ▼ Worse | 85% | 0% |
| case-03 | ✓→✓ | = Same ✓ | 21% | 0% |
| case-01 | ✗→✗ | = Same ✗ | 20% | 0% |
| case-02 | ✗→✗ | = Same ✗ | 111% | 0% |
Challenge each assumption's validity using systematic questioning. Determine which assumptions are well-founded, which are fragile, and which are likely false.
Subagent — spawned via subagent-spawning/spawn-agent skill.
Shared: This SOP is used across multiple strategies (assumption-constraint, conflict-resolution, constraint-breaking) and campaigns.
<!-- BEGIN available-tables (generated) -->
Optional, no fixed order; the final leaf is always a sop.
| SOP | When to use | | --- | --- | | spawn-agent | Spawn a customized CC subagent with full MCP tool access. Used by SOPs that declare execution: subagent. |
<!-- END available-tables (generated) -->
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 18,285 | 21,099 | +15% | 1 | 1 | 0% | 2,908 | 3,496 | +20% | 0 | 0 | — |
case-02 | fail→fail | 8,245 | 14,693 | +78% | 1 | 1 | 0% | 1,239 | 2,620 | +111% | 0 | 0 | — |
case-03 | pass→pass | 12,465 | 14,757 | +18% | 1 | 1 | 0% | 2,175 | 2,629 | +21% | 0 | 0 | — |
case-04 | pass→fail | 9,885 | 13,015 | +32% | 1 | 1 | 0% | 1,680 | 2,357 | +40% | 0 | 0 | — |
case-10 | fail→fail | 14,871 | 18,733 | +26% | 1 | 1 | 0% | 2,360 | 3,274 | +39% | 0 | 0 | — |
case-05 | pass→fail | 14,557 | 24,099 | +66% | 1 | 1 | 0% | 2,195 | 4,055 | +85% | 0 | 0 | — |
case-06 | fail→fail | 15,348 | 14,279 | -7% | 1 | 1 | 0% | 2,439 | 2,459 | +1% | 0 | 0 | — |
case-07 | fail→fail | 12,695 | 14,991 | +18% | 1 | 1 | 0% | 1,949 | 2,388 | +23% | 0 | 0 | — |
case-08 | fail→fail | 16,727 | 17,770 | +6% | 1 | 1 | 0% | 2,613 | 3,080 | +18% | 0 | 0 | — |
case-09 | fail→fail | 16,412 | 13,604 | -17% | 1 | 1 | 0% | 2,663 | 2,311 | -13% | 0 | 0 | — |
case-11 | fail→fail | 15,983 | 16,494 | +3% | 1 | 1 | 0% | 2,376 | 2,767 | +16% | 0 | 0 | — |
case-12 | fail→fail | 16,378 | 11,007 | -33% | 1 | 1 | 0% | 2,562 | 1,776 | -31% | 0 | 0 | — |
case-13 | fail→fail | 18,411 | 14,912 | -19% | 1 | 1 | 0% | 2,837 | 2,461 | -13% | 0 | 0 | — |
case-14 | fail→fail | 15,555 | 17,785 | +14% | 1 | 1 | 0% | 2,209 | 2,813 | +27% | 0 | 0 | — |
case-15 | fail→fail | 14,366 | 18,507 | +29% | 1 | 1 | 0% | 2,290 | 3,171 | +38% | 0 | 0 | — |
case-16 | fail→fail | 15,779 | 13,727 | -13% | 1 | 1 | 0% | 2,439 | 2,311 | -5% | 0 | 0 | — |
case-17 | fail→fail | 22,192 | 14,537 | -34% | 1 | 1 | 0% | 3,119 | 2,423 | -22% | 0 | 0 | — |
case-18 | fail→fail | 17,792 | 16,802 | -6% | 1 | 1 | 0% | 2,802 | 2,717 | -3% | 0 | 0 | — |
case-19 | fail→fail | 15,246 | 13,703 | -10% | 1 | 1 | 0% | 2,340 | 2,221 | -5% | 0 | 0 | — |
case-20 | fail→fail | 16,266 | 18,167 | +12% | 1 | 1 | 0% | 2,524 | 2,916 | +16% | 0 | 0 | — |
case-21 | fail→fail | 17,848 | 16,658 | -7% | 1 | 1 | 0% | 2,754 | 2,831 | +3% | 0 | 0 | — |
case-22 | fail→fail | 13,896 | 20,919 | +51% | 1 | 1 | 0% | 2,052 | 3,328 | +62% | 0 | 0 | — |
case-23 | fail→fail | 17,660 | 18,352 | +4% | 1 | 1 | 0% | 2,579 | 2,846 | +10% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted. The headline lift of -9 percentage points is the difference between those two pass rates over the 23 comparable cases. 2 cases got worse with the skill loaded, and they are included in that figure.
Other measured skills in the registry, with their headline benchmark lift.