Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Autopsy a business idea before you build it: kill-list check, five hard filters, a free-AI one-prompt test, live ad-market verification, and a verdict with a named kill-pattern.
.claude/skills/sickn33-idea-autopsy/SKILL.md| Model | Eval pass | Runs |
|---|---|---|
| gemini-3.6-flashlowest | 88% | 18 |
| gemini-3.1-pro-preview | 100% | 3 |
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 51% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 79% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 135% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 64% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 33% | 0% |
Turns the agent into a ruthless business-idea pathologist: instead of encouraging the user, it hunts for the one sentence that kills an idea — before any money or weeks are spent building it. Built from a real founder kill-list of 42 dead ideas (including a 9/10-scored idea and one that turned out to be federally illegal to charge for). Every autopsy ends in a hard verdict: DEAD with a named kill-pattern, or SURVIVED with the one cheapest test that could still kill it.
If the project contains a REJECTION.md (the user's personal kill-list), read it first. A NICHE match (same niche as a killed row) = verdict DEAD, cite the row, stop. A KILL-PATTERN match alone (new niche, previously-seen pattern) is a strong prior, NOT a verdict: name the matching pattern, then run the specific check for that pattern (the relevant filter or test below) to confirm it actually applies before declaring death. If no kill-list exists, ask the user for permission to create one with exactly this schema — this autopsy writes its first row:
markdown# REJECTION.md — my kill-list ## Killed ideas | # | Idea/Niche | Killed (date) | Hard reason (one line) | Pattern | |---|-----------|---------------|------------------------|---------| ## Survivors under test | Idea | Passed filters (date) | Pending test | Deadline | |------|----------------------|--------------|----------|
Demand evidence, not optimism. One hard NO = dead.
Try to produce the idea's core deliverable with one prompt to a frontier model. If one prompt produces the whole deliverable free, the verdict is DEAD with kill-pattern free-AI: the user doesn't have a product, they have a prompt.
Desk research lies. Walk the user through the Meta Ad Library (or equivalent) with an explicit checklist: number of ACTIVE advertisers, age of the oldest running ad (90+ days = someone is paying because it works), and the three traps — zero ads (wrong-channel), a few giants (incumbent-owned), or hundreds of ads (crowded commodity knife-fight: demand ≠ room for you).
VERDICT: DEAD | SURVIVED
KILL-PATTERN: <name> (if dead)
THE ONE SENTENCE: <the single finding that decided it>
EVIDENCE: <2-4 hard facts with sources/numbers>
NEXT: <if survived: the ONE cheapest test that could still kill it>Then record the result — gated on consent: if REJECTION.md exists, or the user approved creating it in Step 1, append a one-line row (dead) or note the survivor with the date and pending test. If the user declined the kill-list, do NOT create or write the file — print the proposed row as text so they can save it wherever they prefer. The kill-list is the compounding asset — every autopsy makes the next one faster, but only with the user's consent.
User: Should I build a chargeback-response kit for Shopify sellers? Ready-to-send
evidence packets, $29 one-time.
Agent: Running the autopsy... Step 3 kills it: one prompt to a frontier model
produces a complete, usable evidence packet for free.
VERDICT: DEAD. KILL-PATTERN: free-AI.
THE ONE SENTENCE: If AI ships your whole deliverable in one prompt, you
don't have a product — you have a prompt.User: An app that identifies plants from a photo — there are over a thousand
active ads for these, demand is proven!
Agent: Step 4 confirms ~1,100 active advertisers — and that's the trap. The AI is
free, the datasets are public, everyone could build it, so everyone did.
VERDICT: DEAD. KILL-PATTERN: no-moat.
THE ONE SENTENCE: Demand tells you a market exists; it doesn't tell you
there's room for you.REJECTION.md (hence risk: critical). It never edits other files; ask permission before creating the file on first run.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 17,457 | 15,091 | -14% | 1 | 1 | 0% | 2,485 | 3,755 | +51% | 0 | 0 | — |
case-02 | fail→pass | 12,071 | 11,524 | -5% | 1 | 1 | 0% | 1,810 | 3,242 | +79% | 0 | 0 | — |
case-03 | fail→pass | 13,431 | 22,860 | +70% | 1 | 1 | 0% | 2,118 | 4,975 | +135% | 0 | 0 | — |
case-04 | pass→pass | 12,379 | 5,769 | -53% | 1 | 1 | 0% | 1,673 | 2,445 | +46% | 0 | 0 | — |
case-05 | pass→pass | 14,246 | 9,358 | -34% | 1 | 1 | 0% | 2,065 | 2,960 | +43% | 0 | 0 | — |
case-06 | fail→pass | 13,503 | 10,143 | -25% | 1 | 1 | 0% | 1,829 | 2,994 | +64% | 0 | 0 | — |
case-07 | fail→pass | 17,365 | 10,247 | -41% | 1 | 1 | 0% | 2,344 | 3,123 | +33% | 0 | 0 | — |
case-08 | fail→pass | 16,443 | 17,556 | +7% | 1 | 1 | 0% | 2,247 | 3,952 | +76% | 0 | 0 | — |
case-09 | fail→pass | 17,361 | 11,488 | -34% | 1 | 1 | 0% | 2,454 | 3,048 | +24% | 0 | 0 | — |
case-10 | pass→pass | 14,667 | 9,163 | -38% | 1 | 1 | 0% | 2,039 | 2,992 | +47% | 0 | 0 | — |
case-11 | pass→pass | 14,869 | 10,842 | -27% | 1 | 1 | 0% | 2,145 | 2,944 | +37% | 0 | 0 | — |
case-12 | fail→pass | 14,496 | 8,464 | -42% | 1 | 1 | 0% | 2,084 | 2,796 | +34% | 0 | 0 | — |
case-13 | fail→pass | 7,671 | 5,012 | -35% | 1 | 1 | 0% | 1,168 | 2,316 | +98% | 0 | 0 | — |
case-14 | pass→pass | 14,956 | 9,416 | -37% | 1 | 1 | 0% | 2,303 | 2,904 | +26% | 0 | 0 | — |
case-15 | pass→pass | 20,049 | 19,283 | -4% | 1 | 1 | 0% | 2,957 | 4,468 | +51% | 0 | 0 | — |
case-16 | pass→pass | 10,208 | 3,511 | -66% | 1 | 1 | 0% | 1,619 | 2,027 | +25% | 0 | 0 | — |
case-17 | pass→pass | 16,428 | 11,624 | -29% | 1 | 1 | 0% | 2,442 | 3,231 | +32% | 0 | 0 | — |
case-18 | pass→pass | 12,983 | 5,402 | -58% | 1 | 1 | 0% | 1,811 | 2,293 | +27% | 0 | 0 | — |
case-19 | pass→pass | 8,622 | 5,425 | -37% | 1 | 1 | 0% | 1,353 | 2,256 | +67% | 0 | 0 | — |
case-20 | pass→pass | 14,360 | 9,189 | -36% | 1 | 1 | 0% | 2,166 | 2,814 | +30% | 0 | 0 | — |
case-21 | pass→pass | 11,442 | 9,481 | -17% | 1 | 1 | 0% | 1,738 | 3,024 | +74% | 0 | 0 | — |
case-22 | pass→pass | 15,077 | 9,644 | -36% | 1 | 1 | 0% | 2,868 | 3,355 | +17% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +41 percentage points is the difference between those two pass rates over the 22 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.