Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Run an NPI phase-gate review for EVT, DVT, or PVT — exit criteria per phase, open-issue triage, yield readout, waiver discipline, and a go/no-go call. Use when asked to run a gate review, decide EVT exit or DVT entry, review build results, assess whether to proceed to the next build, or triage open issues before a phase gate. Produces a gate review document with criteria scoring, waiver register, yield analysis, and a defensible go/conditional-go/no-go recommendation.
.claude/skills/mohitagw15856-evt-dvt-pvt-gate-review/SKILL.md| Model | Eval pass | Runs |
|---|---|---|
| gemini-3.6-flash | 83% | 6 |
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 152% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -30% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 2% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 30% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 57% | 0% |
Phase gates exist because hardware mistakes compound: an issue waved through EVT costs 10× at DVT and 100× in the field. This skill runs the gate the way a strong NPI lead does — score against written exit criteria, triage every open issue as blocker or waiver, read yield with its denominator, and make a recommendation someone can be held to.
Ask for these if not provided; run the review on partial data but mark unverifiable criteria [no data — cannot score], never assumed-pass:
Reference exit criteria (adapt to the program's own if provided):
| Criterion | EVT exit | DVT exit | PVT exit | |---|---|---|---| | Proves | Design works (works-like) | Design is reliable & certifiable (looks-like/works-like) | Factory can build it at rate | | Tooling | Proto/soft tooling OK | Off near-final tooling | Production tooling, production line | | Functional yield | ≥ ~80% with failures understood | ≥ ~90% | ≥ ~95%, stable across line runs | | Reliability | Key risks tested (thermal, drop samples) | Full reliability suite passed (drop, tumble, thermal cycle, HALT as applicable) | ORT started; Cpk ≥ 1.33 on critical dimensions | | Certs | Pre-scan risks identified | EMC/safety pre-scans passed | Cert filings submitted/granted | | Cost | BOM within ~10% of target | BOM within ~5%, cost-downs planned | COGS at target with yield burdened in | | Open issues | No unresolved blockers | No blockers; waivers classed & expiring | Only Class C waivers, all with limit samples |
Issue triage. Every open issue gets exactly one bucket: Blocker (fails a criterion, fix before gate), Waiver requested (pass the gate with the defect, under discipline below), Defer (not a gate criterion — but say why).
Waiver discipline. Class A — safety/regulatory/data-loss: never waivable. Class B — functional/reliability: waivable only with named owner, expiry date (a specific build or date at which it's fixed or the program stops), and containment for affected units. Class C — cosmetic: waivable against an approved limit sample.
[no data]| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-05 | pass→pass | 15,044 | 21,078 | +40% | 1 | 1 | 0% | 2,405 | 4,405 | +83% | 0 | 0 | — |
case-20 | pass→pass | 12,620 | 14,043 | +11% | 1 | 1 | 0% | 2,024 | 3,622 | +79% | 0 | 0 | — |
case-01 | fail→pass | 62,828 | 24,851 | -60% | 1 | 1 | 0% | 2,084 | 5,247 | +152% | 0 | 0 | — |
case-02 | fail→pass | 36,454 | 17,286 | -53% | 1 | 1 | 0% | 6,365 | 4,473 | -30% | 0 | 0 | — |
case-03 | fail→fail | 65,360 | 23,687 | -64% | 1 | 1 | 0% | 6,007 | 5,518 | -8% | 0 | 0 | — |
case-04 | pass→pass | 16,624 | 19,238 | +16% | 1 | 1 | 0% | 2,860 | 4,602 | +61% | 0 | 0 | — |
case-06 | pass→pass | 26,255 | 28,877 | +10% | 1 | 1 | 0% | 4,404 | 5,647 | +28% | 0 | 0 | — |
case-07 | fail→pass | 16,739 | 11,097 | -34% | 1 | 1 | 0% | 2,795 | 2,864 | +2% | 0 | 0 | — |
case-08 | pass→pass | 13,683 | 8,352 | -39% | 1 | 1 | 0% | 2,083 | 2,414 | +16% | 0 | 0 | — |
case-09 | fail→pass | 11,494 | 7,651 | -33% | 1 | 1 | 0% | 1,833 | 2,379 | +30% | 0 | 0 | — |
case-10 | pass→pass | 12,797 | 13,156 | +3% | 1 | 1 | 0% | 2,171 | 3,349 | +54% | 0 | 0 | — |
case-21 | pass→pass | 12,325 | 10,410 | -16% | 1 | 1 | 0% | 1,952 | 2,690 | +38% | 0 | 0 | — |
case-11 | pass→pass | 13,713 | 8,825 | -36% | 1 | 1 | 0% | 2,190 | 2,546 | +16% | 0 | 0 | — |
case-12 | pass→pass | 25,966 | 10,654 | -59% | 1 | 1 | 0% | 1,691 | 2,712 | +60% | 0 | 0 | — |
case-13 | fail→pass | 11,321 | 12,532 | +11% | 1 | 1 | 0% | 1,746 | 2,748 | +57% | 0 | 0 | — |
case-14 | fail→pass | 15,652 | 11,375 | -27% | 1 | 1 | 0% | 2,242 | 2,831 | +26% | 0 | 0 | — |
case-15 | pass→pass | 10,548 | 9,691 | -8% | 1 | 1 | 0% | 1,700 | 2,574 | +51% | 0 | 0 | — |
case-16 | fail→pass | 12,292 | 10,610 | -14% | 1 | 1 | 0% | 2,104 | 2,877 | +37% | 0 | 0 | — |
case-17 | pass→pass | 16,808 | 12,625 | -25% | 1 | 1 | 0% | 2,389 | 2,920 | +22% | 0 | 0 | — |
case-18 | fail→pass | 29,276 | 11,436 | -61% | 1 | 1 | 0% | 1,235 | 2,801 | +127% | 0 | 0 | — |
case-19 | pass→pass | 13,532 | 11,213 | -17% | 1 | 1 | 0% | 2,108 | 2,571 | +22% | 0 | 0 | — |
case-22 | pass→pass | 14,949 | 15,049 | +1% | 1 | 1 | 0% | 2,219 | 3,397 | +53% | 0 | 0 | — |
case-23 | pass→pass | 12,706 | 10,048 | -21% | 1 | 1 | 0% | 1,838 | 2,748 | +50% | 0 | 0 | — |
case-24 | pass→pass | 15,171 | 18,455 | +22% | 1 | 1 | 0% | 2,234 | 3,682 | +65% | 0 | 0 | — |
case-25 | fail→pass | 16,138 | 12,555 | -22% | 1 | 1 | 0% | 2,365 | 2,968 | +25% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 25 cases were attempted, and 23 counted toward the lift figure. The other 2 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +36 percentage points is the difference between those two pass rates over the 23 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.