Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Write a test strategy document from a feature spec, PRD, or system description. Use when asked to create a test plan, write a test strategy, define QA approach, or plan testing for a feature or release. Produces a complete test strategy with scope, risk assessment, test types, coverage targets, and a prioritised test case outline.
.claude/skills/mohitagw15856-test-strategy-doc/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | -21% | 0% |
| case-03 | ✗→✓ | ▲ Improved | -9% | 0% |
| case-04 | ✗→✓ | ▲ Improved | -5% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 42% | 0% |
| case-09 | ✗→✓ | ▲ Improved | -3% | 0% |
Produces a complete test strategy from a feature spec, PRD, or system description — covering scope, test types, risk areas, coverage requirements, and a prioritised test case outline.
Ask for these if not provided:
In scope:
Out of scope:
Assumptions:
Identify the highest-risk areas first — these drive depth and coverage:
| Area | Risk Level | Why | Test Priority | |---|---|---|---| | e.g. Payment processing] | High | Money movement, regulatory | P0 — exhaustive | | e.g. User authentication] | High | Security boundary | P0 — exhaustive | | e.g. Email notifications] | Medium | External dependency | P1 — happy path + key failures | | e.g. UI copy changes] | Low | Visual only, reversible | P2 — smoke only |
Unit Tests
Integration Tests
End-to-End Tests
Performance Tests (include if any row in the Risk Assessment table has performance as a risk factor, regardless of overall risk level)
Security Tests (include only if risk is high+)
Priority-ordered list of specific test cases:
P0 — Must pass before merge: | Test Case | Type | Expected Outcome | |---|---|---| | e.g. User can log in with valid credentials] | E2E | Redirect to dashboard, session created] | | e.g. Invalid login returns 401] | Integration | Error message displayed, no session] | | e.g. Password is never stored in plain text] | Unit | bcrypt hash in DB] |
P1 — Must pass before release: | Test Case | Type | Expected Outcome | |---|---|---| | e.g. Login fails gracefully when DB is down] | Integration | User sees friendly error, 503] | | e.g. Rate limiting blocks after 5 failed attempts] | Integration | 429 returned, account flagged] |
P2 — Should pass, can ship with known issues tracked: | Test Case | Type | Expected Outcome | |---|---|---| | e.g. Login page renders correctly on mobile] | E2E | Layout matches design] |
Testing is complete when:
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-11 | pass→pass | 25,493 | 48,100 | +89% | 1 | 1 | 0% | 4,740 | 8,280 | +75% | 0 | 0 | — |
case-01 | fail→pass | 57,091 | 33,464 | -41% | 1 | 1 | 0% | 8,330 | 6,591 | -21% | 0 | 0 | — |
case-02 | fail→fail | 64,435 | 37,264 | -42% | 1 | 1 | 0% | 7,528 | 6,273 | -17% | 0 | 0 | — |
case-03 | fail→pass | 72,005 | 42,062 | -42% | 1 | 1 | 0% | 7,542 | 6,845 | -9% | 0 | 0 | — |
case-04 | fail→pass | 41,719 | 33,878 | -19% | 1 | 1 | 0% | 6,537 | 6,218 | -5% | 0 | 0 | — |
case-05 | pass→pass | 36,806 | 39,162 | +6% | 1 | 1 | 0% | 4,255 | 6,290 | +48% | 0 | 0 | — |
case-06 | pass→pass | 30,295 | 36,164 | +19% | 1 | 1 | 0% | 4,675 | 5,773 | +23% | 0 | 0 | — |
case-07 | fail→fail | 34,161 | 27,478 | -20% | 1 | 1 | 0% | 3,940 | 4,726 | +20% | 0 | 0 | — |
case-08 | fail→pass | 30,863 | 31,449 | +2% | 1 | 1 | 0% | 3,645 | 5,181 | +42% | 0 | 0 | — |
case-09 | fail→pass | 48,681 | 65,054 | +34% | 1 | 1 | 0% | 6,460 | 6,255 | -3% | 0 | 0 | — |
case-10 | fail→pass | 22,914 | 5,111 | -78% | 1 | 1 | 0% | 2,414 | 2,233 | -7% | 0 | 0 | — |
case-12 | pass→pass | 23,570 | 24,641 | +5% | 1 | 1 | 0% | 2,753 | 4,278 | +55% | 0 | 0 | — |
case-13 | pass→pass | 22,017 | 20,662 | -6% | 1 | 1 | 0% | 2,489 | 4,694 | +89% | 0 | 0 | — |
case-14 | pass→pass | 30,133 | 19,571 | -35% | 1 | 1 | 0% | 3,898 | 5,445 | +40% | 0 | 0 | — |
case-15 | fail→pass | 27,157 | 35,385 | +30% | 1 | 1 | 0% | 3,026 | 5,411 | +79% | 0 | 0 | — |
case-16 | fail→pass | 48,102 | 33,803 | -30% | 1 | 1 | 0% | 8,234 | 6,299 | -24% | 0 | 0 | — |
case-17 | fail→pass | 36,964 | 28,917 | -22% | 1 | 1 | 0% | 4,891 | 5,138 | +5% | 0 | 0 | — |
case-18 | fail→fail | 29,101 | 28,107 | -3% | 1 | 1 | 0% | 3,485 | 5,260 | +51% | 0 | 0 | — |
case-19 | fail→pass | 31,655 | 32,902 | +4% | 1 | 1 | 0% | 4,097 | 5,115 | +25% | 0 | 0 | — |
case-20 | fail→pass | 27,775 | 20,685 | -26% | 1 | 1 | 0% | 3,487 | 4,591 | +32% | 0 | 0 | — |
case-21 | pass→pass | 48,952 | 36,263 | -26% | 1 | 1 | 0% | 7,511 | 6,550 | -13% | 0 | 0 | — |
case-22 | fail→pass | 27,908 | 28,984 | +4% | 1 | 1 | 0% | 3,652 | 5,460 | +50% | 0 | 0 | — |
case-23 | fail→pass | 34,239 | 43,932 | +28% | 1 | 1 | 0% | 3,957 | 7,014 | +77% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted. The headline lift of +57 percentage points is the difference between those two pass rates over the 23 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.