Loading skill
Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Design and evaluate paid-ad experiments with hypotheses, randomization units, sample-size and duration assumptions, guardrails, platform experiment tools, analysis, and decision rules. Use for A/B test, split test, experiment design, hypothesis, statistical significance, sample size, test duration, or experiment readout.
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | 43% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 34% | 0% |
| case-01 | ✗→✓ | ▲ Improved | 22% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 24% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 29% | 0% |
population, primary metric, guardrails, minimum effect, and stopping rule.
interference, and measurement quality.
effect and uncertainty.
Do not repeatedly peek and stop on a favorable result, call underpowered noise a winner, or generalize beyond the tested population.
Other measured skills in the registry, with their headline benchmark lift.