Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Production readiness checklist for CAST AI cluster onboarding. Use when going live with CAST AI autoscaling, validating Phase 2 setup, or preparing for production cost optimization. Trigger with phrases like "cast ai production", "cast ai go-live", "cast ai checklist", "cast ai launch".
.claude/skills/jeremylongshore-castai-prod-checklist/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 43% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -1% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 0% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 25% | 0% |
| case-11 | ✗→✓ | ▲ Improved | -36% | 0% |
Approve production only when identity, ownership, connectivity, observability, scaling safety, cost interpretation, rollback, and evidence are all explicit. An unchecked or unknown item is a stop, not a soft pass.
Use Read and Grep to prove cluster, organization, region, cloud account, IaC source, reconcilers, and secret references. Confirm one owner for node provisioning, vertical rightsizing, and each HPA.
Use Bash(castctl:_) for supported version or status inspection, Bash(helm:_) for pinned release evidence, and Bash(kubectl:\) for component readiness and warning events. Confirm least-privilege cloud permissions and required outbound connectivity.
Review node templates, maximum CPU, cloud quotas, protected namespaces, workload policy assignment, minimum/maximum requests, apply mode, recommendation confidence, PDB behavior, HPA bounds, and ownership transfer. Reject deprecated cluster minimum CPU configuration.
Confirm baseline source, representative data window, public versus adjusted prices, automation adoption, and reconciliation owner. Do not use available or modeled savings as an unconditional release gate.
Use Bash(terraform:_) for the saved plan and Bash(helm:_) for the reviewed render. Confirm no unexpected deletion, disconnect, IAM expansion, CRD replacement, webhook collision, or automation enablement. Use Write or Edit to record tested rollback steps and thresholds.
Record PASS only when every required item has evidence. Otherwise record HOLD with the exact owner and missing proof. Approve one cluster ring and one automation dimension at a time.
Use Read and Grep for evidence review. Use Write and Edit for the signed gate record. Use Bash(kubectl:_), Bash(helm:_), Bash(terraform:_), and Bash(castctl:_) for non-secret inspection, render, and plan operations; do not perform the production change from the checklist.
A cluster passes observation readiness but holds node automation because Karpenter ownership is unresolved. Another passes a Deferred workload canary while managed HPA takeover remains out of scope.
| Failure | Response | | ----------------------------------------- | ---------------------------------------- | | Evidence is stale or from another cluster | Mark the gate HOLD | | Ownership is shared or implicit | Resolve control authority before release | | Rollback is untested | Limit to observation mode | | Secret appears in an artifact | Stop, rotate, and regenerate evidence |
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 20,509 | 22,441 | +9% | 1 | 1 | 0% | 3,498 | 5,011 | +43% | 0 | 0 | — |
case-02 | fail→pass | 20,786 | 15,678 | -25% | 1 | 1 | 0% | 3,297 | 3,274 | -1% | 0 | 0 | — |
case-03 | fail→fail | 35,214 | 12,125 | -66% | 1 | 1 | 0% | 1,474 | 3,188 | +116% | 0 | 0 | — |
case-04 | pass→pass | 16,971 | 12,607 | -26% | 1 | 1 | 0% | 3,193 | 3,427 | +7% | 0 | 0 | — |
case-05 | pass→pass | 12,137 | 9,721 | -20% | 1 | 1 | 0% | 2,094 | 2,705 | +29% | 0 | 0 | — |
case-06 | pass→pass | 7,556 | 6,316 | -16% | 1 | 1 | 0% | 1,627 | 2,170 | +33% | 0 | 0 | — |
case-07 | pass→pass | 13,841 | 7,211 | -48% | 1 | 1 | 0% | 2,084 | 2,181 | +5% | 0 | 0 | — |
case-08 | fail→pass | 14,901 | 8,186 | -45% | 1 | 1 | 0% | 2,289 | 2,283 | -0% | 0 | 0 | — |
case-09 | pass→pass | 16,327 | 5,070 | -69% | 1 | 1 | 0% | 2,425 | 1,757 | -28% | 0 | 0 | — |
case-10 | fail→pass | 7,546 | 3,542 | -53% | 1 | 1 | 0% | 1,151 | 1,434 | +25% | 0 | 0 | — |
case-11 | fail→pass | 15,775 | 4,184 | -73% | 1 | 1 | 0% | 2,583 | 1,665 | -36% | 0 | 0 | — |
case-12 | fail→pass | 12,652 | 3,418 | -73% | 1 | 1 | 0% | 2,063 | 1,261 | -39% | 0 | 0 | — |
case-13 | pass→pass | 7,101 | 3,103 | -56% | 1 | 1 | 0% | 1,309 | 1,471 | +12% | 0 | 0 | — |
case-14 | pass→pass | 8,236 | 3,636 | -56% | 1 | 1 | 0% | 1,329 | 1,534 | +15% | 0 | 0 | — |
case-15 | pass→pass | 8,375 | 3,066 | -63% | 1 | 1 | 0% | 1,572 | 1,496 | -5% | 0 | 0 | — |
case-16 | pass→pass | 11,764 | 5,520 | -53% | 1 | 1 | 0% | 1,918 | 2,036 | +6% | 0 | 0 | — |
case-17 | fail→pass | 21,596 | 3,862 | -82% | 1 | 1 | 0% | 1,074 | 1,613 | +50% | 0 | 0 | — |
case-18 | fail→pass | 13,530 | 6,452 | -52% | 1 | 1 | 0% | 2,530 | 2,192 | -13% | 0 | 0 | — |
case-19 | fail→pass | 11,008 | 5,800 | -47% | 1 | 1 | 0% | 1,694 | 1,845 | +9% | 0 | 0 | — |
case-20 | pass→pass | 7,264 | 2,772 | -62% | 1 | 1 | 0% | 1,055 | 1,252 | +19% | 0 | 0 | — |
case-21 | fail→pass | 11,946 | 12,007 | +1% | 1 | 1 | 0% | 2,087 | 3,104 | +49% | 0 | 0 | — |
case-22 | fail→pass | 5,450 | 2,098 | -62% | 1 | 1 | 0% | 754 | 1,301 | +73% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 20 counted toward the lift figure. The other 2 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +50 percentage points is the difference between those two pass rates over the 20 comparable cases.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.