Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Set up a local Kubernetes development loop with CAST AI cost monitoring. Use when building cost-aware deployments, testing autoscaler policies, or iterating on Terraform CAST AI configurations locally. Trigger with phrases like "cast ai dev setup", "cast ai local testing", "develop with cast ai", "cast ai terraform dev".
.claude/skills/jeremylongshore-castai-local-dev-loop/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-03 | ✗→✓ | ▲ Improved | -3% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 37% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 5% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 2% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 10% | 0% |
Make configuration feedback fast without treating a live cluster as a test fixture. Render, validate, diff, and policy-check locally; use a sandbox only for the behavior that cannot be proven offline.
Use Read and Grep to locate provider constraints, chart versions, values, annotations, policy definitions, generated files, and secret references. Identify which files are authoritative and which are derived.
Use Bash(terraform:_) to format and validate without applying. Use Bash(helm:_) to lint and render pinned charts. Use Bash(kubectl:\) only for client-side schema checks against rendered objects. Add deterministic tests for region, organization, policy bounds, automation state, HPA ownership, and prohibited secrets.
Use Write or Edit to add sanitized fixtures for missing identity, wrong region, permission denial, malformed configuration, conflicting controllers, unavailable metrics, PDB denial, and unsatisfied node constraints. Assert fail-closed behavior.
When installation behavior must be checked, use Bash(castctl:\) with the documented dry-run against the named sandbox context. Compare detected identity and proposed changes to approved fixtures before any connection.
Use a reviewed plan and one disposable workload or policy assignment. Observe only the intended behavior, never production data. Do not use a personal API key or console edits that bypass the repository source of truth.
Record commit, tool versions, rendered artifact hashes, plan summary, sandbox context, start/end time, cleanup, and remaining untested behavior. Return the sandbox to its declared baseline.
Use Read and Grep for source discovery. Use Write and Edit for fixtures and checks. Use Bash(terraform:_), Bash(helm:_), and Bash(kubectl:_) for offline validation; use Bash(castctl:_) only for a documented dry-run or approved sandbox action.
A policy change is tested against fixture assertions and a rendered workload annotation before one Deferred-mode sandbox canary. Production credentials and clusters never enter the local loop.
| Failure | Response | | -------------------------------- | --------------------------------------- | | Tool versions float | Pin them before comparing results | | Render needs a real secret | Replace it with a reference or sentinel | | Sandbox context is ambiguous | Stop before any cluster command | | Test requires production traffic | Redesign it around sanitized fixtures |
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 12,362 | 5,819 | -53% | 1 | 1 | 0% | 2,279 | 2,424 | +6% | 0 | 0 | — |
case-02 | fail→fail | 10,955 | 7,898 | -28% | 1 | 1 | 0% | 2,206 | 2,701 | +22% | 0 | 0 | — |
case-03 | fail→pass | 13,097 | 5,692 | -57% | 1 | 1 | 0% | 2,360 | 2,300 | -3% | 0 | 0 | — |
case-04 | fail→fail | 15,568 | 9,835 | -37% | 1 | 1 | 0% | 2,977 | 2,922 | -2% | 0 | 0 | — |
case-05 | fail→pass | 7,303 | 4,282 | -41% | 1 | 1 | 0% | 1,406 | 1,925 | +37% | 0 | 0 | — |
case-06 | pass→pass | 8,209 | 2,328 | -72% | 1 | 1 | 0% | 1,433 | 1,563 | +9% | 0 | 0 | — |
case-07 | fail→pass | 12,713 | 4,823 | -62% | 1 | 1 | 0% | 1,878 | 1,977 | +5% | 0 | 0 | — |
case-13 | fail→pass | 12,065 | 5,638 | -53% | 1 | 1 | 0% | 2,078 | 2,123 | +2% | 0 | 0 | — |
case-08 | pass→pass | 6,771 | 3,380 | -50% | 1 | 1 | 0% | 1,255 | 1,725 | +37% | 0 | 0 | — |
case-09 | fail→pass | 10,797 | 4,605 | -57% | 1 | 1 | 0% | 1,779 | 1,963 | +10% | 0 | 0 | — |
case-10 | fail→pass | 11,808 | 5,508 | -53% | 1 | 1 | 0% | 1,978 | 1,731 | -12% | 0 | 0 | — |
case-11 | pass→pass | 8,589 | 3,412 | -60% | 1 | 1 | 0% | 1,373 | 1,693 | +23% | 0 | 0 | — |
case-12 | fail→pass | 9,763 | 4,041 | -59% | 1 | 1 | 0% | 1,551 | 1,917 | +24% | 0 | 0 | — |
case-14 | fail→pass | 6,809 | 2,657 | -61% | 1 | 1 | 0% | 1,029 | 1,529 | +49% | 0 | 0 | — |
case-15 | pass→pass | 10,246 | 5,720 | -44% | 1 | 1 | 0% | 1,710 | 2,212 | +29% | 0 | 0 | — |
case-16 | pass→pass | 10,891 | 4,199 | -61% | 1 | 1 | 0% | 1,754 | 1,859 | +6% | 0 | 0 | — |
case-17 | fail→pass | 4,333 | 2,271 | -48% | 1 | 1 | 0% | 702 | 1,519 | +116% | 0 | 0 | — |
case-18 | fail→pass | 6,799 | 3,543 | -48% | 1 | 1 | 0% | 1,116 | 1,674 | +50% | 0 | 0 | — |
case-19 | pass→pass | 6,678 | 2,332 | -65% | 1 | 1 | 0% | 1,146 | 1,529 | +33% | 0 | 0 | — |
case-20 | pass→pass | 17,072 | 13,928 | -18% | 1 | 1 | 0% | 3,557 | 4,123 | +16% | 0 | 0 | — |
case-21 | pass→pass | 16,132 | 14,515 | -10% | 1 | 1 | 0% | 3,036 | 3,268 | +8% | 0 | 0 | — |
case-22 | pass→pass | 13,107 | 11,551 | -12% | 1 | 1 | 0% | 2,426 | 3,338 | +38% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +45 percentage points is the difference between those two pass rates over the 22 comparable cases.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.