Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Diagnose ClickHouse cluster health and provide concrete remediation.
.claude/skills/frankchen021-diagnose-clickhouse-clusters/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-17 | ✗→✓ | ▲ Improved | 16% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 46% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 77% | 0% |
| case-15 | ✗→✓ | ▲ Improved | 113% | 0% |
| case-18 | ✗→✓ | ▲ Improved | 38% | 0% |
collect_cluster_status before health conclusions about current cluster health.collect_rca_evidence directly when the symptom and target are already clear. Use collect_cluster_status first only when you need current health context, severity/outliers, or help choosing the RCA symptom/scope.high_part_count and unknown.status_analysis_mode="windowed" and reuse the same time window in follow-up calls.visualization skill. Do not emit chart specs directly from this skill.Do not hardcode parts thresholds in responses. Use the thresholds and severities returned by collect_cluster_status.
Use one of these two formats:
Always print a table title line exactly before the table: ### Summary. | Status | Nodes with Issues | Checks Run | Timestamp | |--------|-------------------|------------|-----------| | 🟢 OK / 🟠 WARNING / 🔴 CRITICAL | N | categories | ISO8601 |
Always print a table title line exactly before the table: ### Findings by Category. Use a markdown table (not bullets) with one row per category. Required columns: | Category | Status | Key Metrics | Top Outlier / Scope | Notes | |----------|--------|-------------|----------------------|-------| | parts / errors / replication / ... | 🟢 OK / 🟠 WARNING / 🔴 CRITICAL | concise metric values with thresholds | node/table if present, else - | one short phrase |
Table rules:
collect_cluster_status in stable order.🟠 WARNING), never emoji-only.Key Metrics, put the 1-2 most important metrics only (single-line, semicolon-separated if needed).Notes as compact key/value items (single-line).max_parts_per_table=533 (>500)), avoid prose-heavy sentences. db.table or db ) in all table cells.Notes as compact comma-separated items.Top Outlier / Scope to -.Use compact structure only:
cause | support_score | evidence.In evidence, render up to 3 evidence_for items prefixed with ✓ and up to 2 evidence_against items prefixed with ✗, separated by <br/>. When excluded_candidates is non-empty, include at least one excluded reason as a ✗ item for the most relevant row. Evidence fidelity rules:
candidate.evidence_for and candidate.evidence_against from collect_rca_evidence for that row.observations, other candidates, or status output into the evidence cell.candidate.evidence_for or candidate.evidence_against.indicators_matched/indicators_checked, but never imply more matched checks than the tool returned.Formatting rule: print the line 3. **Possible Actions**, then a blank line, then an indented nested numbered list using exactly 1., 2., 3.. Do not continue the outer top-level numbering for action items.
Formatting rule: print the line 4. **Gaps / Next Checks**, then a blank line, then indented bullets using exactly -.
RCA brevity limits:
collect_cluster_status before giving any opinion on current health.status_analysis_mode="windowed" when user asks for a bounded time window or historical context.collect_rca_evidence. collect_cluster_status is optional unless current health context is needed.gaps[] is non-empty, explicitly state what evidence is missing.support_score < 0.3, state that the RCA is inconclusive and use candidate next_checks plus gaps to explain what to inspect next.0.30-0.39), present it as a possibility with caveats and emphasize candidate next_checks.evidence_for and evidence_against.collect_rca_evidence.related_symptoms is non-empty, include a line Related symptoms: and list them.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 17,152 | 3,985 | -77% | 1 | 1 | 0% | 3,082 | 1,828 | -41% | 0 | 0 | — |
case-17 | fail→pass | 18,488 | 2,805 | -85% | 1 | 1 | 0% | 1,664 | 1,928 | +16% | 0 | 0 | — |
case-22 | fail→fail | 8,694 | 2,557 | -71% | 1 | 1 | 0% | 1,377 | 2,069 | +50% | 0 | 0 | — |
case-02 | fail→fail | 21,030 | 6,926 | -67% | 1 | 1 | 0% | 2,209 | 1,982 | -10% | 0 | 0 | — |
case-03 | fail→fail | 20,475 | 9,068 | -56% | 1 | 1 | 0% | 3,431 | 2,319 | -32% | 0 | 0 | — |
case-04 | pass→fail | 4,737 | 6,311 | +33% | 1 | 1 | 0% | 842 | 2,168 | +157% | 0 | 0 | — |
case-05 | pass→pass | 12,092 | 11,087 | -8% | 1 | 1 | 0% | 2,184 | 3,660 | +68% | 0 | 0 | — |
case-06 | pass→pass | 10,462 | 10,316 | -1% | 1 | 1 | 0% | 1,959 | 3,480 | +78% | 0 | 0 | — |
case-07 | fail→fail | 15,906 | 9,544 | -40% | 1 | 1 | 0% | 3,490 | 2,167 | -38% | 0 | 0 | — |
case-08 | fail→fail | 14,557 | 24,725 | +70% | 1 | 1 | 0% | 2,390 | 2,282 | -5% | 0 | 0 | — |
case-09 | fail→fail | 17,916 | 6,469 | -64% | 1 | 1 | 0% | 3,099 | 1,893 | -39% | 0 | 0 | — |
case-10 | fail→pass | 12,613 | 7,386 | -41% | 1 | 1 | 0% | 1,881 | 2,746 | +46% | 0 | 0 | — |
case-11 | pass→fail | 11,720 | 6,234 | -47% | 1 | 1 | 0% | 1,754 | 1,994 | +14% | 0 | 0 | — |
case-12 | pass→pass | 12,637 | 4,010 | -68% | 1 | 1 | 0% | 1,929 | 2,217 | +15% | 0 | 0 | — |
case-13 | fail→pass | 8,429 | 4,283 | -49% | 1 | 1 | 0% | 1,299 | 2,300 | +77% | 0 | 0 | — |
case-14 | pass→fail | 9,090 | 3,303 | -64% | 1 | 1 | 0% | 1,430 | 2,141 | +50% | 0 | 0 | — |
case-15 | fail→pass | 6,377 | 2,757 | -57% | 1 | 1 | 0% | 910 | 1,942 | +113% | 0 | 0 | — |
case-16 | fail→fail | 9,338 | 3,620 | -61% | 1 | 1 | 0% | 1,448 | 2,143 | +48% | 0 | 0 | — |
case-18 | fail→pass | 10,500 | 4,596 | -56% | 1 | 1 | 0% | 1,742 | 2,411 | +38% | 0 | 0 | — |
case-19 | fail→fail | 16,121 | 4,040 | -75% | 1 | 1 | 0% | 3,082 | 1,936 | -37% | 0 | 0 | — |
case-20 | pass→pass | 11,969 | 5,319 | -56% | 1 | 1 | 0% | 1,832 | 2,473 | +35% | 0 | 0 | — |
case-21 | fail→fail | 11,301 | 3,418 | -70% | 1 | 1 | 0% | 1,886 | 2,141 | +14% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 13 counted toward the lift figure. The other 9 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +9 percentage points is the difference between those two pass rates over the 13 comparable cases. 3 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.