Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Write an internal KYC/AML escalation memo: a factual time-stamped trigger description, customer-profile vs activity mismatch analysis, red-flag taxonomy mapping, outstanding information, and a recommendation with rationale. Use when asked to escalate a KYC alert, document an AML concern, write up unusual-activity findings for compliance review, or prepare an enhanced due diligence referral. Produces a structured internal escalation memo for a compliance team's decision-makers.
.claude/skills/mohitagw15856-kyc-escalation/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 46% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 7% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 48% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 48% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 25% | 0% |
This skill helps compliance teams document internal escalations well: facts separated from inference, red flags mapped to a taxonomy, and a recommendation the MLRO or reporting decision-maker can act on. It documents and escalates — it does not decide, and it does not draft regulatory reports.
Boundaries — apply these without exception. Never draft a SAR/STR or its narrative — that is the designated reporting officer's regulated act; this memo is the internal input to that decision. Never advise on structuring transactions, avoiding detection, or evading monitoring, for any party. Never include tipping-off risk material — the memo is internal-only; do not draft customer-facing language about the investigation.
Ask for what's missing; never fabricate transaction details — mark gaps [not in file]:
1. Trigger description — facts only. What was observed, when, in what amounts, involving whom. Time-stamp everything. No adjectives, no inference — "three cash deposits of 9,400–9,800 on consecutive days", not "obvious structuring".
2. Profile vs activity mismatch. Two columns: what the KYC file says to expect (business type, turnover, geographies, counterparties, channels) vs what was observed. The mismatch — or its absence — is the analytical core. An alert consistent with a well-documented profile may support "clear"; activity inconsistent with the file is what escalates.
3. Red-flag taxonomy mapping. Map observations (never speculation) to categories: structuring patterns (amounts near reporting thresholds, split transactions); rapid movement/pass-through (in-and-out with no business purpose, layering hops); third parties (unexplained payers/payees, funnel patterns); jurisdiction risk (high-risk geography exposure inconsistent with profile); entity opacity (shell characteristics, nominee patterns, circular ownership); source-of-funds gaps (wealth/activity unexplained by the file); behavioural (reluctance to provide documents, unusual urgency, threshold awareness); adverse media / PEP or sanctions proximity (cite the specific source and date). Each flag cites its fact; list relevant categories checked and not present too.
4. Information still needed. What would resolve the ambiguity, and its source (customer outreach — flag tipping-off sensitivity for the decision-maker; internal records; registries; screening re-run). Distinguish "needed before any decision" from "needed for EDD".
5. Recommendation with rationale. Exactly one of: clear (documented, consistent explanation); enhanced due diligence (mismatch resolvable with more information); exit consideration (risk outside appetite regardless of reporting outcome — note exit timing may need the reporting decision-maker's input first); refer to reporting decision-maker (facts that a reasonable person could regard as grounds for suspicion). Two sentences of rationale tying flags to the recommendation.
1. Trigger — time-stamped facts. 2. Customer profile summary — risk rating, expected activity, tenure. 3. Profile vs activity — expected | observed table. 4. Red flags — category | observation | source/date. Plus categories checked, not present. 5. Prior history — earlier alerts and outcomes. 6. Information still needed — item | source | blocking or EDD-stage. 7. Recommendation & rationale — one of the four, two-sentence rationale.
End with: "This memo is analytical support for internal escalation, not a compliance determination. Reporting, exit, and customer-contact decisions rest with your institution's designated decision-makers under its policy and applicable regulation."
[not in file], never filled in[not in file]| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 13,801 | 13,675 | -1% | 1 | 1 | 0% | 2,746 | 4,008 | +46% | 0 | 0 | — |
case-02 | fail→pass | 20,100 | 14,841 | -26% | 1 | 1 | 0% | 3,737 | 4,008 | +7% | 0 | 0 | — |
case-03 | fail→pass | 12,638 | 10,930 | -14% | 1 | 1 | 0% | 2,429 | 3,583 | +48% | 0 | 0 | — |
case-08 | fail→pass | 15,220 | 14,534 | -5% | 1 | 1 | 0% | 2,761 | 4,090 | +48% | 0 | 0 | — |
case-04 | pass→fail | 8,590 | 5,776 | -33% | 1 | 1 | 0% | 1,612 | 2,238 | +39% | 0 | 0 | — |
case-05 | pass→fail | 6,397 | 7,661 | +20% | 1 | 1 | 0% | 1,235 | 2,593 | +110% | 0 | 0 | — |
case-06 | pass→pass | 9,061 | 10,484 | +16% | 1 | 1 | 0% | 1,585 | 3,023 | +91% | 0 | 0 | — |
case-07 | fail→pass | 15,884 | 11,772 | -26% | 1 | 1 | 0% | 2,907 | 3,625 | +25% | 0 | 0 | — |
case-09 | pass→pass | 15,237 | 12,877 | -15% | 1 | 1 | 0% | 2,876 | 4,477 | +56% | 0 | 0 | — |
case-10 | fail→pass | 12,962 | 11,053 | -15% | 1 | 1 | 0% | 2,386 | 3,509 | +47% | 0 | 0 | — |
case-11 | fail→pass | 12,081 | 14,450 | +20% | 1 | 1 | 0% | 2,153 | 4,147 | +93% | 0 | 0 | — |
case-12 | pass→pass | 15,105 | 14,996 | -1% | 1 | 1 | 0% | 2,856 | 4,191 | +47% | 0 | 0 | — |
case-13 | fail→pass | 12,659 | 10,282 | -19% | 1 | 1 | 0% | 2,278 | 3,217 | +41% | 0 | 0 | — |
case-14 | pass→pass | 10,655 | 9,432 | -11% | 1 | 1 | 0% | 2,070 | 2,965 | +43% | 0 | 0 | — |
case-15 | pass→pass | 15,120 | 13,009 | -14% | 1 | 1 | 0% | 2,685 | 3,769 | +40% | 0 | 0 | — |
case-16 | fail→fail | 14,569 | 12,048 | -17% | 1 | 1 | 0% | 2,594 | 3,633 | +40% | 0 | 0 | — |
case-17 | fail→pass | 10,269 | 12,467 | +21% | 1 | 1 | 0% | 1,936 | 3,732 | +93% | 0 | 0 | — |
case-18 | pass→pass | 16,523 | 12,230 | -26% | 1 | 1 | 0% | 2,969 | 3,632 | +22% | 0 | 0 | — |
case-19 | fail→pass | 11,067 | 13,199 | +19% | 1 | 1 | 0% | 2,103 | 3,696 | +76% | 0 | 0 | — |
case-20 | fail→fail | 16,819 | 13,128 | -22% | 1 | 1 | 0% | 3,033 | 3,774 | +24% | 0 | 0 | — |
case-21 | fail→fail | 12,514 | 12,010 | -4% | 1 | 1 | 0% | 2,275 | 3,617 | +59% | 0 | 0 | — |
case-22 | fail→pass | 13,885 | 11,245 | -19% | 1 | 1 | 0% | 2,618 | 3,437 | +31% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +41 percentage points is the difference between those two pass rates over the 22 comparable cases. 2 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.