Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Create a minimal working AssemblyAI transcription example. Use when starting a new AssemblyAI integration, testing your setup, or learning basic transcription patterns. Trigger with phrases like "assemblyai hello world", "assemblyai example", "assemblyai quick start", "simple assemblyai transcription".
.claude/skills/jeremylongshore-assemblyai-hello-world/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 41% | 0% |
| case-11 | ✓→✗ | ▼ Worse | 160% | 0% |
| case-02 | ✓→✓ | = Same ✓ | 83% | 0% |
| case-03 | ✓→✓ | = Same ✓ | 43% | 0% |
| case-04 | ✓→✓ | = Same ✓ | 41% | 0% |
Run a bounded AssemblyAI pre-recorded smoke test with explicit model fallback and safe output. Treat live audio, transcript content, credentials, spend, and destructive state as separately governed boundaries.
Current pre-recorded requests require an explicit speech_models list; there is no default. Submit a consented fixture, keep its transcript ID as the job handle, and require a completed terminal state plus content-safe assertions before declaring success.
For live work, inject ASSEMBLYAI_API_KEY from an approved secret manager and send the raw value only in the AssemblyAI Authorization header to the configured first-party host. Never print, commit, place in a URL, or expose it to an untrusted client. Callback secrets and temporary streaming tokens are separate credentials.
Use Read, Glob, and Grep to inspect repository code, configuration, fixtures, and evidence. Use Write and Edit only for approved implementation or documentation changes. Do not call AssemblyAI, upload audio, open a streaming session, mint a token, replay a callback, deploy, rotate a key, or delete a transcript merely because this skill was invoked.
Require an accountable owner before live audio processing, production credential or endpoint changes, paid model or capacity changes, content retention, callback replay, deployment, or deletion. Read-only repository inspection and synthetic offline validation do not authorize live vendor actions.
Return the operation scope, environment, region, contract surface, authorization class, model and feature decisions, deterministic validation results, content-free identifiers, risks, cleanup or rollback state, and a concise pass/fail receipt. Exclude credentials, signed URLs, audio, transcript text, prompts, and customer-derived content.
Rerun the smallest relevant deterministic check, compare actual state with the requested outcome and current first-party contract, verify sensitive fields are absent from evidence, and confirm rollback, termination, or deletion state before reporting success.
Review the dated first-party evidence map before relying on any model, parameter, limit, price, region, or lifecycle claim.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 7,952 | 5,527 | -30% | 1 | 1 | 0% | 1,661 | 2,345 | +41% | 0 | 0 | — |
case-02 | pass→pass | 5,935 | 4,447 | -25% | 1 | 1 | 0% | 1,146 | 2,096 | +83% | 0 | 0 | — |
case-03 | pass→pass | 8,457 | 5,152 | -39% | 1 | 1 | 0% | 1,463 | 2,091 | +43% | 0 | 0 | — |
case-04 | pass→pass | 7,117 | 4,529 | -36% | 1 | 1 | 0% | 1,512 | 2,129 | +41% | 0 | 0 | — |
case-05 | pass→pass | 7,698 | 3,724 | -52% | 1 | 1 | 0% | 1,448 | 1,952 | +35% | 0 | 0 | — |
case-06 | pass→pass | 8,361 | 3,454 | -59% | 1 | 1 | 0% | 1,469 | 1,812 | +23% | 0 | 0 | — |
case-07 | pass→pass | 7,561 | 7,685 | +2% | 1 | 1 | 0% | 1,542 | 2,782 | +80% | 0 | 0 | — |
case-08 | pass→pass | 8,583 | 3,065 | -64% | 1 | 1 | 0% | 1,670 | 1,786 | +7% | 0 | 0 | — |
case-09 | pass→pass | 2,387 | 1,778 | -26% | 1 | 1 | 0% | 372 | 1,440 | +287% | 0 | 0 | — |
case-10 | pass→pass | 5,662 | 3,431 | -39% | 1 | 1 | 0% | 1,069 | 1,843 | +72% | 0 | 0 | — |
case-11 | pass→fail | 3,159 | 1,634 | -48% | 1 | 1 | 0% | 535 | 1,391 | +160% | 0 | 0 | — |
case-12 | pass→pass | 6,806 | 5,189 | -24% | 1 | 1 | 0% | 1,402 | 2,230 | +59% | 0 | 0 | — |
case-17 | pass→pass | 4,583 | 1,933 | -58% | 1 | 1 | 0% | 820 | 1,509 | +84% | 0 | 0 | — |
case-13 | pass→pass | 10,759 | 6,874 | -36% | 1 | 1 | 0% | 2,153 | 2,417 | +12% | 0 | 0 | — |
case-14 | pass→pass | 8,184 | 4,772 | -42% | 1 | 1 | 0% | 1,544 | 2,201 | +43% | 0 | 0 | — |
case-15 | pass→pass | 6,712 | 3,774 | -44% | 1 | 1 | 0% | 1,167 | 1,857 | +59% | 0 | 0 | — |
case-16 | pass→pass | 3,963 | 2,850 | -28% | 1 | 1 | 0% | 728 | 1,642 | +126% | 0 | 0 | — |
case-22 | pass→pass | 13,132 | 13,278 | +1% | 1 | 1 | 0% | 2,499 | 3,956 | +58% | 0 | 0 | — |
case-18 | pass→pass | 3,063 | 2,214 | -28% | 1 | 1 | 0% | 470 | 1,513 | +222% | 0 | 0 | — |
case-19 | pass→pass | 6,220 | 3,860 | -38% | 1 | 1 | 0% | 1,178 | 1,869 | +59% | 0 | 0 | — |
case-20 | pass→pass | 5,949 | 3,854 | -35% | 1 | 1 | 0% | 1,069 | 1,890 | +77% | 0 | 0 | — |
case-21 | pass→pass | 3,049 | 1,646 | -46% | 1 | 1 | 0% | 488 | 1,397 | +186% | 0 | 0 | — |
case-23 | pass→pass | 10,859 | 11,095 | +2% | 1 | 1 | 0% | 2,189 | 3,584 | +64% | 0 | 0 | — |
case-24 | pass→pass | 8,069 | 7,250 | -10% | 1 | 1 | 0% | 1,468 | 2,592 | +77% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 24 cases were attempted. The headline lift of 0 percentage points is the difference between those two pass rates over the 24 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.