Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Collect AssemblyAI debug evidence for support tickets and troubleshooting. Use when encountering persistent issues, preparing support tickets, or collecting diagnostic information for AssemblyAI problems. Trigger with phrases like "assemblyai debug", "assemblyai support bundle", "collect assemblyai logs", "assemblyai diagnostic".
.claude/skills/jeremylongshore-assemblyai-debug-bundle/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | -17% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 42% | 0% |
| case-04 | ✗→✓ | ▲ Improved | 72% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 113% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 79% | 0% |
Produce a redacted AssemblyAI diagnostic bundle for support and incident triage. Treat live audio, transcript content, credentials, spend, and destructive state as separately governed boundaries.
Allowlisted evidence includes hosts, SDK version, operation, models, hashed IDs, status transitions, timestamps, durations, response or close codes, delivery attempts, and termination state. Exclude authorization values, signed URLs, token responses, audio, transcript text, prompts, and full callback bodies by default.
For live work, inject ASSEMBLYAI_API_KEY from an approved secret manager and send the raw value only in the AssemblyAI Authorization header to the configured first-party host. Never print, commit, place in a URL, or expose it to an untrusted client. Callback secrets and temporary streaming tokens are separate credentials.
Use Read, Glob, and Grep to inspect repository code, configuration, fixtures, and evidence. Use Write and Edit only for approved implementation or documentation changes. Do not call AssemblyAI, upload audio, open a streaming session, mint a token, replay a callback, deploy, rotate a key, or delete a transcript merely because this skill was invoked.
Require an accountable owner before live audio processing, production credential or endpoint changes, paid model or capacity changes, content retention, callback replay, deployment, or deletion. Read-only repository inspection and synthetic offline validation do not authorize live vendor actions.
Return the operation scope, environment, region, contract surface, authorization class, model and feature decisions, deterministic validation results, content-free identifiers, risks, cleanup or rollback state, and a concise pass/fail receipt. Exclude credentials, signed URLs, audio, transcript text, prompts, and customer-derived content.
Rerun the smallest relevant deterministic check, compare actual state with the requested outcome and current first-party contract, verify sensitive fields are absent from evidence, and confirm rollback, termination, or deletion state before reporting success.
Review the dated first-party evidence map before relying on any model, parameter, limit, price, region, or lifecycle claim.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 20,491 | 8,088 | -61% | 1 | 1 | 0% | 4,287 | 3,550 | -17% | 0 | 0 | — |
case-02 | fail→pass | 12,755 | 10,406 | -18% | 1 | 1 | 0% | 3,000 | 4,264 | +42% | 0 | 0 | — |
case-03 | fail→fail | 12,363 | 11,047 | -11% | 1 | 1 | 0% | 2,417 | 3,638 | +51% | 0 | 0 | — |
case-04 | fail→pass | 10,183 | 8,034 | -21% | 1 | 1 | 0% | 1,821 | 3,127 | +72% | 0 | 0 | — |
case-05 | pass→pass | 6,096 | 3,714 | -39% | 1 | 1 | 0% | 1,068 | 2,237 | +109% | 0 | 0 | — |
case-06 | fail→fail | 28,508 | 3,055 | -89% | 1 | 1 | 0% | 1,050 | 2,204 | +110% | 0 | 0 | — |
case-07 | fail→pass | 5,775 | 2,779 | -52% | 1 | 1 | 0% | 964 | 2,056 | +113% | 0 | 0 | — |
case-08 | fail→pass | 9,410 | 8,262 | -12% | 1 | 1 | 0% | 1,839 | 3,290 | +79% | 0 | 0 | — |
case-09 | pass→pass | 11,399 | 3,223 | -72% | 1 | 1 | 0% | 1,927 | 2,283 | +18% | 0 | 0 | — |
case-10 | fail→fail | 6,397 | 2,156 | -66% | 1 | 1 | 0% | 1,070 | 2,037 | +90% | 0 | 0 | — |
case-11 | pass→pass | 11,366 | 4,427 | -61% | 1 | 1 | 0% | 2,109 | 2,486 | +18% | 0 | 0 | — |
case-12 | pass→pass | 10,947 | 3,327 | -70% | 1 | 1 | 0% | 1,949 | 2,340 | +20% | 0 | 0 | — |
case-13 | fail→pass | 4,816 | 3,151 | -35% | 1 | 1 | 0% | 788 | 2,186 | +177% | 0 | 0 | — |
case-14 | fail→pass | 13,483 | 3,690 | -73% | 1 | 1 | 0% | 2,412 | 2,387 | -1% | 0 | 0 | — |
case-15 | fail→pass | 8,880 | 2,606 | -71% | 1 | 1 | 0% | 1,400 | 2,102 | +50% | 0 | 0 | — |
case-16 | fail→pass | 12,115 | 1,851 | -85% | 1 | 1 | 0% | 2,018 | 1,989 | -1% | 0 | 0 | — |
case-17 | pass→pass | 6,506 | 2,693 | -59% | 1 | 1 | 0% | 1,064 | 2,071 | +95% | 0 | 0 | — |
case-18 | fail→pass | 8,732 | 3,021 | -65% | 1 | 1 | 0% | 1,421 | 2,219 | +56% | 0 | 0 | — |
case-19 | fail→pass | 11,995 | 4,361 | -64% | 1 | 1 | 0% | 2,079 | 2,433 | +17% | 0 | 0 | — |
case-20 | pass→pass | 11,548 | 8,619 | -25% | 1 | 1 | 0% | 2,455 | 3,352 | +37% | 0 | 0 | — |
case-21 | pass→pass | 9,424 | 4,233 | -55% | 1 | 1 | 0% | 1,656 | 2,523 | +52% | 0 | 0 | — |
case-22 | pass→pass | 12,187 | 11,593 | -5% | 1 | 1 | 0% | 2,459 | 4,112 | +67% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 21 counted toward the lift figure. The other 1 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +50 percentage points is the difference between those two pass rates over the 21 comparable cases.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.