Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Build a complete RAG pipeline with Cohere Chat, Embed, and Rerank. Use when implementing retrieval-augmented generation, building grounded Q&A systems, or combining search with LLM generation. Trigger with phrases like "cohere RAG", "cohere retrieval", "cohere grounded generation", "cohere search and answer".
.claude/skills/jeremylongshore-cohere-core-workflow-a/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 9% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 3% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 22% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 29% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 84% | 0% |
Create a measurable retrieval pipeline that separates indexing, query embedding, candidate retrieval, reranking, generation, and citation checks.
Use Read, Glob, and Grep to inspect code, configuration, and evidence. Use WebFetch only for current Cohere primary documentation. Use Write or Edit only when the user requested implementation and the exact target files are known; never write credentials or customer content.
input_type=search_document and user queries with input_type=search_query.Use an environment-specific key injected from an approved secret manager. Never print, persist, commit, or place CO_API_KEY in an example. Confirm access with the least costly bounded operation appropriate to the task, and treat key creation, rotation, revocation, role changes, and production-capacity requests as owner-approved actions.
Do not expose or rotate keys, change Cohere Team roles, accept commercial terms, enable sensitive production data, increase spend or capacity, switch production models, send a support bundle, or execute model-proposed side effects without the accountable owner's approval. Keep diagnosis read-only unless implementation was requested.
Return the resolved API and model contract, files or settings inspected, evidence collected, validation result, remaining risk, owner, and rollback or next action. Redact keys, authorization headers, prompts, retrieved documents, embeddings, customer identifiers, and unrestricted environment output.
| Condition | Response | |---|---| | Dimension mismatch | Stop and rebuild or select the index matching its recorded model contract. | | No candidates | Return a grounded no-answer result instead of unconstrained generation. | | Invalid citation | Reject or flag the answer and retain the evidence bundle. | | Context overflow | Reduce or rechunk evidence; v2 does not provide legacy prompt truncation. |
Use this compact handoff shape to keep the selected scope, validation evidence, and operational result reviewable.
Input:
textquery=approved-fixture; retrieve=40; rerank=8; require-citations=true
Expected handoff:
textanswer=grounded; citations=valid; selected-doc-ids=recorded; eval=pass
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 16,848 | 10,867 | -35% | 1 | 1 | 0% | 3,805 | 4,142 | +9% | 0 | 0 | — |
case-02 | fail→pass | 18,539 | 10,629 | -43% | 1 | 1 | 0% | 4,076 | 4,202 | +3% | 0 | 0 | — |
case-03 | fail→pass | 17,630 | 14,532 | -18% | 1 | 1 | 0% | 4,121 | 5,009 | +22% | 0 | 0 | — |
case-04 | pass→pass | 8,764 | 5,204 | -41% | 1 | 1 | 0% | 1,597 | 2,620 | +64% | 0 | 0 | — |
case-05 | fail→pass | 9,192 | 3,219 | -65% | 1 | 1 | 0% | 1,745 | 2,255 | +29% | 0 | 0 | — |
case-06 | pass→pass | 3,806 | 2,862 | -25% | 1 | 1 | 0% | 755 | 2,206 | +192% | 0 | 0 | — |
case-07 | pass→pass | 4,657 | 3,897 | -16% | 1 | 1 | 0% | 901 | 2,332 | +159% | 0 | 0 | — |
case-08 | fail→pass | 6,160 | 2,426 | -61% | 1 | 1 | 0% | 1,140 | 2,102 | +84% | 0 | 0 | — |
case-09 | fail→pass | 13,794 | 6,327 | -54% | 1 | 1 | 0% | 2,638 | 2,854 | +8% | 0 | 0 | — |
case-10 | pass→pass | 5,380 | 5,384 | +0% | 1 | 1 | 0% | 1,038 | 2,686 | +159% | 0 | 0 | — |
case-11 | fail→pass | 15,596 | 8,820 | -43% | 1 | 1 | 0% | 3,015 | 3,278 | +9% | 0 | 0 | — |
case-12 | pass→pass | 11,710 | 8,791 | -25% | 1 | 1 | 0% | 2,104 | 3,406 | +62% | 0 | 0 | — |
case-13 | pass→pass | 12,989 | 1,822 | -86% | 1 | 1 | 0% | 2,681 | 2,014 | -25% | 0 | 0 | — |
case-14 | fail→pass | 6,284 | 2,323 | -63% | 1 | 1 | 0% | 1,131 | 2,086 | +84% | 0 | 0 | — |
case-15 | pass→pass | 8,066 | 2,272 | -72% | 1 | 1 | 0% | 1,554 | 2,084 | +34% | 0 | 0 | — |
case-16 | pass→pass | 5,898 | 3,038 | -48% | 1 | 1 | 0% | 1,107 | 2,165 | +96% | 0 | 0 | — |
case-17 | pass→pass | 3,975 | 2,773 | -30% | 1 | 1 | 0% | 698 | 2,099 | +201% | 0 | 0 | — |
case-18 | pass→pass | 8,525 | 4,416 | -48% | 1 | 1 | 0% | 1,752 | 2,501 | +43% | 0 | 0 | — |
case-19 | fail→fail | 18,502 | 15,053 | -19% | 1 | 1 | 0% | 3,690 | 4,768 | +29% | 0 | 0 | — |
case-20 | pass→pass | 15,929 | 14,517 | -9% | 1 | 1 | 0% | 3,020 | 4,703 | +56% | 0 | 0 | — |
case-21 | fail→fail | 9,081 | 5,990 | -34% | 1 | 1 | 0% | 1,764 | 2,957 | +68% | 0 | 0 | — |
case-22 | fail→fail | 13,774 | 10,701 | -22% | 1 | 1 | 0% | 2,798 | 4,099 | +46% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +36 percentage points is the difference between those two pass rates over the 22 comparable cases.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.