Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Tune CodeRabbit review configuration: learnings, code guidelines, and noise reduction. Use when fine-tuning review quality, training CodeRabbit with team preferences, adding code guidelines, or reducing false positives. Trigger with phrases like "coderabbit tune reviews", "coderabbit learnings", "coderabbit guidelines", "reduce coderabbit noise", "coderabbit false positives".
.claude/skills/jeremylongshore-coderabbit-core-workflow-b/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 27% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 42% | 0% |
| case-04 | ✗→✓ | ▲ Improved | -16% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 45% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 90% | 0% |
Improve relevance using supported knowledge layers. Avoid invented configuration keys and judge changes against labeled real findings.
references/official-docs.md and re-check any time-sensitive contract before execution.CLAUDE.md and AGENTS.md.Treat Git-provider sessions, CodeRabbit web sessions, CLI credentials, and CodeRabbit API keys as separate credentials. Use only an already-approved session or secret-manager reference, never print a secret, and do not place credentials in .coderabbit.yaml, source files, logs, or deliverables.
Require owner approval before changing shared learnings, global overrides, or exclusions. Keep analysis and drafts local until approval is explicit, and record who approved the action and its scope.
A labeled sample, hypothesis, supported change, comparison metrics, and rollback decision. Include source dates, unknowns, and the exact boundary between observed fact and recommendation.
| Condition | Response | |---|---| | Current contract is unclear or docs disagree | Stop mutation, cite both sources, and request owner resolution. | | Required access or approval is missing | Produce a draft and evidence plan only. | | Validation or pilot behavior differs from expectation | Restore the prior state and retain the failed evidence. | | Output contains secrets or private code | Stop, quarantine the artifact, redact it, and notify the data owner. |
Move a durable convention into AGENTS.md and verify detection.
Replace unsupported custom_patterns with documented path instructions.
references/official-docs.md.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 16,706 | 10,785 | -35% | 1 | 1 | 0% | 2,801 | 3,565 | +27% | 0 | 0 | — |
case-02 | fail→fail | 18,256 | 12,243 | -33% | 1 | 1 | 0% | 3,370 | 4,127 | +22% | 0 | 0 | — |
case-03 | fail→pass | 10,628 | 4,713 | -56% | 1 | 1 | 0% | 1,836 | 2,606 | +42% | 0 | 0 | — |
case-04 | fail→pass | 13,053 | 3,309 | -75% | 1 | 1 | 0% | 2,639 | 2,217 | -16% | 0 | 0 | — |
case-05 | pass→pass | 9,401 | 5,010 | -47% | 1 | 1 | 0% | 1,825 | 2,615 | +43% | 0 | 0 | — |
case-06 | pass→pass | 9,526 | 4,580 | -52% | 1 | 1 | 0% | 1,611 | 2,473 | +54% | 0 | 0 | — |
case-07 | fail→fail | 12,245 | 4,073 | -67% | 1 | 1 | 0% | 2,083 | 2,399 | +15% | 0 | 0 | — |
case-08 | fail→pass | 9,412 | 3,470 | -63% | 1 | 1 | 0% | 1,604 | 2,326 | +45% | 0 | 0 | — |
case-09 | pass→pass | 6,745 | 3,244 | -52% | 1 | 1 | 0% | 1,247 | 2,207 | +77% | 0 | 0 | — |
case-10 | fail→pass | 7,566 | 3,309 | -56% | 1 | 1 | 0% | 1,200 | 2,280 | +90% | 0 | 0 | — |
case-11 | pass→pass | 7,678 | 3,277 | -57% | 1 | 1 | 0% | 1,230 | 2,287 | +86% | 0 | 0 | — |
case-12 | fail→pass | 15,092 | 9,716 | -36% | 1 | 1 | 0% | 2,395 | 3,471 | +45% | 0 | 0 | — |
case-13 | pass→pass | 10,432 | 3,317 | -68% | 1 | 1 | 0% | 1,733 | 2,234 | +29% | 0 | 0 | — |
case-14 | pass→pass | 11,302 | 4,414 | -61% | 1 | 1 | 0% | 1,768 | 2,296 | +30% | 0 | 0 | — |
case-15 | fail→pass | 10,427 | 3,834 | -63% | 1 | 1 | 0% | 1,734 | 2,441 | +41% | 0 | 0 | — |
case-16 | fail→pass | 10,329 | 5,752 | -44% | 1 | 1 | 0% | 1,621 | 2,594 | +60% | 0 | 0 | — |
case-17 | fail→pass | 8,442 | 7,578 | -10% | 1 | 1 | 0% | 1,419 | 2,947 | +108% | 0 | 0 | — |
case-18 | pass→pass | 9,382 | 6,225 | -34% | 1 | 1 | 0% | 1,733 | 2,780 | +60% | 0 | 0 | — |
case-19 | pass→pass | 12,110 | 3,901 | -68% | 1 | 1 | 0% | 1,999 | 2,307 | +15% | 0 | 0 | — |
case-20 | fail→fail | 13,030 | 9,818 | -25% | 1 | 1 | 0% | 2,423 | 3,533 | +46% | 0 | 0 | — |
case-21 | fail→fail | 11,937 | 7,196 | -40% | 1 | 1 | 0% | 2,207 | 2,995 | +36% | 0 | 0 | — |
case-22 | fail→fail | 12,242 | 9,546 | -22% | 1 | 1 | 0% | 2,533 | 3,546 | +40% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +41 percentage points is the difference between those two pass rates over the 22 comparable cases.
The publisher has shipped newer versions since this run, so these numbers describe v1, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.