Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Writes strong OKRs (Objectives and Key Results) with outcome-focused, inspiring objectives and measurable, ambitious, time-bound key results. Use this skill when the user asks to write, draft, review, critique, or improve OKRs, goals, or quarterly/annual objectives, when they mention "objectives and key results", "set goals for the team", "turn this strategy into OKRs", "are these good key results", or when distinguishing outcomes from outputs/tasks for planning and performance.
.claude/skills/jayrha-okr-writer/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-10 | ✗→✓ | ▲ Improved | 100% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 68% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 66% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 53% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 90% | 0% |
OKRs (Objectives and Key Results) are a goal-setting framework that links an inspiring qualitative Objective ("what we want to achieve") to 2-5 quantitative Key Results ("how we measure progress"). This skill produces OKRs that are outcome-focused (changes in behavior, value, or state — not task lists), measurable (a number anyone can verify), and appropriately ambitious.
Keywords: OKR, objective, key result, goal setting, KPI vs OKR, outcome vs output, stretch goal, quarterly planning, north star, measurable goals, scoring, grading.
Use this skill to: draft new OKRs from a vague goal or strategy, rewrite weak/task-based OKRs into outcome-based ones, review and score a draft set, or cascade company OKRs to teams.
Follow these steps in order. Do not skip the diagnosis step — most weak OKRs come from skipping it.
references/outcome-vs-output.md.references/objective-patterns.md.<verb> <metric> from <baseline> to <target> by <date>. Cover the outcome from multiple angles and include at least one counter-balancing / quality KR so the team can't game the headline number. See references/key-result-formulas.md.references/review-rubric.md (or scripts/okr_lint.py for automated checks). Fix any KR that is a task, lacks a number, lacks a baseline, or just measures activity.templates/okr-template.md. Include owner, period, confidence, and the initiatives separately from the KRs.Every key result should be reducible to:
> {Direction verb} {metric} from {baseline} to {target} by {deadline}.
Examples:
If you can write "Done / Not done" instead of a number, it is probably a task, not a key result. Binary milestones are allowed sparingly (e.g., "Achieve SOC 2 Type II certification") but prefer graded metrics.
Weak input: "Improve our mobile app and ship the redesign."
Diagnosis: This is an output. The implied outcome is usage/satisfaction. Baseline needed.
Strong OKR:
See examples/saas-growth-okrs.md and examples/rewrite-weak-okrs.md for full examples.
references/outcome-vs-output.md — how to convert outputs into outcomes, the "so that" laddering technique, and a large lookup table.references/objective-patterns.md — sentence patterns and a bank of strong/weak objective examples.references/key-result-formulas.md — KR metric types, the formula, counter-metrics, and verb lists.references/review-rubric.md — the scoring rubric to grade any draft OKR set.templates/okr-template.md — fill-in template for presenting a final OKR set.examples/saas-growth-okrs.md — a complete company → team cascade example.examples/rewrite-weak-okrs.md — before/after rewrites of common bad OKRs.scripts/okr_lint.py — runnable linter that flags task-like, baseline-less, or number-less key results.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 12,312 | 10,721 | -13% | 1 | 1 | 0% | 2,348 | 3,933 | +68% | 0 | 0 | — |
case-02 | pass→pass | 11,255 | 9,051 | -20% | 1 | 1 | 0% | 1,879 | 3,396 | +81% | 0 | 0 | — |
case-03 | pass→fail | 23,970 | 23,633 | -1% | 1 | 1 | 0% | 4,208 | 6,437 | +53% | 0 | 0 | — |
case-04 | pass→fail | 17,940 | 15,697 | -13% | 1 | 1 | 0% | 3,356 | 4,710 | +40% | 0 | 0 | — |
case-10 | fail→pass | 10,644 | 11,871 | +12% | 1 | 1 | 0% | 1,922 | 3,847 | +100% | 0 | 0 | — |
case-05 | pass→pass | 9,804 | 7,666 | -22% | 1 | 1 | 0% | 1,740 | 3,169 | +82% | 0 | 0 | — |
case-06 | pass→pass | 10,607 | 9,773 | -8% | 1 | 1 | 0% | 2,010 | 3,583 | +78% | 0 | 0 | — |
case-07 | fail→pass | 14,069 | 12,320 | -12% | 1 | 1 | 0% | 2,471 | 4,140 | +68% | 0 | 0 | — |
case-08 | fail→pass | 12,697 | 9,421 | -26% | 1 | 1 | 0% | 2,091 | 3,466 | +66% | 0 | 0 | — |
case-09 | fail→fail | 13,083 | 16,521 | +26% | 1 | 1 | 0% | 2,588 | 4,780 | +85% | 0 | 0 | — |
case-11 | fail→pass | 12,826 | 9,382 | -27% | 1 | 1 | 0% | 2,337 | 3,567 | +53% | 0 | 0 | — |
case-12 | fail→fail | 13,704 | 16,947 | +24% | 1 | 1 | 0% | 1,640 | 3,173 | +93% | 0 | 0 | — |
case-13 | fail→pass | 11,291 | 9,672 | -14% | 1 | 1 | 0% | 1,904 | 3,622 | +90% | 0 | 0 | — |
case-14 | pass→pass | 14,824 | 7,638 | -48% | 1 | 1 | 0% | 1,891 | 3,273 | +73% | 0 | 0 | — |
case-15 | pass→pass | 12,624 | 11,165 | -12% | 1 | 1 | 0% | 2,183 | 3,842 | +76% | 0 | 0 | — |
case-16 | pass→pass | 10,192 | 10,579 | +4% | 1 | 1 | 0% | 1,629 | 3,602 | +121% | 0 | 0 | — |
case-17 | pass→pass | 11,082 | 8,661 | -22% | 1 | 1 | 0% | 1,625 | 3,639 | +124% | 0 | 0 | — |
case-18 | pass→pass | 7,319 | 7,249 | -1% | 1 | 1 | 0% | 1,398 | 3,099 | +122% | 0 | 0 | — |
case-19 | pass→pass | 8,379 | 9,826 | +17% | 1 | 1 | 0% | 1,458 | 3,204 | +120% | 0 | 0 | — |
case-20 | pass→pass | 11,422 | 7,471 | -35% | 1 | 1 | 0% | 1,985 | 3,286 | +66% | 0 | 0 | — |
case-21 | pass→pass | 18,150 | 9,788 | -46% | 1 | 1 | 0% | 2,480 | 3,643 | +47% | 0 | 0 | — |
case-22 | pass→pass | 8,815 | 8,109 | -8% | 1 | 1 | 0% | 1,666 | 3,351 | +101% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +14 percentage points is the difference between those two pass rates over the 22 comparable cases. 2 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 8/3/2026 | +13% |
Other measured skills in the registry, with their headline benchmark lift.