Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Generate a structured interview scorecard and interview guide for any role. Use when asked to create a hiring rubric, interview scorecard, structured interview guide, or assessment criteria for a job. Produces a scorecard with competencies, behavioural questions, and scoring guidance.
.claude/skills/mohitagw15856-hiring-rubric/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | 18% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 89% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 74% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 107% | 0% |
| case-15 | ✗→✓ | ▲ Improved | -21% | 0% |
This skill generates a complete structured interview scorecard and guide for any role. It reduces hiring bias, enables consistent evaluation across interviewers, and produces better hiring decisions.
Ask the user for these if not provided:
Level: Junior / Mid / Senior / Staff / Manager] Team: Team name] Created: Date]
Each competency is scored 1–4:
Hiring recommendation:
For each competency (generate 4–6 based on the role):
Why this matters for this role: One sentence — connects to actual job requirements]
What 4 looks like (Strong Yes): Specific, observable behaviours. "Proactively decomposed an ambiguous problem into a structured approach without prompting. Could articulate tradeoffs clearly and made assumptions explicit."]
What 2 looks like (Lean No): Specific, observable behaviours at the lower end. "Could answer direct questions but struggled when the interviewer removed scaffolding. Required significant prompting to reach a structured answer."]
Interview Questions (2–3 per competency):
If the role requires a technical screen, describe:]
2–3 values-based questions aligned to company values if provided, or general culture fit questions:]
5–7 specific red flags relevant to this role and level:]
Suggest how to divide competencies across interview rounds to avoid repetition:
| Round | Interviewer | Competencies to Assess | |---|---|---| | 1 — Recruiter Screen | Recruiter | Motivation, career narrative, basics | | 2 — Hiring Manager | Role] | Assign 2 competencies] | | 3 — Peer Interview | Role] | Assign 2 competencies] | | 4 — Stakeholder | Role] | Assign 1–2 competencies + culture] |
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-04 | pass→pass | 10,532 | 7,735 | -27% | 1 | 1 | 0% | 2,148 | 2,691 | +25% | 0 | 0 | — |
case-05 | pass→pass | 6,216 | 5,860 | -6% | 1 | 1 | 0% | 1,164 | 2,232 | +92% | 0 | 0 | — |
case-01 | fail→fail | 39,845 | 29,158 | -27% | 1 | 1 | 0% | 6,293 | 6,507 | +3% | 0 | 0 | — |
case-02 | fail→pass | 28,828 | 27,570 | -4% | 1 | 1 | 0% | 4,971 | 5,867 | +18% | 0 | 0 | — |
case-03 | fail→fail | 20,909 | 18,916 | -10% | 1 | 1 | 0% | 4,462 | 4,717 | +6% | 0 | 0 | — |
case-06 | pass→pass | 19,374 | 14,834 | -23% | 1 | 1 | 0% | 3,081 | 3,700 | +20% | 0 | 0 | — |
case-07 | fail→fail | 29,413 | 36,344 | +24% | 1 | 1 | 0% | 5,388 | 7,460 | +38% | 0 | 0 | — |
case-08 | fail→pass | 17,793 | 22,508 | +26% | 1 | 1 | 0% | 2,891 | 5,457 | +89% | 0 | 0 | — |
case-09 | fail→pass | 14,553 | 20,081 | +38% | 1 | 1 | 0% | 2,541 | 4,417 | +74% | 0 | 0 | — |
case-10 | fail→fail | 14,057 | 22,652 | +61% | 1 | 1 | 0% | 2,468 | 5,188 | +110% | 0 | 0 | — |
case-11 | fail→fail | 16,096 | 26,403 | +64% | 1 | 1 | 0% | 2,744 | 5,452 | +99% | 0 | 0 | — |
case-12 | fail→pass | 12,617 | 18,187 | +44% | 1 | 1 | 0% | 2,053 | 4,242 | +107% | 0 | 0 | — |
case-13 | pass→fail | 14,341 | 23,185 | +62% | 1 | 1 | 0% | 2,415 | 5,024 | +108% | 0 | 0 | — |
case-14 | pass→pass | 15,781 | 21,169 | +34% | 1 | 1 | 0% | 2,684 | 4,699 | +75% | 0 | 0 | — |
case-15 | fail→pass | 12,237 | 2,622 | -79% | 1 | 1 | 0% | 2,145 | 1,698 | -21% | 0 | 0 | — |
case-16 | fail→fail | 16,147 | 21,995 | +36% | 1 | 1 | 0% | 3,384 | 5,039 | +49% | 0 | 0 | — |
case-17 | pass→pass | 15,786 | 23,672 | +50% | 1 | 1 | 0% | 2,830 | 4,986 | +76% | 0 | 0 | — |
case-18 | fail→fail | 19,886 | 27,654 | +39% | 1 | 1 | 0% | 3,444 | 5,910 | +72% | 0 | 0 | — |
case-19 | fail→fail | 9,114 | 16,610 | +82% | 1 | 1 | 0% | 1,591 | 4,236 | +166% | 0 | 0 | — |
case-20 | fail→pass | 18,267 | 22,242 | +22% | 1 | 1 | 0% | 3,047 | 5,040 | +65% | 0 | 0 | — |
case-21 | pass→pass | 17,597 | 23,353 | +33% | 1 | 1 | 0% | 3,053 | 4,966 | +63% | 0 | 0 | — |
case-22 | pass→pass | 15,956 | 23,314 | +46% | 1 | 1 | 0% | 2,837 | 5,131 | +81% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +23 percentage points is the difference between those two pass rates over the 22 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.