Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Pre-design research for Maestro cards: creates or validates a same-card research.md receipt, maps context, stakeholders, hosting, unknowns, and the first design fork. Use before maestro-design when the user brings a new idea, zero-context feature, external/pasted plan, unfamiliar domain, stakeholder-heavy request, hosting-unclear work, or when research is missing, stale, skipped-risky, or needed for READY_FOR_DESIGN.
.claude/skills/reinamaccredy-maestro-research/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-21 | ✗→✓ | ▲ Improved | 55% | 0% |
| case-22 | ✗→✓ | ▲ Improved | 110% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 92% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 138% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 105% | 0% |
Research is breadth. Design is depth. Use this skill to decide whether a card has enough valid context to enter maestro-design.
Research writes a same-card research.md receipt or a skip receipt. It does not lock decisions, write acceptance criteria, finalize handoffs, create tasks, start implementation, browse by default, or approve build.
Activate with a known session id: maestro hook record --event skill_activation --skill maestro-research --session <session_id>
Exact command signatures live in reference/cli.md when present. If the local binary does not list a verb, do not invent it.
Use this skill when the request is zero-context, unfamiliar-domain, externally pasted, stakeholder-heavy, hosting-unclear, or when maestro-design lacks fresh research.
Skip research only with a durable skip receipt when one condition is true:
Fresh research means the same card and the same problem statement, as_of within 7 days, unchanged invalidates_when, and a still-valid READY_FOR_DESIGN gate or skip.
Run maestro status, then maestro active. If the idea is external, sandbox-bound, or hosting-unclear, record that before writing into the current repo. Use the same card for research and later design. Exploratory ideas may start as Idea cards; intended product work may start as proposed Features. Completion criterion: one owning card and one Hosting value are named, or the gate is not READY_FOR_DESIGN.
Search precedent with maestro grep "<topic>" corpus:memory, then read the pasted source, relevant repo artifacts, and stakeholder facts already available. If evidence answers a question, inspect it instead of asking the user. Completion criterion: the brief can cite what is known from artifacts, what is assumed, and what remains unknown.
Create or update same-card research.md. If replacing prior research, preserve the old brief as research.md.bak or a Superseded pointer, and add a Maestro note with the reason and changed facts. Keep exactly one current research.md.
Required receipt sections:
skipped, skip_reason, skipped_bycurrent-repo, sandbox-repo, or external, with rationaleas_of, invalidates_whenCompletion criterion: every required section is present, even when the section says None.
Choose exactly one:
READY_FOR_DESIGNNEEDS_STAKEHOLDERNEEDS_EVIDENCEPIVOTSTOPREADY_FOR_DESIGN requires no blocking unknowns, no open stakeholder actions, compatible hosting, whitelisted skip evidence when skipped, and one concrete first design fork. Duplicate or prior-art discovery gates STOP or PIVOT with a pointer to the existing card.
Completion criterion: the final response names the gate, the durable receipt, and the next Maestro route.
Landscape is a map, not a decision. It may list directions such as browser extension, omnichannel assistant, or dedicated inbox. The first design fork is one entry question for depth, such as "Where should Copilot live in the Sales workflow?"
Stakeholder actions have status:
textopen | resolved | superseded
A resolved answer may create new blockers. If a deferred unknown becomes acceptance-critical during the first design fork, maestro-design may bounce back to this skill.
High-risk user skips are allowed only as user-directed risk:
textskipped: true skipped_by: user unresolved_risks: - <security, legal, customer-impacting, or data risk>
midstream.
READY_FOR_DESIGN without a first design fork.guidance.
Use reference/examples.md for regression examples, including Sales Copilot, skip receipts, stale research, and hosting mismatch.
READY_FOR_DESIGN plus compatible hosting -> maestro-design reads research.md first.
NEEDS_STAKEHOLDER or NEEDS_EVIDENCE -> stay on the same card and resolve the listed blocker.
PIVOT -> update the problem statement or create a linked successor.
STOP -> preserve the receipt and archive or close the candidate when the owning lifecycle allows it.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-21 | fail→pass | 10,637 | 7,198 | -32% | 1 | 1 | 0% | 1,458 | 2,265 | +55% | 0 | 0 | — |
case-22 | fail→pass | 10,924 | 12,684 | +16% | 1 | 1 | 0% | 1,271 | 2,668 | +110% | 0 | 0 | — |
case-01 | fail→fail | 14,695 | 6,817 | -54% | 1 | 1 | 0% | 377 | 1,753 | +365% | 0 | 0 | — |
case-02 | fail→fail | 7,582 | 23,390 | +208% | 1 | 1 | 0% | 293 | 4,964 | +1594% | 0 | 0 | — |
case-03 | fail→fail | 14,516 | 8,319 | -43% | 1 | 1 | 0% | 1,659 | 1,681 | +1% | 0 | 0 | — |
case-04 | pass→fail | 23,827 | 24,919 | +5% | 1 | 1 | 0% | 3,515 | 4,780 | +36% | 0 | 0 | — |
case-15 | pass→pass | 11,552 | 8,583 | -26% | 1 | 1 | 0% | 1,467 | 2,519 | +72% | 0 | 0 | — |
case-05 | pass→fail | 9,697 | 20,625 | +113% | 1 | 1 | 0% | 1,487 | 4,145 | +179% | 0 | 0 | — |
case-06 | pass→fail | 21,662 | 10,409 | -52% | 1 | 1 | 0% | 3,054 | 1,730 | -43% | 0 | 0 | — |
case-07 | fail→fail | 19,614 | 6,833 | -65% | 1 | 1 | 0% | 2,241 | 1,659 | -26% | 0 | 0 | — |
case-08 | fail→pass | 8,721 | 4,659 | -47% | 1 | 1 | 0% | 1,097 | 2,105 | +92% | 0 | 0 | — |
case-09 | fail→pass | 13,141 | 20,602 | +57% | 1 | 1 | 0% | 1,628 | 3,882 | +138% | 0 | 0 | — |
case-10 | fail→fail | 4,474 | 10,096 | +126% | 1 | 1 | 0% | 423 | 2,733 | +546% | 0 | 0 | — |
case-11 | fail→pass | 8,095 | 3,882 | -52% | 1 | 1 | 0% | 991 | 2,033 | +105% | 0 | 0 | — |
case-12 | fail→pass | 91,927 | 4,734 | -95% | 1 | 1 | 0% | 2,251 | 1,922 | -15% | 0 | 0 | — |
case-13 | fail→pass | 13,642 | 12,195 | -11% | 1 | 1 | 0% | 1,905 | 2,952 | +55% | 0 | 0 | — |
case-14 | pass→pass | 13,025 | 6,308 | -52% | 1 | 1 | 0% | 1,157 | 2,423 | +109% | 0 | 0 | — |
case-16 | fail→pass | 14,155 | 11,022 | -22% | 1 | 1 | 0% | 1,831 | 2,398 | +31% | 0 | 0 | — |
case-17 | pass→pass | 8,041 | 6,792 | -16% | 1 | 1 | 0% | 1,060 | 2,339 | +121% | 0 | 0 | — |
case-18 | fail→pass | 11,333 | 5,153 | -55% | 1 | 1 | 0% | 1,488 | 2,232 | +50% | 0 | 0 | — |
case-19 | fail→pass | 9,447 | 7,110 | -25% | 1 | 1 | 0% | 1,431 | 2,263 | +58% | 0 | 0 | — |
case-20 | fail→pass | 16,257 | 2,212 | -86% | 1 | 1 | 0% | 1,578 | 1,649 | +4% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 17 counted toward the lift figure. The other 5 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +36 percentage points is the difference between those two pass rates over the 17 comparable cases. 4 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.