Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Full crystallization strategy for users who have no research direction at all. Covers actor profiling, landscape reconnaissance, direction narrowing, obstacle analysis, goal decomposition, and north-star synthesis. Use when the user's first message reveals zero specificity about what they want to research.
.claude/skills/yogsoth-ai-cold-start/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-06 | ✗→✓ | ▲ Improved | -16% | 0% |
| case-08 | ✗→✓ | ▲ Improved | -3% | 0% |
| case-14 | ✗→✓ | ▲ Improved | 34% | 0% |
| case-17 | ✗→✓ | ▲ Improved | -1% | 0% |
| case-19 | ✗→✓ | ▲ Improved | 27% | 0% |
The user knows nothing — they want to publish at a top venue but have no idea what to research.
All SOPs in this strategy follow these rules:
| Tactic | Purpose | |--------|---------| | actor-profiling | Understand who the user is | | landscape-reconnaissance | Broad, shallow field exploration | | direction-narrowing | Focus within chosen field(s) | | obstacle-analysis | Identify and mitigate barriers | | goal-decomposition | KAOS-style AND/OR goal structuring | | north-star-synthesis | Converge into North Star + ResearchBrief |
actor-profiling → landscape-reconnaissance → direction-narrowing
→ obstacle-analysis → goal-decomposition → north-star-synthesisThis is a reference, not a mandate. You decide the actual execution path.
You are the general. This strategy gives you:
What you decide:
The only non-negotiable: the process ends with north-star-synthesis producing a North Star + ResearchBrief that the user confirms.
<!-- BEGIN available-tables (generated) -->
Optional, no fixed order; the final leaf is always a sop.
| Tactic | When to use | | --- | --- | | actor-profiling | Understand who the user is — background, resources, constraints, and deep motivations. Produces an ActorProfile that informs all downstream decisions. Use this tactic at the start of any crystallization process to build a model of the user's capabilities, limitations, and intent. | | direction-narrowing | Focus within the user's chosen field(s). Identify specific sub-directions through deep paper and web research, then present ranked candidates. Use after landscape-reconnaissance has identified fields of interest. | | goal-decomposition | Structure the user's chosen direction into a formal goal tree using KAOS-style AND/OR decomposition. Validate feasibility against ActorProfile and ObstacleReport. Use after obstacle-analysis confirms the direction is viable. | | landscape-reconnaissance | Broad, shallow exploration of candidate research fields. Understand what's out there before narrowing. Use when the user needs to discover which fields are available to them — especially in cold-start and warm-start scenarios. | | north-star-synthesis | Converge all accumulated context into a crystallized North Star statement and structured ResearchBrief. Performs self-review before presenting to user. Use as the final tactic in any start mode — this is where everything comes together. | | obstacle-analysis | Identify what blocks the user from pursuing their chosen direction, assess severity, propose mitigations with search-validated evidence, and get user acceptance. Use after direction-narrowing has identified a specific direction. |
<!-- END available-tables (generated) -->
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 17,356 | 5,032 | -71% | 1 | 1 | 0% | 2,707 | 1,758 | -35% | 0 | 0 | — |
case-02 | fail→fail | 17,439 | 4,462 | -74% | 1 | 1 | 0% | 2,648 | 1,574 | -41% | 0 | 0 | — |
case-03 | fail→fail | 8,839 | 4,611 | -48% | 1 | 1 | 0% | 1,384 | 1,565 | +13% | 0 | 0 | — |
case-04 | fail→fail | 12,445 | 4,341 | -65% | 1 | 1 | 0% | 1,927 | 1,505 | -22% | 0 | 0 | — |
case-05 | pass→pass | 10,145 | 5,237 | -48% | 1 | 1 | 0% | 1,567 | 1,671 | +7% | 0 | 0 | — |
case-06 | fail→pass | 10,883 | 3,739 | -66% | 1 | 1 | 0% | 1,720 | 1,438 | -16% | 0 | 0 | — |
case-07 | pass→pass | 6,464 | 5,303 | -18% | 1 | 1 | 0% | 1,030 | 1,568 | +52% | 0 | 0 | — |
case-08 | fail→pass | 12,593 | 6,518 | -48% | 1 | 1 | 0% | 1,891 | 1,840 | -3% | 0 | 0 | — |
case-09 | pass→pass | 9,669 | 5,393 | -44% | 1 | 1 | 0% | 1,413 | 1,725 | +22% | 0 | 0 | — |
case-10 | pass→pass | 12,642 | 6,432 | -49% | 1 | 1 | 0% | 1,955 | 1,939 | -1% | 0 | 0 | — |
case-11 | pass→fail | 17,082 | 6,244 | -63% | 1 | 1 | 0% | 2,661 | 1,901 | -29% | 0 | 0 | — |
case-12 | pass→pass | 15,600 | 3,916 | -75% | 1 | 1 | 0% | 2,382 | 1,542 | -35% | 0 | 0 | — |
case-13 | pass→pass | 13,093 | 3,561 | -73% | 1 | 1 | 0% | 2,148 | 1,411 | -34% | 0 | 0 | — |
case-14 | fail→pass | 9,905 | 6,567 | -34% | 1 | 1 | 0% | 1,517 | 2,027 | +34% | 0 | 0 | — |
case-15 | pass→pass | 15,310 | 6,343 | -59% | 1 | 1 | 0% | 2,508 | 1,948 | -22% | 0 | 0 | — |
case-16 | pass→pass | 14,830 | 4,974 | -66% | 1 | 1 | 0% | 2,377 | 1,680 | -29% | 0 | 0 | — |
case-17 | fail→pass | 10,125 | 4,265 | -58% | 1 | 1 | 0% | 1,499 | 1,488 | -1% | 0 | 0 | — |
case-18 | fail→fail | 4,908 | 2,883 | -41% | 1 | 1 | 0% | 755 | 1,296 | +72% | 0 | 0 | — |
case-19 | fail→pass | 12,597 | 10,398 | -17% | 1 | 1 | 0% | 1,975 | 2,500 | +27% | 0 | 0 | — |
case-20 | pass→fail | 17,763 | 7,976 | -55% | 1 | 1 | 0% | 2,649 | 2,211 | -17% | 0 | 0 | — |
case-21 | pass→fail | 14,411 | 6,725 | -53% | 1 | 1 | 0% | 2,788 | 2,096 | -25% | 0 | 0 | — |
case-22 | fail→fail | 4,050 | 3,144 | -22% | 1 | 1 | 0% | 619 | 1,395 | +125% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +9 percentage points is the difference between those two pass rates over the 22 comparable cases. 3 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.