Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Append research process and results to the current Phase's context file. Each append MUST contain >=500 lines of markdown covering both process and results. Use this skill at plan-designated checkpoint points — typically after each strategy completes or at key decision nodes within a research Phase.
.claude/skills/yogsoth-ai-context-checkpoint/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 111% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 22% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 236% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 142% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 13% | 0% |
Append research process and results to the current Phase's context file.
Multiple times within a Phase, at plan-designated checkpoint points:
step: "import context-management:context-checkpoint"Typically triggered after each strategy completes or at key decision nodes.
Import context-init. This is idempotent — if the context file for the current Phase already exists, it skips creation and returns the existing file path.
Determine the current Phase's context file path by:
context/INDEX.md for the most recent entry matching the current PhaseAppend a new section to the context file:
markdown--- ## Checkpoint: <Descriptive Name> <CC writes substantive content here covering process + results>
Content format: CC has full autonomy. A default semi-structured template is available as guidance but not mandatory:
markdown--- ## Checkpoint: <Descriptive Name> ### Objective What this stage aimed to accomplish. ### Process Summary What was done — searches performed, papers read, methods applied, decisions made along the way. ### Key Findings The substantive results — discoveries, patterns, important papers, technical details. ### Decisions Made Choices made during this stage and their rationale. ### Open Questions What remains unresolved, what needs further investigation.
CC may use this template, modify it, combine sections, add new sections, or write in completely free-form style. The only requirement is coverage of both process and results.
Update the row for the current context file:
date +%Y-%m-%d-%H-%M for current time)The checkpoint is a detailed record for future reference. Write as if the reader has zero context about what happened during this research stage. Include:
Sparse checkpoints are useless for recovery. Write generously where the content warrants it — this is a research log, not a summary.
<!-- BEGIN available-tables (generated) -->
Optional, no fixed order; the final leaf is always a sop.
| SOP | When to use | | --- | --- | | context-init | Create a new context file for a research Phase. Called once at Phase start to initialize the file that subsequent context-checkpoint calls will append to. Use this skill whenever a new research Phase begins and a fresh context file is needed. |
<!-- END available-tables (generated) -->
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 14,691 | 23,813 | +62% | 1 | 1 | 0% | 2,340 | 4,934 | +111% | 0 | 0 | — |
case-02 | fail→pass | 17,413 | 20,852 | +20% | 1 | 1 | 0% | 2,942 | 3,590 | +22% | 0 | 0 | — |
case-03 | fail→fail | 13,281 | 5,541 | -58% | 1 | 1 | 0% | 2,269 | 1,237 | -45% | 0 | 0 | — |
case-04 | pass→pass | 13,590 | 11,281 | -17% | 1 | 1 | 0% | 2,253 | 2,700 | +20% | 0 | 0 | — |
case-05 | fail→fail | 1,969 | 8,879 | +351% | 1 | 1 | 0% | 281 | 1,288 | +358% | 0 | 0 | — |
case-06 | pass→fail | 5,242 | 6,773 | +29% | 1 | 1 | 0% | 417 | 1,299 | +212% | 0 | 0 | — |
case-07 | fail→fail | 11,893 | 4,972 | -58% | 1 | 1 | 0% | 2,370 | 1,092 | -54% | 0 | 0 | — |
case-08 | fail→pass | 5,611 | 12,650 | +125% | 1 | 1 | 0% | 775 | 2,605 | +236% | 0 | 0 | — |
case-09 | pass→fail | 6,878 | 5,746 | -16% | 1 | 1 | 0% | 1,049 | 1,264 | +20% | 0 | 0 | — |
case-10 | fail→pass | 6,149 | 12,657 | +106% | 1 | 1 | 0% | 1,216 | 2,939 | +142% | 0 | 0 | — |
case-11 | fail→fail | 7,174 | 6,771 | -6% | 1 | 1 | 0% | 1,173 | 1,174 | +0% | 0 | 0 | — |
case-12 | fail→pass | 12,192 | 10,271 | -16% | 1 | 1 | 0% | 1,676 | 1,901 | +13% | 0 | 0 | — |
case-13 | fail→pass | 8,119 | 2,566 | -68% | 1 | 1 | 0% | 1,365 | 1,235 | -10% | 0 | 0 | — |
case-14 | fail→fail | 14,367 | 18,002 | +25% | 1 | 1 | 0% | 2,237 | 3,514 | +57% | 0 | 0 | — |
case-15 | pass→pass | 3,749 | 6,450 | +72% | 1 | 1 | 0% | 754 | 2,121 | +181% | 0 | 0 | — |
case-16 | pass→pass | 6,930 | 10,450 | +51% | 1 | 1 | 0% | 1,331 | 2,903 | +118% | 0 | 0 | — |
case-17 | pass→fail | 9,598 | 4,891 | -49% | 1 | 1 | 0% | 1,867 | 1,207 | -35% | 0 | 0 | — |
case-18 | fail→pass | 8,413 | 1,924 | -77% | 1 | 1 | 0% | 1,480 | 1,094 | -26% | 0 | 0 | — |
case-19 | pass→pass | 14,380 | 7,136 | -50% | 1 | 1 | 0% | 1,126 | 1,844 | +64% | 0 | 0 | — |
case-20 | fail→pass | 10,106 | 12,841 | +27% | 1 | 1 | 0% | 1,647 | 2,483 | +51% | 0 | 0 | — |
case-21 | fail→pass | 10,596 | 16,746 | +58% | 1 | 1 | 0% | 2,076 | 3,408 | +64% | 0 | 0 | — |
case-22 | fail→fail | 4,800 | 5,050 | +5% | 1 | 1 | 0% | 709 | 1,098 | +55% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 14 counted toward the lift figure. The other 8 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +27 percentage points is the difference between those two pass rates over the 14 comparable cases. 4 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.