Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Cut reasoning-token weight. Use when thinking or reasoning tokens dominate the bill, when asked to cap a thinking budget, lower effort, or stop simple tasks from burning extended thinking. Light tasks get light thinking: matches thinking budget to task weight, then verifies accuracy held before keeping any cap. Never caps blind.
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | -37% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -47% | 0% |
| case-03 | ✗→✓ | ▲ Improved | -42% | 0% |
| case-04 | ✗→✓ | ▲ Improved | -47% | 0% |
| case-10 | ✗→✓ | ▲ Improved | -48% | 0% |
Reasoning tokens bill like output tokens, and models usually lock their answer early in the chain; the tail is often narration you pay for. Light tasks get light thinking. Save the heavy reasoning for the problems that are actually heavy.
questions with one right answer. Extended thinking rarely changes these.
boundaries, anything where the first idea is usually wrong. Thinking earns its tokens here.
a thinking-budget setting, or per-request flags. Cut the budget hard for light tasks; leave heavy tasks alone at first. Prefer per-session or per-request controls: a cap written into settings outlives the task, and a cap in project settings caps your teammates too. Never ship a cap in shared config.
capped pile and compare results against what you got uncapped. Same quality: keep the cap. Worse: raise it back and say so.
the saving is a receipt, not a feeling.
Other measured skills in the registry, with their headline benchmark lift.