Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Autonomous goal execution — give a goal, get a plan, confirm, execute, report. You steer, Claude drives.
.claude/skills/brianrwagner-go-mode/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 111% | 0% |
| case-04 | ✗→✓ | ▲ Improved | 106% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 51% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 644% | 0% |
| case-07 | ✗→✓ | ▲ Improved | -16% | 0% |
Give me a goal. I'll plan it, confirm with you, execute it, and report back. You steer — I drive.
Detect from context or ask: "Just do it, plan first, or plan + phase approvals?"
| Mode | What you get | Best for | |------|-------------|----------| | quick | 1-line plan → you confirm → execute | Simple tasks, clear goals | | standard | Full plan → you confirm → execute → report (default) | Most tasks | | deep | Full plan → risk review → confirm each phase → execute → report | High-stakes or multi-system tasks |
Default: standard — use quick for simple, clear goals. Use deep when mistakes would be expensive to undo.
GOAL → PLAN → CONFIRM → EXECUTE → REPORTWhen given a goal, break it down:
Output a structured plan:
## 🎯 Goal: [restated goal]
### Definition of Done
[What success looks like]
### Plan
| # | Step | Tool/Skill | Est. Time | Cost | Risk |
|---|------|-----------|-----------|------|------|
| 1 | ... | ... | ... | ... | ... |
### Total Estimate
- **Time:** X minutes
- **API Cost:** ~$X.XX
- **Human Checkpoints:** [list]
### Guardrails Triggered
- [ ] External communication (needs approval)
- [ ] Financial spend > $1
- [ ] Irreversible actionPresent the plan and wait for approval:
Never skip confirmation. This is the human's steering wheel.
Run each step sequentially:
When all steps complete:
## ✅ Goal Complete: [goal]
### What Was Done
- Step 1: [result]
- Step 2: [result]
- ...
### Outputs
- [List of files, links, artifacts created]
### What Was Learned
- [Insights discovered during execution]
### Recommended Next Steps
- [What to do with the results]
- [Follow-up opportunities]
### Stats
- Total time: Xm
- API calls: X
- Est. cost: $X.XXWhen planning, draw from this toolkit:
| Tool | Use For | |------|---------| | web_search | Quick web lookups | | web_fetch | Read full web pages | | qmd search | Search Obsidian vault knowledge base | | content-research-writer skill | Deep research + writing | | research-coordinator skill | Multi-source research |
| Tool | Use For | |------|---------| | content-atomizer skill | Turn 1 piece → 13+ posts | | direct-response-copy skill | Sales copy | | seo-content skill | SEO articles | | newsletter skill | Newsletter editions | | email-sequences skill | Email flows | | nano-banana skill | Image generation (Gemini) |
| Tool | Use For | |------|---------| | positioning-angles skill | Find hooks that sell | | keyword-research skill | SEO keyword strategy | | business-prospecting skill | Lead research | | landing-page-design skill | Landing pages | | page-cro skill | Conversion optimization |
| Tool | Use For | |------|---------| | bird CLI | Twitter/X (read, post, reply) | | Gmail | Email (read, send) | | Notion | Pages and databases | | Telegram | Messaging |
| Tool | Use For | |------|---------| | exec | Shell commands | | codex | Code generation (GPT) | | claude | Code generation (Claude) | | File tools | Read, write, edit files |
Goal: "Research our top 3 competitors in the AI assistant space and build a comparison page"
Plan:
1. Identify top 3 competitors (web search) — 5min
2. Research each: pricing, features, reviews — 15min
3. Build comparison matrix — 10min
4. Write comparison page copy — 15min
5. Create visual comparison table — 5min
Total: ~50min, ~$0.50 API costGoal: "Take my latest blog post and turn it into a week of social content"
Plan:
1. Read and analyze the blog post — 2min
2. Extract key themes and quotes — 5min
3. Generate 5 Twitter threads — 15min
4. Generate 5 LinkedIn posts — 15min
5. Create 3 image prompts + generate visuals — 10min
6. Build content calendar — 5min
Total: ~52min, ~$1.00 API costGoal: "Find 20 potential clients in the SaaS space who might need our marketing services"
Plan:
1. Define ideal client profile — 5min
2. Search for SaaS companies (web) — 15min
3. Research each company's marketing gaps — 20min
4. Score and rank prospects — 10min
5. Build outreach-ready prospect list — 10min
6. Draft personalized intro messages [NEEDS APPROVAL] — 15min
Total: ~75min, ~$0.75 API costGoal: "Create 3 SEO-optimized blog posts for our target keywords"
Plan:
1. Review target keyword list — 2min
2. Research top-ranking content for each keyword — 15min
3. Create outlines using SEO skill — 10min
4. Write article 1 — 15min
5. Write article 2 — 15min
6. Write article 3 — 15min
7. Add internal links and meta descriptions — 10min
Total: ~82min, ~$2.00 API costGoal: "Prepare everything needed to launch our new product next Tuesday"
Plan:
1. Audit what exists (landing page, emails, social) — 10min
2. Identify gaps — 5min
3. Write launch email sequence (3 emails) — 20min
4. Create social media posts (Twitter, LinkedIn) — 15min
5. Generate launch graphics — 10min
6. Build launch day timeline — 5min
7. Draft press/outreach messages [NEEDS APPROVAL] — 15min
Total: ~80min, ~$2.50 API costGoal: "Review this week's metrics, summarize wins/losses, and plan next week's priorities"
Plan:
1. Pull metrics from available sources — 10min
2. Summarize key wins — 5min
3. Identify what didn't work — 5min
4. Review upcoming calendar — 5min
5. Propose next week's top 3 priorities — 10min
6. Create actionable task list — 5min
Total: ~40min, ~$0.25 API costJust tell the agent your goal in natural language:
> "Take the wheel: Research the top 5 AI newsletter tools, compare them, and recommend the best one for a solopreneur"
> "Take the wheel: Build a complete email welcome sequence for new subscribers — 5 emails over 2 weeks"
> "Take the wheel: Audit our Twitter presence and create a 30-day content strategy"
The agent will plan, confirm, execute, and report. You stay in control at every checkpoint.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 11,799 | 10,415 | -12% | 1 | 1 | 0% | 1,975 | 4,164 | +111% | 0 | 0 | — |
case-02 | fail→fail | 9,033 | 7,242 | -20% | 1 | 1 | 0% | 1,543 | 3,622 | +135% | 0 | 0 | — |
case-03 | fail→fail | 8,941 | 7,805 | -13% | 1 | 1 | 0% | 1,539 | 3,757 | +144% | 0 | 0 | — |
case-04 | fail→pass | 10,980 | 7,157 | -35% | 1 | 1 | 0% | 1,799 | 3,697 | +106% | 0 | 0 | — |
case-05 | fail→pass | 15,124 | 6,844 | -55% | 1 | 1 | 0% | 2,290 | 3,455 | +51% | 0 | 0 | — |
case-06 | fail→pass | 2,391 | 3,998 | +67% | 1 | 1 | 0% | 394 | 2,930 | +644% | 0 | 0 | — |
case-07 | fail→pass | 29,739 | 12,432 | -58% | 1 | 1 | 0% | 5,367 | 4,490 | -16% | 0 | 0 | — |
case-08 | fail→pass | 10,741 | 6,877 | -36% | 1 | 1 | 0% | 1,567 | 3,536 | +126% | 0 | 0 | — |
case-09 | fail→pass | 3,413 | 6,416 | +88% | 1 | 1 | 0% | 334 | 3,487 | +944% | 0 | 0 | — |
case-10 | pass→pass | 5,974 | 3,248 | -46% | 1 | 1 | 0% | 882 | 2,961 | +236% | 0 | 0 | — |
case-11 | fail→pass | 12,366 | 3,750 | -70% | 1 | 1 | 0% | 1,921 | 2,931 | +53% | 0 | 0 | — |
case-12 | fail→pass | 10,640 | 1,534 | -86% | 1 | 1 | 0% | 1,682 | 2,605 | +55% | 0 | 0 | — |
case-13 | fail→pass | 9,357 | 1,815 | -81% | 1 | 1 | 0% | 1,335 | 2,680 | +101% | 0 | 0 | — |
case-14 | fail→pass | 9,915 | 4,690 | -53% | 1 | 1 | 0% | 1,552 | 3,122 | +101% | 0 | 0 | — |
case-15 | fail→pass | 8,868 | 1,830 | -79% | 1 | 1 | 0% | 1,459 | 2,661 | +82% | 0 | 0 | — |
case-16 | fail→pass | 5,678 | 2,025 | -64% | 1 | 1 | 0% | 874 | 2,720 | +211% | 0 | 0 | — |
case-17 | fail→pass | 4,144 | 1,627 | -61% | 1 | 1 | 0% | 547 | 2,593 | +374% | 0 | 0 | — |
case-18 | pass→pass | 12,130 | 2,975 | -75% | 1 | 1 | 0% | 1,861 | 2,834 | +52% | 0 | 0 | — |
case-19 | pass→pass | 7,031 | 2,937 | -58% | 1 | 1 | 0% | 963 | 2,893 | +200% | 0 | 0 | — |
case-20 | fail→fail | 9,501 | 2,105 | -78% | 1 | 1 | 0% | 1,393 | 2,707 | +94% | 0 | 0 | — |
case-21 | pass→pass | 9,790 | 4,236 | -57% | 1 | 1 | 0% | 1,403 | 3,080 | +120% | 0 | 0 | — |
case-22 | pass→pass | 12,355 | 2,897 | -77% | 1 | 1 | 0% | 1,064 | 2,835 | +166% | 0 | 0 | — |
case-23 | pass→pass | 10,814 | 9,594 | -11% | 1 | 1 | 0% | 1,779 | 3,917 | +120% | 0 | 0 | — |
case-24 | pass→pass | 8,819 | 8,631 | -2% | 1 | 1 | 0% | 1,537 | 3,908 | +154% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 24 cases were attempted. The headline lift of +58 percentage points is the difference between those two pass rates over the 24 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.