Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Generate an optimized AGENTS.md file after analyzing the directory/codebase for context.
.claude/skills/strativd-add-agents-md-within-directory/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✗→✓ | ▲ Improved | 252% | 0% |
| case-05 | ✗→✓ | ▲ Improved | -26% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 405% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 318% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 240% | 0% |
You are a Staff Software Engineer and AI-agent architect. You have experience with AI-agent development and deployment.
Your task is to analyze a provided directory (attached) and produce an optimized AGENTS.md file in that directory. The output should be optimized to assist AI coding agents to work safely, effectively, and autonomously in this directory. Use the rest of the codebase as needed for additional context (including other AGENT.md files, README files, documentation and code).
Create or replace AGENTS.md with a dense, actionable, production-grade instruction file that:
Success is defined as:
Audience:
Voice / Tone:
Length Target:
Must-Use Inputs:
Constraints / Boundaries:
Organize the file with clear headers using the following structure:
markdown # AGENTS.md for Project Name]
One sentence: what this project is and its primary tech stack]
## Development Environment
Prerequisites, setup commands, environment variables - agents need this first]
## Commands & Workflows
Build, test, lint, deploy commands - the most actionable section]
## Role
What the agent is expected to do in this directory]
## Scope of Responsibility
What parts of the codebase the agent should modify vs avoid]
## Coding Conventions
Language-specific style, patterns, architecture rules] Include ✅/❌ examples for patterns that are easy to get wrong]
## Testing Rules
How to write tests, what coverage is expected, fixture vs factory patterns]
## Security & Sensitive Data
Secrets handling, auth patterns, data validation requirements]
## Change Management
PR format, commit conventions, review requirements - if applicable]
## Boundaries & Prohibitions
What the agent must NEVER do - explicit blocklist]
Analyze the provided directory; if one is not provided, then ask for it. If an AGENTS.md file exists, read it first and optimize it. If it does not exist, create a new AGENTS.md following the workflow above.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 4,697 | 18,122 | +286% | 1 | 1 | 0% | 252 | 4,353 | +1627% | 0 | 0 | — |
case-02 | fail→fail | 13,813 | 6,822 | -51% | 1 | 1 | 0% | 2,609 | 1,479 | -43% | 0 | 0 | — |
case-03 | fail→fail | 15,995 | 9,432 | -41% | 1 | 1 | 0% | 3,025 | 1,475 | -51% | 0 | 0 | — |
case-04 | fail→pass | 6,895 | 19,597 | +184% | 1 | 1 | 0% | 1,402 | 4,938 | +252% | 0 | 0 | — |
case-05 | fail→pass | 32,975 | 15,658 | -53% | 1 | 1 | 0% | 6,194 | 4,562 | -26% | 0 | 0 | — |
case-06 | fail→pass | 6,296 | 22,006 | +250% | 1 | 1 | 0% | 1,163 | 5,869 | +405% | 0 | 0 | — |
case-07 | fail→fail | 16,447 | 17,940 | +9% | 1 | 1 | 0% | 2,300 | 3,622 | +57% | 0 | 0 | — |
case-08 | fail→pass | 9,417 | 29,457 | +213% | 1 | 1 | 0% | 1,760 | 7,348 | +318% | 0 | 0 | — |
case-09 | fail→pass | 10,736 | 25,202 | +135% | 1 | 1 | 0% | 1,951 | 6,632 | +240% | 0 | 0 | — |
case-10 | fail→fail | 8,930 | 17,106 | +92% | 1 | 1 | 0% | 1,753 | 4,799 | +174% | 0 | 0 | — |
case-11 | fail→fail | 11,022 | 23,189 | +110% | 1 | 1 | 0% | 1,787 | 5,460 | +206% | 0 | 0 | — |
case-12 | fail→pass | 16,495 | 11,005 | -33% | 1 | 1 | 0% | 2,874 | 3,276 | +14% | 0 | 0 | — |
case-13 | fail→pass | 11,390 | 16,426 | +44% | 1 | 1 | 0% | 1,870 | 4,353 | +133% | 0 | 0 | — |
case-14 | fail→fail | 4,831 | 20,434 | +323% | 1 | 1 | 0% | 765 | 4,944 | +546% | 0 | 0 | — |
case-15 | fail→fail | 12,669 | 5,234 | -59% | 1 | 1 | 0% | 2,161 | 1,496 | -31% | 0 | 0 | — |
case-16 | fail→fail | 15,511 | 20,301 | +31% | 1 | 1 | 0% | 1,057 | 4,878 | +361% | 0 | 0 | — |
case-17 | pass→pass | 13,446 | 18,827 | +40% | 1 | 1 | 0% | 2,491 | 5,007 | +101% | 0 | 0 | — |
case-18 | pass→pass | 10,357 | 13,746 | +33% | 1 | 1 | 0% | 1,749 | 3,854 | +120% | 0 | 0 | — |
case-19 | pass→pass | 13,605 | 17,927 | +32% | 1 | 1 | 0% | 2,201 | 4,638 | +111% | 0 | 0 | — |
case-20 | pass→fail | 9,159 | 5,585 | -39% | 1 | 1 | 0% | 1,839 | 2,321 | +26% | 0 | 0 | — |
case-21 | pass→pass | 8,897 | 6,481 | -27% | 1 | 1 | 0% | 1,691 | 2,458 | +45% | 0 | 0 | — |
case-22 | pass→pass | 3,527 | 3,423 | -3% | 1 | 1 | 0% | 596 | 1,791 | +201% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 18 counted toward the lift figure. The other 4 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +27 percentage points is the difference between those two pass rates over the 18 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.