Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Generate a structured response-to-referees document from a referee report and the revised manuscript. Maps each referee comment to the specific revision, classifies coverage (addressed / partially / deferred / disagreement), and drafts polite but firm responses. Use during the R&R (revise-and-resubmit) stage of paper revision.
.claude/skills/pedrohcgs-respond-to-referees/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-07 | ✗→✓ | ▲ Improved | 94% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 61% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 72% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 92% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 91% | 0% |
Produce a complete response-to-referees document by cross-referencing the referee report against the revised manuscript. Classify every concern, draft a courteous response for each, and flag anything unaddressed before submission.
$0 — path to the referee report$1 — path to the revised manuscriptSupported formats and how to read them. In the commands below, FILE stands for the input path being converted — either $0 (referee report) or $1 (revised manuscript). Always use mktemp for the temp file (not a predictable /tmp/... name) so paths with spaces and concurrent runs don't collide, and so untrusted FILE paths can't clobber other temp files via symlink races.
| Format | How to extract text | | --- | --- | | .tex, .qmd, .md, .txt | Read directly with the Read tool. | | .pdf | TMP=$(mktemp --suffix=.txt) && pdftotext "FILE" "$TMP" (poppler-utils; use mktemp -t ... on macOS if --suffix is unsupported). Grep "$TMP". | | .docx | TMP=$(mktemp --suffix=.txt) && pandoc "FILE" -t plain -o "$TMP" (or docx2txt "FILE" "$TMP"). Grep "$TMP". | | .html | TMP=$(mktemp --suffix=.txt) && pandoc "FILE" -t plain -o "$TMP". Grep "$TMP". |
If a required tool is missing or extraction fails, ask the user to provide a plain-text version (.txt or .md) and stop.
Before any parsing or grep, convert non-text inputs (.pdf, .docx, .html) to plain text using the table above. Keep both the temp text file (for grep) and the original (for citation page references).
For every concern:
Grep the plain-text version of the revised manuscript for those terms (and synonyms). Note: Grep only works on text — if the original was any non-text format (for example, .pdf, .docx, or .html), grep the converted temp file from Step 0.Read the surrounding context (±20 lines) to confirm the change addresses the concern.Assign one of four labels to each concern:
| Label | Meaning | | --- | --- | | Addressed | The revision directly resolves the concern with a specific change you can point to. | | Partially addressed | The revision moves in the requested direction but does not fully resolve the concern (e.g., added one robustness check when the referee asked for two). | | Deferred | The revision does not change the manuscript on this point but the response will explain why (out of scope, separate paper, conflicting referee, etc.). | | Disagreement | The author respectfully disagrees with the referee's premise. The response will explain the reasoning and any compromise. |
If you cannot find any evidence of a revision OR a deliberate decision to defer/disagree, mark the concern UNADDRESSED — REQUIRES AUTHOR INPUT and surface it in the warning summary at the end.
For every concern, write a 3–6 sentence response in this structure:
Tone conventions: courteous but firm; never defensive; never quote the referee back at length; use "we" for the author team; avoid "the referee is wrong" — prefer "we respectfully retain our original framing because…".
Write the output to response-to-referees.md (matching the template filename) or a path the user specifies. Use the structure in templates/response-to-referees.md:
The response document's most hallucination-prone content is the set of "we added X on page Y" claims. Hallucinating these gets a paper desk-rejected on sight. Before declaring the response document final, run the Post-Flight Verification protocol from .claude/rules/post-flight-verification.md.
Steps:
claim-verifier via the Agent tool with subagent_type=claim-verifier and context=fork. Hand it: the claims table, the verification questions, the path to the revised manuscript. Do NOT include the response draft.Downgrade to the classification the evidence supports:
AddressedPartially addressed or Deferred with an honest rationaleOpt-out: --no-verify flag. Not recommended — the referee will run this check themselves.
After the document is written, include this summary in your final chat message to the user (NOT inside the response document):
## Unaddressed concerns requiring author input
- R1.3: [summary] — no evidence of revision found
- R2.7: [summary] — flagged as deferred but no rationale yet draftedIf everything is covered, the final message should say All concerns addressed or explicitly classified.
response-to-referees.md — the deliverable (filename matches templates/response-to-referees.md)response-to-referees-matrix.csv — machine-readable concern-to-response mapping for tracking across revisionsTip. Before drafting your response, consider running /review-paper --peer --r2 <journal> on the revised manuscript first. It simulates the next referee round against your revisions — catching the "Resolved / Partial / Not addressed" classification mistakes before the real referee does. See .claude/skills/review-paper/SKILL.md.
/review-paper./slide-excellence (works on .tex manuscripts via the domain-reviewer agent).quality_reports/ if you want a permanent record alongside other quality reports.Before reporting completion:
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 54,739 | 36,303 | -34% | 1 | 1 | 0% | 4,536 | 2,662 | -41% | 0 | 0 | — |
case-02 | fail→fail | 46,033 | 36,593 | -21% | 1 | 1 | 0% | 7,790 | 2,773 | -64% | 0 | 0 | — |
case-08 | pass→pass | 11,436 | 5,785 | -49% | 1 | 1 | 0% | 2,005 | 3,204 | +60% | 0 | 0 | — |
case-03 | fail→fail | 44,467 | 5,004 | -89% | 1 | 1 | 0% | 6,029 | 2,413 | -60% | 0 | 0 | — |
case-04 | fail→fail | 11,173 | 37,387 | +235% | 1 | 1 | 0% | 1,263 | 2,455 | +94% | 0 | 0 | — |
case-05 | pass→fail | 8,847 | 5,028 | -43% | 1 | 1 | 0% | 1,302 | 2,444 | +88% | 0 | 0 | — |
case-06 | pass→pass | 6,520 | 11,038 | +69% | 1 | 1 | 0% | 1,062 | 3,338 | +214% | 0 | 0 | — |
case-07 | fail→pass | 9,729 | 5,554 | -43% | 1 | 1 | 0% | 1,628 | 3,162 | +94% | 0 | 0 | — |
case-09 | fail→pass | 9,946 | 4,395 | -56% | 1 | 1 | 0% | 1,797 | 2,892 | +61% | 0 | 0 | — |
case-10 | fail→fail | 12,815 | 5,167 | -60% | 1 | 1 | 0% | 2,205 | 3,027 | +37% | 0 | 0 | — |
case-11 | fail→pass | 11,662 | 8,560 | -27% | 1 | 1 | 0% | 2,162 | 3,729 | +72% | 0 | 0 | — |
case-12 | fail→pass | 8,239 | 3,074 | -63% | 1 | 1 | 0% | 1,362 | 2,611 | +92% | 0 | 0 | — |
case-13 | fail→pass | 7,991 | 2,879 | -64% | 1 | 1 | 0% | 1,390 | 2,648 | +91% | 0 | 0 | — |
case-14 | fail→pass | 15,658 | 3,621 | -77% | 1 | 1 | 0% | 1,995 | 2,933 | +47% | 0 | 0 | — |
case-15 | fail→pass | 9,538 | 6,678 | -30% | 1 | 1 | 0% | 1,643 | 3,279 | +100% | 0 | 0 | — |
case-16 | pass→pass | 13,969 | 11,028 | -21% | 1 | 1 | 0% | 2,312 | 3,878 | +68% | 0 | 0 | — |
case-17 | fail→pass | 13,717 | 6,356 | -54% | 1 | 1 | 0% | 2,362 | 3,304 | +40% | 0 | 0 | — |
case-18 | fail→fail | 7,573 | 2,855 | -62% | 1 | 1 | 0% | 1,234 | 2,649 | +115% | 0 | 0 | — |
case-19 | fail→pass | 16,744 | 2,013 | -88% | 1 | 1 | 0% | 3,119 | 2,504 | -20% | 0 | 0 | — |
case-20 | fail→pass | 14,644 | 3,418 | -77% | 1 | 1 | 0% | 1,295 | 2,794 | +116% | 0 | 0 | — |
case-21 | fail→pass | 10,224 | 3,462 | -66% | 1 | 1 | 0% | 1,550 | 2,794 | +80% | 0 | 0 | — |
case-22 | fail→pass | 11,810 | 16,244 | +38% | 1 | 1 | 0% | 1,875 | 3,004 | +60% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 17 counted toward the lift figure. The other 5 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +50 percentage points is the difference between those two pass rates over the 17 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 8/13/2026 | +40% |
Other measured skills in the registry, with their headline benchmark lift.