Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when drafting the CoRL rebuttal — a single-page PDF due days after reviews are released, aimed at reviewers, the Area Chair, and the discussion window that follows. Covers triage against the first-round gate, one-page layout economy, presenting new numbers compactly, and tone for an exchange that becomes public on acceptance.
.claude/skills/brycewang-stanford-corl-author-response/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-22 | ✗→✓ | ▲ Improved | 62% | 0% |
| case-01 | ✗→✓ | ▲ Improved | 45% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 10% | 0% |
| case-10 | ✗→✓ | ▲ Improved | 38% | 0% |
| case-16 | ✗→✓ | ▲ Improved | 71% | 0% |
CoRL's rebuttal is a one-page PDF, not a text box — a format that rewards preparation and layout discipline over volume. In the 2026 cycle it was due August 11, 11:59 PM AoE, with reviewer–AC discussion following August 12–19 (corl.org rebuttal instructions, read 2026-07-08). The window between reviews arriving and the deadline is measured in days; the winning move is to have experiments and a compiled skeleton ready before reviews exist (see corl-workflow).
CoRL 2026 applied a first-round rejection rule — no weak-accept-or-above score from any reviewer or the AC makes the paper a candidate for rejection without rebuttal. So the opening triage is not "what do I answer?" but "who is my advocate?": identify the most positive scorer, protect their reasons for optimism, and target the specific objections that keep others below threshold.
A page holds roughly 700–900 words in the template's footprint, less once you add a results table. Spend it by expected score movement:
| Budget slice | Content | Why it earns its space | |---|---|---| | ~15% | Header: thanks (one line) + a 2–3 line summary of what the rebuttal adds | ACs skim this first; make the additions enumerable | | ~50% | The 2–3 concerns the AC flagged or multiple reviewers share | Core-concern focus is what the official instructions ask for | | ~25% | One compact table/figure of new numbers (seeds, episodes, baseline) | Evidence moves scores; prose reassurance does not | | ~10% | Per-reviewer one-liners for smaller points, keyed as R1/R2/R3 | Shows nothing was ducked |
What does not fit on the page: re-explaining the paper, quoting long review passages back, answering every minor point at equal depth, or promising future work as a substitute for a number you could have produced.
New robot-learning evidence compresses well if you design for it:
latex% One-page rebuttal skeleton (compile with the venue-permitted format) \section*{Response to Reviews — Paper \#XXXX} We thank the reviewers. This response adds: (i) 5-seed results on all tasks (R1, R3), (ii) a real-robot spot-check of the sim claim (R2), (iii) the missing diffusion-policy baseline (R1). \paragraph{Seeds and episodes (R1.W1, R3.W2).} \begin{tabular}{lccc} Task suite & Ours & Ours (5 seeds) & DP baseline \\ Kitchen-6 & 78\% & 76.4 $\pm$ 2.1\% & 61.2 $\pm$ 3.0\% \\ ... \end{tabular} % cite episode counts inline: "each cell = 5 seeds x 50 eval episodes" \paragraph{Sim-to-real scope (R2.W1).} We ran task A on hardware: 64\% over 25 trials (video frames in Fig.~1); we will scope the claim to tasks A--B in revision and state the transfer gap explicitly.
Rules of thumb: every number states its seeds × episodes provenance inline; every promised revision names the section it will change; anything that cannot be shown in one table is summarized with its single headline number.
dispersion and whether conclusions changed. This is the cheapest score-mover at this venue.
feasible; if genuinely infeasible in the window, explain the concrete blocker (hardware access, closed weights) and give the nearest fair proxy.
the smallest real-robot result that preserves the paper's core; a scoped-down true claim beats a defended overclaim in the discussion window.
will add; this section is mandatory at CoRL and reviewers audit it.
flat in tone; flag persistently wrong low-confidence reviews to the AC via a confidential comment, not in the shared page.
warrants it — so check OpenReview daily; a reviewer follow-up unanswered for three days reads as concession.
("see Table 1 of our response, row 2").
a final short summary comment ("changes we commit to in revision: …") gives them a clean artifact to quote.
Reviews and rebuttals of accepted CoRL papers are made publicly available. Write the page so you would be comfortable with it linked from the paper forever: no sarcasm, no reviewer-blaming, no overpromising. Concessions phrased as improvements ("we agree, and the revision will…") age well in public.
text[ ] At least one gate-passing score exists (else: resubmission branch) [ ] Page compiles to exactly 1 page, venue-permitted format [ ] AC-flagged / shared concerns get the majority of the space [ ] Every new number carries seeds x episodes provenance [ ] Every commitment names its revision location [ ] Anonymity preserved (no identity, no identifying links) [ ] Submitted well before Aug 11 23:59 AoE; calendar alert for the Aug 12-19 discussion window
Rebuttal format, deadlines, and the gate rule are re-set annually — reconfirm at https://www.corl.org/contributions/instruction-for-rebuttal for the live cycle.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-22 | fail→pass | 21,255 | 17,132 | -19% | 1 | 1 | 0% | 1,982 | 3,210 | +62% | 0 | 0 | — |
case-01 | fail→pass | 32,198 | 31,478 | -2% | 1 | 1 | 0% | 3,868 | 5,592 | +45% | 0 | 0 | — |
case-02 | fail→fail | 29,938 | 32,067 | +7% | 1 | 1 | 0% | 4,068 | 6,134 | +51% | 0 | 0 | — |
case-09 | pass→pass | 16,144 | 15,180 | -6% | 1 | 1 | 0% | 2,498 | 3,121 | +25% | 0 | 0 | — |
case-03 | fail→pass | 40,033 | 25,799 | -36% | 1 | 1 | 0% | 4,936 | 5,445 | +10% | 0 | 0 | — |
case-04 | pass→pass | 20,957 | 18,741 | -11% | 1 | 1 | 0% | 2,484 | 3,635 | +46% | 0 | 0 | — |
case-05 | pass→pass | 21,995 | 22,552 | +3% | 1 | 1 | 0% | 2,565 | 4,356 | +70% | 0 | 0 | — |
case-06 | pass→pass | 25,167 | 23,153 | -8% | 1 | 1 | 0% | 3,337 | 4,399 | +32% | 0 | 0 | — |
case-07 | pass→pass | 21,603 | 22,827 | +6% | 1 | 1 | 0% | 2,511 | 4,028 | +60% | 0 | 0 | — |
case-08 | fail→fail | 19,905 | 10,598 | -47% | 1 | 1 | 0% | 2,207 | 3,060 | +39% | 0 | 0 | — |
case-10 | fail→pass | 20,339 | 18,808 | -8% | 1 | 1 | 0% | 2,535 | 3,500 | +38% | 0 | 0 | — |
case-11 | pass→pass | 17,978 | 18,091 | +1% | 1 | 1 | 0% | 1,981 | 3,163 | +60% | 0 | 0 | — |
case-12 | pass→pass | 18,734 | 15,035 | -20% | 1 | 1 | 0% | 2,040 | 3,687 | +81% | 0 | 0 | — |
case-13 | pass→pass | 17,858 | 15,856 | -11% | 1 | 1 | 0% | 1,948 | 3,524 | +81% | 0 | 0 | — |
case-14 | fail→fail | 12,305 | 18,460 | +50% | 1 | 1 | 0% | 1,832 | 3,031 | +65% | 0 | 0 | — |
case-15 | fail→fail | 13,681 | 13,851 | +1% | 1 | 1 | 0% | 1,841 | 3,072 | +67% | 0 | 0 | — |
case-16 | fail→pass | 21,639 | 24,572 | +14% | 1 | 1 | 0% | 2,202 | 3,762 | +71% | 0 | 0 | — |
case-17 | fail→pass | 17,803 | 5,721 | -68% | 1 | 1 | 0% | 1,603 | 2,342 | +46% | 0 | 0 | — |
case-18 | pass→pass | 24,251 | 18,556 | -23% | 1 | 1 | 0% | 2,292 | 3,815 | +66% | 0 | 0 | — |
case-19 | pass→pass | 21,049 | 20,441 | -3% | 1 | 1 | 0% | 2,484 | 3,873 | +56% | 0 | 0 | — |
case-20 | pass→pass | 23,651 | 23,969 | +1% | 1 | 1 | 0% | 2,372 | 4,038 | +70% | 0 | 0 | — |
case-21 | fail→pass | 21,238 | 18,867 | -11% | 1 | 1 | 0% | 2,011 | 3,234 | +61% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +32 percentage points is the difference between those two pass rates over the 22 comparable cases. 2 cases got worse with the skill loaded, and they are included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.