Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when planning or drafting the CVPR one-page rebuttal after reviews are released, covering the official rebuttal template, the ban on new contributions and external links, triaging multiple reviews at 16k-submission scale, choosing which numbers fit in one page, and writing for the AC who reads the discussion.
.claude/skills/brycewang-stanford-cvpr-author-response/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 22% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 10% | 0% |
| case-03 | ✗→✓ | ▲ Improved | -9% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 67% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 39% | 0% |
Use this when CVPR reviews land. In the 2026 cycle, reviews reached authors on January 22 and the rebuttal closed January 29, with reviewer–AC discussion running January 30 – February 5 and decisions on February 20. One week, one page — the format is the strategy.
The rebuttal is an optional, single-page PDF built on the rebuttal.tex template in the official author kit. Two rules from the Author Guidelines carry teeth:
technically fit because margins or fonts were squeezed — altered formatting counts as over-length.
is explicitly not for new contributions or unrequested new experiments, reviewers are told not to demand significant new experiments, and it must stay anonymous with no links to external material (no anonymous repos, no video URLs, no project pages).
That last clause surprises people arriving from venues with revision uploads: at CVPR you cannot attach anything. If a number matters, it goes in the page as a sentence or a five-row table.
With ~16,000 papers in the 2026 pool, your reviewers wrote many reviews and your AC is tracking dozens of papers. The rebuttal's real reader is the AC during the discussion window. Triage accordingly:
| Review signal | Rebuttal move | Budget | |---|---|---| | Factual error ("no ablation on X" when Table 4 has it) | Quote the reviewer, cite the exact table/line, one sentence | 2–3 lines each | | Misread contribution / scope | Restate the claim in one crisp sentence, point to where the paper says it | 4–6 lines | | Requested clarification (protocol, metric, split) | Answer directly; give the number | 3–5 lines | | Requested small experiment | Run it if it fits the week; report as a mini-table | 6–10 lines | | Demand for a full new benchmark or method variant | Decline politely; explain why out of scope for a rebuttal | 2–3 lines | | Subjective "not novel enough" | Delta-list against the 2–3 closest citations, no adjectives | 4–6 lines |
Answer every reviewer, but not every sentence of every reviewer. Ignoring one review entirely reads as concession in the discussion phase.
latex% rebuttal.tex skeleton — headings are navigation for the AC \section*{Shared concerns (R1, R3): fairness of the baseline comparison} % one paragraph + mini-table with matched-backbone numbers \section*{R1: reported speed} % cite Sec. 4.2 line refs; give ms/frame on the named GPU from the CRF hardware \section*{R2: missing citation [Foo 2025]} % concede, state the delta in one sentence, promise camera-ready citation \section*{R3: qualitative failures} % point to supplement Sec. C; summarize the two failure modes in text
Group shared objections once instead of three times; name reviewers in headings; lead with the concern you can kill with a fact, because the first inches of the page set the AC's prior about who is being careful.
the contested points.
happens without you in the room; give allies quotable sentences.
discussion; "significantly better" does not.
Draft rebuttal prose tends to argue; shipping rebuttal prose demonstrates. Compare:
> Before: "We respectfully disagree with R2's claim that our comparison is unfair. > As explained in the paper, our method uses standard settings, and it is common > practice in the field to compare this way. We believe the improvement is clearly > significant."
> After: "R2: 'the baseline uses a weaker backbone.' Both rows of Table 2 use > ResNet-50 with the schedule from 24] (Sec. 4.1, L412–418). Under R2's suggested > ViT-B setting, the gap is +1.9 AP (mini-table below), consistent with the paper's > +2.1."
The rewrite quotes the objection, cites the paper by line, gives a number, and never characterizes its own results. Every paragraph on the page should survive the same transformation: objection verbatim → paper location → number.
the first-read emotional register writes bad rebuttals.
fixability tag: factual / clarification / small-run / out-of-scope.
much of the week the paper deserves versus the resubmission draft.
cvpr-workflow and set the internal freeze 24 hours beforethe OpenReview deadline.
The rebuttal cannot rescue a paper whose evidence is missing, and reviewers are instructed not to expect new experiments. If every review converges on the same absent comparison, the plan is the next cycle (or ICCV/ECCV), not one page of protest. Spend the week where movement is possible: factual corrections, misreadings, and small requested numbers.
formats at CVF venues before.
text[Rebuttal plan] file / skip (reason) [Shared concerns] <grouped items with reviewer IDs> [Factual corrections] <claim → paper location> [Numbers to include] <experiment → mini-table budget> [Declined requests] <item → one-line rationale> [Page budget] <lines allocated per section, total ≤ 1 page>
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 21,584 | 18,825 | -13% | 1 | 1 | 0% | 2,912 | 3,566 | +22% | 0 | 0 | — |
case-02 | fail→pass | 33,368 | 32,127 | -4% | 1 | 1 | 0% | 5,024 | 5,508 | +10% | 0 | 0 | — |
case-03 | fail→pass | 22,503 | 8,665 | -61% | 1 | 1 | 0% | 3,388 | 3,094 | -9% | 0 | 0 | — |
case-04 | fail→fail | 11,968 | 17,873 | +49% | 1 | 1 | 0% | 1,986 | 3,532 | +78% | 0 | 0 | — |
case-05 | pass→pass | 18,788 | 13,196 | -30% | 1 | 1 | 0% | 2,246 | 3,481 | +55% | 0 | 0 | — |
case-06 | pass→pass | 18,842 | 18,291 | -3% | 1 | 1 | 0% | 2,049 | 3,215 | +57% | 0 | 0 | — |
case-07 | pass→pass | 17,375 | 9,439 | -46% | 1 | 1 | 0% | 1,866 | 2,973 | +59% | 0 | 0 | — |
case-08 | fail→pass | 11,877 | 15,687 | +32% | 1 | 1 | 0% | 2,044 | 3,419 | +67% | 0 | 0 | — |
case-09 | fail→pass | 21,929 | 17,900 | -18% | 1 | 1 | 0% | 2,591 | 3,601 | +39% | 0 | 0 | — |
case-10 | fail→pass | 15,066 | 13,813 | -8% | 1 | 1 | 0% | 1,643 | 2,886 | +76% | 0 | 0 | — |
case-11 | fail→pass | 23,110 | 10,561 | -54% | 1 | 1 | 0% | 1,595 | 2,843 | +78% | 0 | 0 | — |
case-12 | pass→pass | 11,507 | 13,343 | +16% | 1 | 1 | 0% | 1,831 | 2,899 | +58% | 0 | 0 | — |
case-13 | fail→pass | 16,098 | 14,180 | -12% | 1 | 1 | 0% | 1,772 | 2,873 | +62% | 0 | 0 | — |
case-14 | pass→pass | 18,550 | 8,151 | -56% | 1 | 1 | 0% | 2,108 | 2,851 | +35% | 0 | 0 | — |
case-15 | pass→pass | 20,829 | 19,160 | -8% | 1 | 1 | 0% | 2,308 | 3,750 | +62% | 0 | 0 | — |
case-16 | fail→pass | 12,119 | 9,026 | -26% | 1 | 1 | 0% | 1,771 | 2,663 | +50% | 0 | 0 | — |
case-17 | pass→pass | 11,898 | 13,268 | +12% | 1 | 1 | 0% | 1,681 | 2,797 | +66% | 0 | 0 | — |
case-18 | fail→pass | 17,328 | 13,449 | -22% | 1 | 1 | 0% | 1,675 | 2,523 | +51% | 0 | 0 | — |
case-19 | fail→fail | 16,204 | 12,712 | -22% | 1 | 1 | 0% | 1,897 | 2,861 | +51% | 0 | 0 | — |
case-20 | fail→fail | 20,095 | 15,117 | -25% | 1 | 1 | 0% | 2,653 | 3,417 | +29% | 0 | 0 | — |
case-21 | fail→pass | 15,216 | 16,076 | +6% | 1 | 1 | 0% | 2,077 | 2,971 | +43% | 0 | 0 | — |
case-22 | fail→pass | 22,051 | 21,821 | -1% | 1 | 1 | 0% | 2,264 | 3,922 | +73% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted. The headline lift of +55 percentage points is the difference between those two pass rates over the 22 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.