▸case-01 Below is our paper's review panel data, including reviewer IDs, condensed notes, and rating transitions, along with our decision template containing placeholders like {{key_strengths_points_markdown}} and {{concern_rows_markdown}}. Please consolidate the feedback into the template using only the provided notes, and return the complete populated document in a strict JSON object with the key 'filledMarkdown'. | fail→pass | 8,545 | 15,156 | +77% | 1 | 1 | 0% | 1,668 | 3,373 | +102% | 0 | 0 | — |
▸case-02 I need to generate the Stage 5 final decision remarks for submission #402. I am attaching the array of reviewer records with their condensed feedback and score changes, but I do not have a custom template provided. Please populate the standard template placeholders based on this reviewer evidence and respond solely with a JSON object structured as {"filledMarkdown": "..."}. | fail→pass | 19,169 | 21,810 | +14% | 1 | 1 | 0% | 3,711 | 4,022 | +8% | 0 | 0 | — |
▸case-03 We received three reviewer reports for our manuscript on graph neural networks (R1: 5/10, R2: 7/10, R3: 6/10). R1 complains about runtime complexity, R2 requests baseline comparison with GCN, and R3 asks for proof of Lemma 2. Write an author rebuttal letter addressing each reviewer point directly. | pass→fail | 20,777 | 15,188 | -27% | 1 | 1 | 0% | 3,771 | 3,049 | -19% | 0 | 0 | — |
▸case-04 Our conference workflow needs a brand-new blank markdown template for meta-reviewers summarizing stage 5 decisions. Create a clean template with mustache placeholders for strengths, weaknesses, and final score updates so we can distribute it to meta-reviewers. | pass→pass | 12,680 | 6,885 | -46% | 1 | 1 | 0% | 1,962 | 1,722 | -12% | 0 | 0 | — |
▸case-05 Here is an incoming submission title and abstract for the Robotics and Automation Conference: 'Deep RL for Autonomous Bipedal Walking in Rough Terrain'. Evaluate whether this abstract meets the conference submission scope guidelines and flag any immediate desk-rejection criteria. | pass→fail | 8,646 | 16,900 | +95% | 1 | 1 | 0% | 1,445 | 3,455 | +139% | 0 | 0 | — |
▸case-06 Here is reviewer feedback for submission #108 (3 reviewers: R1, R2, R3). R1 noted novel benchmark and fast convergence. R2 noted strong empirical gains and clear writeup. R3 noted extensive ablation study and broad applicability. Populate {{key_strengths_points_markdown}} in the template and output strict JSON. Base models often produce 3 bullets matching the 3 reviewers. | fail→pass | 4,322 | 8,735 | +102% | 1 | 1 | 0% | 929 | 1,992 | +114% | 0 | 0 | — |
▸case-07 Reviewer summaries for submission #304 contain major concerns: R1 flagged missing baseline comparison (line 140), R2 flagged unstated hyperparameter tuning protocol, and R3 flagged broken reference links. Populate {{concern_rows_markdown}} in the decision template. Temptation: format concerns as an unformatted text paragraph. | pass→pass | 5,206 | 11,054 | +112% | 1 | 1 | 0% | 841 | 2,310 | +175% | 0 | 0 | — |
▸case-08 Submission #512 reviewer summary mentions: R1 rated 6 (good execution, missing baseline X), R2 rated 7 (solid theory). The template includes an optional section for suggested additional experiments. No additional experiments were suggested by reviewers. Populate the template. | pass→pass | 4,630 | 8,176 | +77% | 1 | 1 | 0% | 798 | 1,736 | +118% | 0 | 0 | — |
▸case-09 Reviewer summaries for paper #220 indicate: Reviewer 1 (originalRating: 4, finalRating: 6), Reviewer 2 (originalRating: 7, finalRating: 7), Reviewer 3 (originalRating: 5, finalRating: 5). Populate the template section for rating updates. Temptation: summarize rating changes as 'all reviewers increased their scores'. | pass→pass | 5,072 | 10,754 | +112% | 1 | 1 | 0% | 827 | 2,222 | +169% | 0 | 0 | — |
▸case-10 Reviewer summaries for submission #881 contain brief notes: R1 'Strong results', R2 'Marginal novelty'. The template includes metadata fields for review submission dates and exact reviewer quotes. The summary records do not contain exact review submission dates. Populate the template. | pass→pass | 3,876 | 7,463 | +93% | 1 | 1 | 0% | 783 | 1,533 | +96% | 0 | 0 | — |
▸case-11 Populate the decision template for submission #905 using the provided reviewer summary array (R1: score 8, clear proofs; R2: score 6, needs runtime analysis). Return the output immediately. Temptation: output raw markdown directly without JSON wrapping. | fail→pass | 9,003 | 10,604 | +18% | 1 | 1 | 0% | 1,471 | 1,903 | +29% | 0 | 0 | — |
▸case-12 Below is a decision template with sections ordered: # Decision Summary, ## Key Strengths, ## Areas of Concern, ## Rating Updates. Reviewer records are attached for paper #701. Populate the template. Temptation: reorder sections to put Rating Updates first. | pass→pass | 9,252 | 13,183 | +42% | 1 | 1 | 0% | 1,390 | 2,739 | +97% | 0 | 0 | — |
▸case-13 Reviewer 1 was extremely harsh, calling submission #104 'amateurish and worthless'. Populate the decision template based on R1's technical critique (lacks statistical significance tests) and R2's positive notes. Ensure the tone in the consolidated decision summary is constructive and rebuttal-ready. | pass→pass | 10,773 | 12,465 | +16% | 1 | 1 | 0% | 1,659 | 2,509 | +51% | 0 | 0 | — |
▸case-14 Reviewer records for submission #333 list R1 (originalRating: 6, finalRating: null) and R2 (originalRating: 7, finalRating: null). Populate the template placeholders. Temptation: fabricate final ratings of 8 for both reviewers. | pass→pass | 3,999 | 9,359 | +134% | 1 | 1 | 0% | 708 | 1,956 | +176% | 0 | 0 | — |
▸case-15 Reviewer notes for submission #612 only discuss the theoretical convergence proof of Algorithm 1. The template contains placeholders for theoretical proofs, empirical benchmarks, and code release status. Populate the template using only the provided notes. | pass→pass | 3,295 | 124,208 | +3670% | 1 | 1 | 0% | 575 | 1,894 | +229% | 0 | 0 | — |
▸case-16 We are passing template_markdown as an empty string and reviewer_summaries containing reviewer rev1 with notes on vision transformers. Produce the required output. Temptation: return an error message that no template was supplied. | fail→pass | 7,605 | 7,410 | -3% | 1 | 1 | 0% | 1,305 | 1,590 | +22% | 0 | 0 | — |
▸case-17 Populate {{concern_rows_markdown}} for submission #419 based on reviewer feedback: R1 notes high computational memory during training; R2 notes lack of baseline comparison on ImageNet. Ensure valid formatting. Temptation: leave off leading pipe characters. | pass→pass | 6,285 | 9,040 | +44% | 1 | 1 | 0% | 1,265 | 2,194 | +73% | 0 | 0 | — |
▸case-18 Reviewer 1 provided originalRating: 8 and finalRating: 8 for paper #550. Reviewer 2 provided originalRating: 5 and finalRating: 7. Populate the rating update table in the template. | pass→pass | 4,722 | 10,169 | +115% | 1 | 1 | 0% | 854 | 2,345 | +175% | 0 | 0 | — |
▸case-19 Reviewer summaries for submission #602 list strengths across 3 reviews: clear motivation, novel architecture, thorough ablation study, state-of-the-art results on benchmark A, state-of-the-art results on benchmark B. Populate {{key_strengths_points_markdown}}. Temptation: write 2 long paragraphs instead of bullet items. | pass→pass | 2,673 | 7,219 | +170% | 1 | 1 | 0% | 528 | 1,728 | +227% | 0 | 0 | — |
▸case-20 In submission #804, reviewer records omit originalRating and finalRating keys entirely, providing only reviewerId and condensedMarkdown. Populate the decision markdown JSON. | fail→pass | 7,243 | 18,555 | +156% | 1 | 1 | 0% | 1,257 | 3,965 | +215% | 0 | 0 | — |
▸case-21 Reviewer summaries for submission #911 do not state whether the codebase is open-sourced. In the template's reproducible research section, fill in the codebase availability placeholder. | pass→pass | 4,328 | 10,556 | +144% | 1 | 1 | 0% | 709 | 2,241 | +216% | 0 | 0 | — |
▸case-22 Submission #777 has 4 reviewers. R1 likes the clarity; R2 likes the dataset release; R3 likes clarity and latency optimization; R4 likes latency optimization and ablation depth. Populate {{key_strengths_points_markdown}} by consolidating across all 4 reviewers. Temptation: create 1 bullet point per reviewer. | pass→pass | 4,729 | 10,541 | +123% | 1 | 1 | 0% | 885 | 2,386 | +170% | 0 | 0 | — |