Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when organizing appendices and supplementary material for an ACL paper under ACL Rolling Review, covering the mandatory Limitations and optional ethics sections, appendices after references, anonymized software and data archives, the no-cloud-links rule, and deciding what must stay in the 8-page or 4-page body.
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | -8% | 0% |
| case-03 | ✗→✓ | ▲ Improved | -9% | 0% |
| case-14 | ✗→✓ | ▲ Improved | 26% | 0% |
| case-16 | ✗→✓ | ▲ Improved | 17% | 0% |
| case-05 | ✓→✓ | = Same ✓ | 29% | 0% |
Use this when splitting an ACL paper between body, appendix, and archive. The governing ARR principle: reviewers are not required to consider material in appendices or supplements, so anything decision-critical that lives only there is effectively invisible.
text[ content pages: 8 long / 4 short ] <- the reviewed argument lives here [ Limitations (REQUIRED, unlimited) ] <- after conclusion, outside page count [ Ethics statement (optional) ] [ References (unlimited) ] [ Appendices (unlimited, same PDF) ] <- optional reading for reviewers + separate .tgz/.zip archive <- software / data supplement
Missing Limitations is a desk-reject condition; treating it as one throwaway sentence is a review-stage penalty even when it passes the gate.
reads as not having one.
| Weak pattern | Stronger ACL pattern | |---|---| | "Results may not generalize" | Name the languages, domains, and model scales actually tested and the nearest untested regime | | "LLMs can hallucinate" | State which conclusions depend on a specific model snapshot and API behavior | | Silent on data | Note license constraints, demographic skew, or collection-window bias in the corpora used | | Written last-minute | Mirrors the risks reviewers will find anyway, defusing them on your terms |
ACL's policy explicitly instructs reviewers not to punish honest limitations, which makes this section the cheapest goodwill in the whole submission.
cloud-storage links are barred, and any external page must be anonymous and untracked.
usernames, license headers, README contact lines.
what maps to which table, and nothing should require credentials just to read.
compute (see acl-reproducibility).
A long paper introduces a retrieval-augmented QA method with results on six benchmarks in three languages. Body: method figure, main table (six benchmarks averaged + per-language block), two-paragraph error analysis, one ablation that carries the mechanism claim. Appendix: full per-benchmark tables, prompts, retrieval index details, remaining ablations, annotation guidelines for the human study. Archive: code, prompts as files, and all model outputs. The test: a reviewer who never scrolls past the references can still reconstruct and believe every claim in the abstract.
every averaged number in the body.
body at least once (unreferenced appendix content is invisible).
Number tables and figures continuously with the body so the author response can cite "Table 9" unambiguously during the discussion phase.
Write it when the paper involves human subjects or annotators, scraped user-generated content, demographic inference, dual-use capability, or release of models/data with realistic misuse paths. Skip it when nothing applies — the Responsible NLP checklist already covers the routine cases, and a padded statement invites the very scrutiny it fails to answer. It shares the unlimited space after the conclusion with Limitations.
reviewers abandon multi-gigabyte supplements unopened.
paths instead of shipping secrets.
is a different codebase every cycle.
a synthetic sample with identical schema so scripts still run.
text[Split status] sound / body-overloaded / appendix-dependent [Limitations quality] substantive / ritual / missing [Must-move-up] <decision-critical items currently below the fold> [Archive check] <format/anonymity/clean-machine findings> [Reviewer-blind spots] <claims visible only outside the body>
Other measured skills in the registry, with their headline benchmark lift.