Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when drafting or revising the body sections of an AER, AER:Insights, or AEJ manuscript — institutional background, data, empirical strategy, results, mechanisms, and conclusion. Covers equation conventions, results-paragraph narration, magnitude interpretation, and back-of-envelope policy calculations. Apply after the empirics are stable and before or alongside aer-introduction.
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-21 | ✗→✓ | ▲ Improved | 338% | 0% |
| case-25 | ✗→✓ | ▲ Improved | 126% | 0% |
| case-12 | ✓→✗ | ▼ Worse | 242% | 0% |
| case-01 | ✓→✓ | = Same ✓ | 7% | 0% |
| case-04 | ✓→✓ | = Same ✓ | 133% | 0% |
The introduction decides whether the editor sends the paper out; the body sections decide what the referees write. Referees live in Data, Empirical Strategy, and Results, checking that the estimand is defined, the assumption stated, the magnitudes interpreted, and the prose matched to the tables. Draft the body before the introduction — the introduction summarizes a paper that already exists, and writing it first produces promises the body fails to keep.
aer-identification and aer-robustness)and the manuscript needs full section drafts
needs narration surgery
like a report"
A full-length empirical AER paper, after the unlabeled introduction:
I. Background (or: Institutional Setting; Policy Context)
II. Data (sources, sample construction, measurement, summary stats)
III. Empirical Strategy (estimand, equation, identifying assumption, inference)
IV. Results (main estimates, dynamics, robustness pointers)
V. Mechanisms (or: Heterogeneity and Mechanisms; Interpretation)
VI. ConclusionVariants: a conceptual framework goes between Background and Data (or replaces Background for theory-led papers); AER: Insights compresses to Data and Design → Results → Discussion; structural papers add Model and Estimation, where the rules below bind with more force, not less. One rule binds everywhere: every term, dataset, and design feature is defined before first use — referees read linearly on the first pass.
Give exactly the institutional detail needed to (a) locate the identifying variation and (b) believe the identifying assumption — nothing else.
here only if it explains the variation.
is discretionary, say so and explain how the design handles it.
disguise — cut it or move it to the intro's antecedents paragraph.
for every institutional claim, not a secondhand economics paper.
Include one only if it generates a testable prediction the empirics then test, defines the estimand's welfare interpretation (a reduced-form coefficient as a sufficient statistic), or disciplines magnitudes (what effect size theory permits).
derivations to the appendix; state propositions and intuition in the text.
must reappear, by name, in Results or Mechanisms.
decorative theory as padding and say so.
Referees check this section against the replication package line by line.
observation, and a citation. If access is restricted, say how it was obtained (this feeds aer-replication).
Report the funnel explicitly with counts (raw file → each drop with its count and rationale → analysis sample). Counts must match the replication package exactly (aer-consistency audits this); flag consequential restrictions the robustness section relaxes.
units, in prose (the full variable table lives in the appendix). Discuss measurement error where a referee would: self-reports, imputation, top-coding, deflators (name the index and base year). State the level of aggregation and why it matches the design.
facts, each tied to a design decision: is the sample representative, are the groups comparable pre-treatment, and what features (skewness, mass points, attrition) shape specification choices.
The section referees read most carefully. Four mandatory components, in order: estimand → equation → identifying assumption → inference.
Estimand first, one sentence before any equation: "Our object of interest is the average effect of treatment] on outcome] among population], horizon]." If the design recovers a local effect (LATE, effect at the cutoff, ATT for switchers), say so here, not in the conclusion's limitations paragraph.
latexY_{ict} = \beta\, D_{ct} + \alpha_c + \gamma_t + X_{ict}'\delta + \varepsilon_{ict}
included (what each index ranges over; what $D_{ct}$ codes). Keep subscript order consistent across equations and tables. Number displayed equations; refer to "equation (1)."
identifies it once controls and fixed effects absorb the rest.
than pretending the design is equation-(1) TWFE — the estimator choice comes from aer-identification.
the evidence. The unit of writing is "assumption — threat — evidence"; name the two or three most plausible violations and point to where each is met.
design's variation, number of clusters, and any small-cluster correction (wild cluster bootstrap) or design-specific inference (AR confidence sets, randomization inference). Never leave inference to the table notes alone.
One paragraph per claim, not per table. Each results paragraph:
magnitude, not the table's existence: "The reform raises earnings by 4.2 percent (Table 3, column 4)," never "Table 3 presents the results."
specifications (added controls, finer fixed effects, alternative samples).
benchmark it.
Column-by-column narration without a finding-first sentence is the most reliable marker of a weak results section.
Every headline coefficient gets three conversions: (1) native units — log-outcome coefficients are log points; use the exact 100·(e^β − 1) whenever |β| > 0.10, and never confuse percent with percentage points; (2) relative to the sample — against the dependent variable's mean or SD; (3) relative to the literature or a policy lever. Report the baseline rate next to every binary marginal effect. If the implied magnitude is implausible, say so and investigate before a referee does. Worked conversions: examples/results-section-example.md; the percent/percentage-point table lives in aer-consistency.
them. "Significant at the 5 percent level," never "almost significant." For nulls, report the CI and what it rules out, distinguishing a precise zero from an uninformative one. Line-level rules: docs/style-guide.md.
calculation: every input and its source stated, multiplied through in one visible chain, rounded honestly — not a cascade of unsourced products. If it needs more than a paragraph, it becomes its own short section. Worked example: examples/results-section-example.md.
The results section cites robustness, it does not contain it: one paragraph summarizing the aer-robustness battery ("stable across clustering, sample, and specification variants; Appendix B reports the full set") plus the single most important check shown inline.
Organize by candidate explanation, not by table:
the ones you will rule out.
auxiliary outcomes ("if the channel is skill-biased adoption, effects should concentrate in tradeable services — they do, Table 5").
then the evidence against it. Steelman first, then answer.
prove the mechanism"). Mechanism evidence is consistency evidence; only the main effect carries design-based identification (see aer-identification).
Short (half a page to one page), doing four things:
sentences copied verbatim (aer-consistency flags duplication).
the complier/locality caveat if the design implies one.
beyond the results, never three.
answerable, not "more research is needed."
No new results, no new citations, no new caveats that belong in Results.
one-to-one onto body sections; nothing promised is undelivered and nothing major appears unannounced.
everywhere — aer-consistency audits this before submission.
and historical events.
docs/style-guide.md — finding-first sentences, no fillertransitions, no AI-pattern tics.
Bundled with the installed skill, no repository checkout needed --- read it before the repo resources below:
references/section-skeletons.md --- per-section skeletons and conclusion-first results narration rulesLoad only the relevant resource:
examples/results-section-example.md
docs/style-guide.mddocs/methods-reference.md
examples/intro-example.md
templates/stata/,templates/r/, or templates/python/
textSECTIONS DRAFTED: <Background | Framework | Data | Strategy | Results | Mechanisms | Conclusion> ESTIMAND STATED: <yes / no — sentence> SAMPLE FUNNEL REPORTED: <yes / no> HEADLINE MAGNITUDE CONVERSIONS: <log-points / percent / vs-mean / vs-literature> BACK-OF-ENVELOPE CALCULATION: <present / not applicable> MECHANISM CHANNELS: <favored + ruled-out list> NEXT SKILL: <aer-introduction | aer-tables-figures | aer-consistency>
reports...")
never benchmarking the magnitude at all
sections
the identified main effect
recycled verbatim into the abstract
Other measured skills in the registry, with their headline benchmark lift.