Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when appraising the cumulative evidence and ensuring balance in an Academy of Management Annals (Annals) review — weighing conflicting findings by credibility, steelmanning rival schools, and handling the author's own work even-handedly. Audits evidence quality and fairness; it does not design the framework (amann-organizing-framework) or build exhibits (amann-tables-figures).
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | 33% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 58% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 8% | 0% |
| case-14 | ✗→✓ | ▲ Improved | 73% | 0% |
| case-01 | ✓→✓ | = Same ✓ | 26% | 0% |
An Annals review carries no estimates of its own — but it is not therefore neutral. The "attitude" is disciplined critical appraisal: you judge how good the cumulative evidence is, reconcile conflicts by credibility, and state where the field's confidence is and is not warranted. This is the review-craft replacement for primary-research robustness checks: you are appraising other people's designs, not defending your own.
A review must be comprehensive in coverage yet selective in emphasis — long (~50 pages) but not an inventory. Resolve the tension by tiering the corpus:
| Tier | Treatment | |------|-----------| | Foundational / field-defining | discussed in text — what it established and its limits | | Important contributions | grouped and weighed within the framework's cells; cited with their finding | | Confirmatory / incremental | cited in clusters ("see also …") to show coverage without bloating prose | | Tangential | cited only where it bears on a specific claim |
Comprehensiveness is proven by the citation set (the saturation log from amann-literature-synthesis); selectivity is exercised in the prose. Equal-length summaries of every paper abdicate the editorial judgment that is the review's value.
Management findings conflict constantly. Reconcile them by why they differ, never by tally:
Occasionally a credibility judgment cannot be settled by reading alone: the review's account of a controversy hinges on whether a staggered-adoption TWFE estimate survives modern corrections, or an apparent consensus may melt once publication bias is priced in. When a load-bearing magnitude comes with a replication package, audit it rather than adjudicate by prose — the shared playbook execution-with-mcp maps each design family to callable StatsPAI / Stata MCP tools (bacon_decomposition to expose bad-comparison weighting, callaway_santanna to re-estimate, honest_did_from_result for pre-trend fragility). Any number produced this way must come from an actual run and be labelled as the review's own re-analysis. This is the exception, not the Annals default: most appraisal here stays qualitative.
Annals is the review-of-the-field: its account of a debate becomes the field's shared reference, and the surveyed authors often referee the review. Balance is therefore both ethical and strategic.
text【Tiering】corpus split foundational/important/confirmatory/tangential? Y/N 【Comprehensiveness evidence】saturation log + citation set support "nothing important missing"? Y/N 【Conflict handling】reconciled by credibility + construct + context (not vote-count)? Y/N 【Evidence quality】strong/mixed/thin stated per claim? Y/N 【Steelman】each rival school stated at its strongest? Y/N 【Self-citation audit】own work at warranted tier; emphasis identity-blind? Y/N 【Attitude labelled】provocative reads marked as author's judgment? Y/N 【Next skill】→ amann-tables-figures (who-found-what tables + framework figure)
Other measured skills in the registry, with their headline benchmark lift.