Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Audit figures, tables, captions, cross-references, and statistical notes.
.claude/skills/brycewang-stanford-figure-table-audit/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-03 | ✗→✓ | ▲ Improved | 17% | 0% |
| case-04 | ✗→✓ | ▲ Improved | 38% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 39% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 54% | 0% |
| case-22 | ✗→✓ | ▲ Improved | 28% | 0% |
This is an original Open Science Skills workflow for manuscript QA. It remixes general figure/table and citation-compliance ideas from Cheng-I Wu's Academic Research Skills for Claude Code (CC BY-NC 4.0), but is rewritten for open-science social-science manuscripts. It is not a visual hallucination engine: when a claim requires reading plotted values from an image, prefer source data or mark the issue as needing author verification.
This is the end-stage auditor. For figure design and production guidance during drafting, use the figures skill; for table design, use the tables skill. Run figure-table-audit once the figure and table set is stable and you are preparing for submission.
Identify:
If only a PDF is available, state that cross-reference and value checks are lower confidence.
Build an inventory with:
Check:
For each figure/table used to support a substantive claim:
Do not infer exact values by eyeballing a plot unless the figure encodes labeled values. If source data are unavailable, write VISUAL READ ONLY - AUTHOR VERIFY.
Captions and table notes should let a reader understand the evidence without hunting:
For conjoint, list-experiment, topic-modeling, LLM-classification, and OCR studies, invoke or recommend the relevant sibling skill when table/figure interpretation depends on method-specific standards.
Flag:
Check whether:
Produce a Figure and Table Audit Report:
# Figure and Table Audit Report
Scope:
Inputs checked:
Build/source status:
Summary: <N blocking, N recommended, N minor, N author-verification>
## Inventory
| ID | Path/location | Caption/title | First callout | Source/script |
## Blocking Issues
| Location | Figure/table | Issue | Evidence | Fix |
## Recommended Fixes
| Location | Figure/table | Issue | Fix |
## Minor / Production Issues
| Figure/table | Issue | Fix |
## Author Verification Needed
| Figure/table | Why verification is needed |
## Readiness Checklist
| Dimension | PASS/FAIL/PARTIAL/NA | Notes |Severity:
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 31,118 | 12,733 | -59% | 1 | 1 | 0% | 5,467 | 3,544 | -35% | 0 | 0 | — |
case-02 | fail→fail | 3,285 | 3,941 | +20% | 1 | 1 | 0% | 198 | 1,565 | +690% | 0 | 0 | — |
case-03 | fail→pass | 28,824 | 24,659 | -14% | 1 | 1 | 0% | 5,394 | 6,334 | +17% | 0 | 0 | — |
case-04 | fail→pass | 14,399 | 13,898 | -3% | 1 | 1 | 0% | 2,885 | 3,984 | +38% | 0 | 0 | — |
case-05 | pass→pass | 17,057 | 15,999 | -6% | 1 | 1 | 0% | 3,125 | 4,300 | +38% | 0 | 0 | — |
case-06 | pass→pass | 7,171 | 8,199 | +14% | 1 | 1 | 0% | 1,228 | 2,580 | +110% | 0 | 0 | — |
case-07 | pass→pass | 9,518 | 9,013 | -5% | 1 | 1 | 0% | 1,852 | 2,944 | +59% | 0 | 0 | — |
case-08 | fail→pass | 10,603 | 7,008 | -34% | 1 | 1 | 0% | 1,886 | 2,616 | +39% | 0 | 0 | — |
case-09 | fail→pass | 16,214 | 15,782 | -3% | 1 | 1 | 0% | 2,772 | 4,271 | +54% | 0 | 0 | — |
case-10 | pass→pass | 6,007 | 5,099 | -15% | 1 | 1 | 0% | 1,244 | 2,274 | +83% | 0 | 0 | — |
case-11 | pass→pass | 7,707 | 4,902 | -36% | 1 | 1 | 0% | 1,448 | 2,305 | +59% | 0 | 0 | — |
case-12 | pass→pass | 9,503 | 7,444 | -22% | 1 | 1 | 0% | 1,860 | 2,728 | +47% | 0 | 0 | — |
case-13 | pass→pass | 10,363 | 7,776 | -25% | 1 | 1 | 0% | 1,885 | 2,865 | +52% | 0 | 0 | — |
case-14 | pass→pass | 13,370 | 7,747 | -42% | 1 | 1 | 0% | 2,269 | 2,682 | +18% | 0 | 0 | — |
case-15 | pass→pass | 9,475 | 3,900 | -59% | 1 | 1 | 0% | 1,441 | 2,066 | +43% | 0 | 0 | — |
case-16 | pass→pass | 9,020 | 4,251 | -53% | 1 | 1 | 0% | 1,447 | 2,009 | +39% | 0 | 0 | — |
case-17 | pass→pass | 8,261 | 6,261 | -24% | 1 | 1 | 0% | 1,473 | 2,417 | +64% | 0 | 0 | — |
case-18 | pass→pass | 6,156 | 4,979 | -19% | 1 | 1 | 0% | 1,142 | 2,230 | +95% | 0 | 0 | — |
case-19 | pass→pass | 8,140 | 3,705 | -54% | 1 | 1 | 0% | 1,432 | 1,969 | +38% | 0 | 0 | — |
case-20 | pass→pass | 11,295 | 11,897 | +5% | 1 | 1 | 0% | 1,978 | 3,290 | +66% | 0 | 0 | — |
case-21 | pass→pass | 7,206 | 4,715 | -35% | 1 | 1 | 0% | 1,144 | 2,108 | +84% | 0 | 0 | — |
case-22 | fail→pass | 8,042 | 3,250 | -60% | 1 | 1 | 0% | 1,403 | 1,799 | +28% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 21 counted toward the lift figure. The other 1 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +23 percentage points is the difference between those two pass rates over the 21 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.