Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Structure and write comprehensive literature reviews for any field
.claude/skills/brycewang-stanford-literature-review-writing/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | 27% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 35% | 0% |
| case-19 | ✗→✓ | ▲ Improved | 50% | 0% |
| case-03 | ✓→✓ | = Same ✓ | 56% | 0% |
| case-04 | ✓→✓ | = Same ✓ | 93% | 0% |
A skill for structuring, synthesizing, and writing literature reviews. Covers narrative reviews, systematic reviews, scoping reviews, and literature review sections within empirical papers. Provides frameworks for organizing sources, identifying themes, and writing with synthesis rather than summary.
| Type | Purpose | Search | Analysis | Length | |------|---------|--------|----------|--------| | Narrative | Broad overview of a topic | Selective | Qualitative synthesis | 5-30 pages | | Systematic | Answer a specific question exhaustively | Exhaustive, documented | May include meta-analysis | 10-40 pages | | Scoping | Map the extent of research on a topic | Broad, systematic | Charting and categorizing | 10-20 pages | | Integrative | Synthesize diverse methodologies | Targeted | Conceptual synthesis | 10-25 pages | | Umbrella | Review of systematic reviews | Focused on SRs | Synthesis of syntheses | 10-20 pages |
pythondef create_literature_matrix(sources: list[dict]) -> dict: """ Create a structured matrix for organizing reviewed literature. Args: sources: List of dicts with keys: author, year, title, method, sample, key_findings, themes, limitations """ matrix = { "headers": [ "Author (Year)", "Research Question", "Method", "Sample/Data", "Key Findings", "Themes", "Limitations" ], "rows": [], "theme_index": {} } for src in sources: row = { "citation": f"{src['author']} ({src['year']})", "question": src.get("research_question", ""), "method": src.get("method", ""), "sample": src.get("sample", ""), "findings": src.get("key_findings", ""), "themes": src.get("themes", []), "limitations": src.get("limitations", "") } matrix["rows"].append(row) for theme in src.get("themes", []): matrix["theme_index"].setdefault(theme, []).append( row["citation"] ) return matrix
Chronological (rarely ideal for reviews):
"In 2010, Smith found X. Then in 2012, Jones found Y.
In 2015, Lee found Z."
Problem: Reads like an annotated bibliography, not a synthesis.
Thematic (recommended):
"Three factors have been identified as predictors of X.
First, [factor A] has been consistently supported (Smith, 2010;
Jones, 2012; Lee, 2015). Second, [factor B] shows mixed results..."
Advantage: Synthesizes findings around conceptual themes.
Methodological (useful for systematic reviews):
"Studies using qualitative methods (n=12) found [pattern],
while quantitative studies (n=25) reported [different pattern].
This methodological divide suggests..."Summary (weak -- describes one paper at a time):
"Smith (2020) studied 200 undergraduates and found that sleep
quality predicted academic performance. Jones (2021) surveyed
150 graduate students and found a similar relationship."
Synthesis (strong -- integrates multiple sources around a point):
"Sleep quality has been consistently linked to academic performance
across both undergraduate (Smith, 2020; Lee, 2019) and graduate
(Jones, 2021) populations, with effect sizes ranging from r=0.25
to r=0.42. However, this relationship may be confounded by
socioeconomic factors (Park, 2022), which only two studies
controlled for."Agreement:
"There is broad consensus that..."
"Multiple studies converge on the finding that..."
"This finding has been replicated across [contexts]..."
Disagreement:
"However, findings diverge regarding..."
"In contrast to the majority view, [author] argues..."
"The evidence is mixed, with some studies reporting [X] and others [Y]..."
Gap identification:
"Notably absent from this literature is..."
"While [aspect] has been well studied, [gap] remains unexplored..."
"No studies to date have examined [specific gap]..."
Transition:
"Taken together, these findings suggest..."
"Building on this body of work, recent studies have begun to..."
"This line of research has evolved from [earlier focus] to [current focus]..."1. Introduction (1-2 pages)
- Define the topic and scope
- Explain why this review is needed (gap, timeliness, controversy)
- State the review's objectives or research questions
2. Methods (for systematic/scoping reviews)
- Search strategy, databases, date range
- Inclusion/exclusion criteria
- Screening process (PRISMA flow diagram)
3. Body: Thematic Sections (bulk of the review)
- Each section covers a theme, construct, or sub-question
- Synthesize rather than summarize
- Use tables to compare studies when appropriate
4. Discussion / Synthesis
- What is the overall state of knowledge?
- Where do studies agree and disagree?
- What are the gaps?
5. Conclusion / Future Directions
- Summarize the key takeaways
- Propose a research agenda addressing identified gapsProblem: "Laundry list" structure
- Each paragraph describes one study in isolation
Fix: Group studies by theme and synthesize across them
Problem: Missing recent literature
- Review stops at 2020 in a fast-moving field
Fix: Search within the last 12 months before submission
Problem: Uncritical acceptance
- All studies treated as equally valid
Fix: Evaluate methodological quality and note study limitations
Problem: No conceptual framework
- Sources listed without a guiding structure
Fix: Start with a framework (theoretical, conceptual, or thematic map)
Problem: Omitting contradictory evidence
- Only citing studies that support the authors' position
Fix: Actively seek and discuss disconfirming evidence| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 19,975 | 23,483 | +18% | 1 | 1 | 0% | 4,213 | 5,073 | +20% | 0 | 0 | — |
case-02 | fail→pass | 24,065 | 17,087 | -29% | 1 | 1 | 0% | 3,528 | 4,479 | +27% | 0 | 0 | — |
case-03 | pass→pass | 17,021 | 13,275 | -22% | 1 | 1 | 0% | 2,408 | 3,748 | +56% | 0 | 0 | — |
case-04 | pass→pass | 10,246 | 22,391 | +119% | 1 | 1 | 0% | 1,737 | 3,345 | +93% | 0 | 0 | — |
case-05 | fail→fail | 18,311 | 14,875 | -19% | 1 | 1 | 0% | 2,504 | 4,133 | +65% | 0 | 0 | — |
case-06 | fail→pass | 12,441 | 3,394 | -73% | 1 | 1 | 0% | 1,613 | 2,171 | +35% | 0 | 0 | — |
case-07 | pass→pass | 7,917 | 6,027 | -24% | 1 | 1 | 0% | 1,011 | 2,456 | +143% | 0 | 0 | — |
case-08 | pass→pass | 10,766 | 10,114 | -6% | 1 | 1 | 0% | 1,756 | 3,349 | +91% | 0 | 0 | — |
case-09 | pass→pass | 13,499 | 10,738 | -20% | 1 | 1 | 0% | 1,860 | 3,288 | +77% | 0 | 0 | — |
case-10 | pass→pass | 12,075 | 13,679 | +13% | 1 | 1 | 0% | 1,862 | 3,645 | +96% | 0 | 0 | — |
case-11 | pass→pass | 13,894 | 15,165 | +9% | 1 | 1 | 0% | 2,517 | 4,279 | +70% | 0 | 0 | — |
case-12 | pass→pass | 16,642 | 12,366 | -26% | 1 | 1 | 0% | 2,317 | 3,510 | +51% | 0 | 0 | — |
case-13 | pass→pass | 14,271 | 13,654 | -4% | 1 | 1 | 0% | 2,130 | 3,865 | +81% | 0 | 0 | — |
case-14 | pass→pass | 8,369 | 7,592 | -9% | 1 | 1 | 0% | 1,102 | 2,751 | +150% | 0 | 0 | — |
case-15 | pass→pass | 30,684 | 7,219 | -76% | 1 | 1 | 0% | 2,183 | 2,742 | +26% | 0 | 0 | — |
case-16 | pass→pass | 15,914 | 15,089 | -5% | 1 | 1 | 0% | 2,288 | 3,927 | +72% | 0 | 0 | — |
case-17 | pass→pass | 11,278 | 7,388 | -34% | 1 | 1 | 0% | 1,867 | 2,715 | +45% | 0 | 0 | — |
case-18 | fail→fail | 17,800 | 14,333 | -19% | 1 | 1 | 0% | 2,488 | 3,902 | +57% | 0 | 0 | — |
case-19 | fail→pass | 19,081 | 14,722 | -23% | 1 | 1 | 0% | 2,883 | 4,323 | +50% | 0 | 0 | — |
case-20 | pass→pass | 16,271 | 13,999 | -14% | 1 | 1 | 0% | 2,664 | 3,968 | +49% | 0 | 0 | — |
case-21 | pass→pass | 8,696 | 10,316 | +19% | 1 | 1 | 0% | 1,667 | 3,180 | +91% | 0 | 0 | — |
case-22 | pass→pass | 16,028 | 17,144 | +7% | 1 | 1 | 0% | 2,340 | 4,469 | +91% | 0 | 0 | — |
case-23 | pass→pass | 23,762 | 20,602 | -13% | 1 | 1 | 0% | 3,399 | 4,946 | +46% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted. The headline lift of +13 percentage points is the difference between those two pass rates over the 23 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.