Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Apply structured critical thinking — identifying claims, evidence, reasoning chains, hidden assumptions, and logical fallacies — to evaluate or construct specific written arguments rigorously. Use this skill when the user presents a concrete argument, claim, op-ed, research finding, or piece of reasoning to be analyzed for logical validity or flaws, even if they say 'is this argument valid', 'what logical fallacies are in this', or 'what assumptions am I making in this thesis'. Do NOT use for ca
.claude/skills/asgard-ai-platform-hum-critical-thinking/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 3% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 51% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 22% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 38% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 61% | 0% |
Critical thinking systematically evaluates arguments by decomposing them into claims, evidence, reasoning, and assumptions. It identifies where arguments are strong, weak, or fallacious — not to "win" debates but to arrive at better-justified conclusions.
IRON LAW: Separate the Argument from the Person
Evaluate the ARGUMENT (claim + evidence + reasoning), not the person
making it. A bad person can make a good argument. A trusted expert
can make a bad argument. Ad hominem (attacking the person) and appeal
to authority (trusting the person) are both fallacies.Every argument has four components:
Step 1: Identify the claim — What exactly is being argued? Restate in one sentence.
Step 2: Examine the evidence
Step 3: Evaluate the reasoning
Step 4: Surface assumptions
| Fallacy | What It Does | Example | |---------|-------------|---------| | Ad hominem | Attacks the person, not the argument | "You can't talk about economics, you're not an economist" | | Straw man | Distorts the opponent's argument to attack a weaker version | "You want to reduce military spending? So you want us defenseless?" | | False dichotomy | Presents only two options when more exist | "You're either with us or against us" | | Slippery slope | Claims one event will inevitably lead to extreme consequences | "If we allow remote work, soon no one will come to the office ever" | | Appeal to authority | Uses authority status instead of evidence | "The CEO says AI will replace all jobs, so it must be true" | | Hasty generalization | Draws broad conclusion from limited cases | "My two friends who studied art are unemployed, so art degrees are useless" | | Red herring | Introduces irrelevant information to distract | "Yes, our product has bugs, but look at our amazing company culture" | | Circular reasoning | Conclusion is assumed in the premise | "This is the best approach because there's no better one" |
markdown# Argument Analysis: {Topic} ## Claim {One-sentence restatement of the core argument} ## Evidence Assessment | Evidence | Type | Sufficient? | Relevant? | Current? | |----------|------|------------|-----------|----------| | {evidence 1} | {fact/anecdote/expert/stat} | Y/N | Y/N | Y/N | ## Reasoning Evaluation - Logical validity: {valid / fallacious} - Fallacies detected: {list with explanation} - Alternative explanations: {what else could explain the evidence} ## Hidden Assumptions 1. {assumption} — reasonable? {Y/N, why} ## Verdict - Argument strength: Strong / Moderate / Weak - Key weakness: {the biggest flaw} - What would strengthen it: {what evidence or reasoning is missing}
Scenario: Evaluating the claim "Remote work reduces productivity"
references/formal-logic.mdreferences/fallacy-catalog.md| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 19,450 | 16,015 | -18% | 1 | 1 | 0% | 3,491 | 3,584 | +3% | 0 | 0 | — |
case-02 | fail→pass | 13,963 | 13,850 | -1% | 1 | 1 | 0% | 2,421 | 3,650 | +51% | 0 | 0 | — |
case-03 | fail→pass | 16,134 | 12,012 | -26% | 1 | 1 | 0% | 2,784 | 3,387 | +22% | 0 | 0 | — |
case-04 | pass→pass | 9,558 | 8,776 | -8% | 1 | 1 | 0% | 2,130 | 3,369 | +58% | 0 | 0 | — |
case-05 | pass→pass | 13,677 | 15,541 | +14% | 1 | 1 | 0% | 2,326 | 3,808 | +64% | 0 | 0 | — |
case-06 | pass→pass | 13,993 | 14,418 | +3% | 1 | 1 | 0% | 2,374 | 3,611 | +52% | 0 | 0 | — |
case-07 | pass→pass | 14,019 | 9,946 | -29% | 1 | 1 | 0% | 2,133 | 3,164 | +48% | 0 | 0 | — |
case-08 | fail→pass | 15,218 | 10,979 | -28% | 1 | 1 | 0% | 2,255 | 3,111 | +38% | 0 | 0 | — |
case-09 | fail→pass | 12,791 | 12,282 | -4% | 1 | 1 | 0% | 2,116 | 3,408 | +61% | 0 | 0 | — |
case-10 | pass→pass | 13,272 | 10,249 | -23% | 1 | 1 | 0% | 2,121 | 2,790 | +32% | 0 | 0 | — |
case-11 | fail→pass | 15,325 | 11,389 | -26% | 1 | 1 | 0% | 2,414 | 2,885 | +20% | 0 | 0 | — |
case-12 | fail→pass | 18,342 | 10,554 | -42% | 1 | 1 | 0% | 2,266 | 3,072 | +36% | 0 | 0 | — |
case-13 | pass→pass | 12,580 | 10,545 | -16% | 1 | 1 | 0% | 2,107 | 3,005 | +43% | 0 | 0 | — |
case-14 | pass→pass | 15,903 | 15,204 | -4% | 1 | 1 | 0% | 2,230 | 3,627 | +63% | 0 | 0 | — |
case-15 | pass→pass | 14,567 | 9,641 | -34% | 1 | 1 | 0% | 2,272 | 2,889 | +27% | 0 | 0 | — |
case-16 | fail→pass | 12,423 | 8,256 | -34% | 1 | 1 | 0% | 1,806 | 2,727 | +51% | 0 | 0 | — |
case-17 | fail→pass | 14,825 | 15,203 | +3% | 1 | 1 | 0% | 2,144 | 3,596 | +68% | 0 | 0 | — |
case-18 | pass→pass | 10,096 | 8,391 | -17% | 1 | 1 | 0% | 1,694 | 2,637 | +56% | 0 | 0 | — |
case-19 | fail→fail | 14,043 | 12,576 | -10% | 1 | 1 | 0% | 2,223 | 3,055 | +37% | 0 | 0 | — |
case-20 | pass→pass | 15,714 | 11,320 | -28% | 1 | 1 | 0% | 2,266 | 3,098 | +37% | 0 | 0 | — |
case-21 | fail→pass | 11,352 | 12,522 | +10% | 1 | 1 | 0% | 1,788 | 3,183 | +78% | 0 | 0 | — |
case-22 | fail→pass | 13,037 | 25,107 | +93% | 1 | 1 | 0% | 1,973 | 3,926 | +99% | 0 | 0 | — |
case-23 | pass→pass | 11,770 | 14,552 | +24% | 1 | 1 | 0% | 1,966 | 3,455 | +76% | 0 | 0 | — |
case-24 | pass→fail | 16,219 | 12,536 | -23% | 1 | 1 | 0% | 2,364 | 3,460 | +46% | 0 | 0 | — |
case-25 | fail→pass | 15,113 | 11,571 | -23% | 1 | 1 | 0% | 2,102 | 2,884 | +37% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 25 cases were attempted. The headline lift of +44 percentage points is the difference between those two pass rates over the 25 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.