Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Scan and summarize recent academic papers from Semantic Scholar
.claude/skills/paperclipai-paper-digest/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-20 | ✓→✓ | = Same ✓ | -33% | 0% |
| case-21 | ✓→✓ | = Same ✓ | 23% | 0% |
| case-22 | ✓→✓ | = Same ✓ | 34% | 0% |
| case-01 | ✗→✗ | = Same ✗ | -6% | 0% |
| case-02 | ✗→✗ | = Same ✗ | 43% | 0% |
Scans Semantic Scholar for recent papers matching tracked topics and produces curated summaries.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 13,263 | 18,185 | +37% | 1 | 1 | 0% | 2,177 | 2,045 | -6% | 0 | 0 | — |
case-02 | fail→fail | 14,986 | 19,503 | +30% | 1 | 1 | 0% | 2,447 | 3,499 | +43% | 0 | 0 | — |
case-03 | fail→fail | 20,859 | 17,361 | -17% | 1 | 1 | 0% | 3,701 | 3,198 | -14% | 0 | 0 | — |
case-04 | fail→fail | 13,134 | 15,638 | +19% | 1 | 1 | 0% | 2,112 | 2,599 | +23% | 0 | 0 | — |
case-05 | fail→fail | 15,893 | 6,825 | -57% | 1 | 1 | 0% | 2,547 | 285 | -89% | 0 | 0 | — |
case-06 | fail→fail | 15,034 | 18,437 | +23% | 1 | 1 | 0% | 2,622 | 3,264 | +24% | 0 | 0 | — |
case-07 | fail→fail | 14,681 | 12,269 | -16% | 1 | 1 | 0% | 2,673 | 2,318 | -13% | 0 | 0 | — |
case-08 | fail→fail | 15,090 | 14,079 | -7% | 1 | 1 | 0% | 2,358 | 1,919 | -19% | 0 | 0 | — |
case-09 | fail→fail | 17,199 | 19,069 | +11% | 1 | 1 | 0% | 2,665 | 2,972 | +12% | 0 | 0 | — |
case-10 | fail→fail | 16,859 | 20,321 | +21% | 1 | 1 | 0% | 2,980 | 2,725 | -9% | 0 | 0 | — |
case-11 | fail→fail | 13,351 | 11,717 | -12% | 1 | 1 | 0% | 2,280 | 1,895 | -17% | 0 | 0 | — |
case-12 | fail→fail | 15,634 | 19,696 | +26% | 1 | 1 | 0% | 2,510 | 2,452 | -2% | 0 | 0 | — |
case-13 | fail→fail | 17,713 | 34,900 | +97% | 1 | 1 | 0% | 3,165 | 2,715 | -14% | 0 | 0 | — |
case-14 | fail→fail | 14,609 | 10,902 | -25% | 1 | 1 | 0% | 2,381 | 1,709 | -28% | 0 | 0 | — |
case-15 | fail→fail | 17,957 | 22,077 | +23% | 1 | 1 | 0% | 3,209 | 3,513 | +9% | 0 | 0 | — |
case-16 | fail→fail | 14,963 | 19,589 | +31% | 1 | 1 | 0% | 2,328 | 2,699 | +16% | 0 | 0 | — |
case-17 | fail→fail | 14,385 | 15,442 | +7% | 1 | 1 | 0% | 2,088 | 2,665 | +28% | 0 | 0 | — |
case-18 | fail→fail | 20,374 | 15,585 | -24% | 1 | 1 | 0% | 2,923 | 2,788 | -5% | 0 | 0 | — |
case-19 | fail→fail | 17,624 | 16,525 | -6% | 1 | 1 | 0% | 2,923 | 2,653 | -9% | 0 | 0 | — |
case-20 | pass→pass | 7,856 | 6,040 | -23% | 1 | 1 | 0% | 1,943 | 1,300 | -33% | 0 | 0 | — |
case-21 | pass→pass | 8,190 | 8,187 | -0% | 1 | 1 | 0% | 1,434 | 1,763 | +23% | 0 | 0 | — |
case-22 | pass→pass | 6,841 | 9,730 | +42% | 1 | 1 | 0% | 1,184 | 1,587 | +34% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 21 counted toward the lift figure. The other 1 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of 0 percentage points is the difference between those two pass rates over the 21 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Other measured skills in the registry, with their headline benchmark lift.