Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Poll RSS, JSON APIs, and GitHub with watermark dedup.
.claude/skills/nousresearch-watchers/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-10 | ✗→✓ | ▲ Improved | -22% | 0% |
| case-09 | ✗→✓ | ▲ Improved | -13% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 33% | 0% |
| case-07 | ✗→✓ | ▲ Improved | -19% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 15% | 0% |
Poll external sources on an interval and react only to new items. Three ready-made scripts plus a shared watermark helper; wire them into a cron job (or run them ad-hoc from the terminal).
A watcher is just a script that:
The scripts below handle all three. The agent runs them via the terminal tool — from a cron job, a webhook, or an interactive chat — and reports what's new.
All three live in $HERMES_HOME/skills/devops/watchers/scripts/ once the skill is installed. Each reads WATCHER_STATE_DIR (defaults to $HERMES_HOME/watcher-state/) for its state file, keyed by the --name argument.
| Script | What it watches | Dedup key | |---|---|---| | watch_rss.py | RSS 2.0 or Atom feed URL | <guid> / <id> | | watch_http_json.py | Any JSON endpoint returning a list of objects | Configurable id field | | watch_github.py | GitHub issues / pulls / releases / commits for a repo | id / sha |
All three:
## <title>\n<url>\n\n<optional body> per itemRun a watcher directly from the terminal tool:
bashpython $HERMES_HOME/skills/devops/watchers/scripts/watch_rss.py \ --name hn --url https://news.ycombinator.com/rss --max 5
Watch a GitHub repo (set GITHUB_TOKEN in ${HERMES_HOME:-~/.hermes}/.env to avoid the 60 req/hr anonymous rate limit):
bashpython $HERMES_HOME/skills/devops/watchers/scripts/watch_github.py \ --name hermes-issues --repo NousResearch/hermes-agent --scope issues
Poll an arbitrary JSON API:
bashpython $HERMES_HOME/skills/devops/watchers/scripts/watch_http_json.py \ --name api --url https://api.example.com/events \ --id-field event_id --items-path data.events
Ask the agent to schedule a cron job with a prompt like:
> Every 15 minutes, run watch_rss.py --name hn --url https://news.ycombinator.com/rss. If it prints anything, summarize the headlines and deliver them. If it prints nothing, stay silent.
The agent invokes the script via the terminal tool inside the cron job's agent loop; no changes to cron's built-in --script flag are needed.
Every watcher writes $HERMES_HOME/watcher-state/<name>.json. Inspect:
bashcat $HERMES_HOME/watcher-state/hn.json
Force a replay (next run treated as first poll):
bashrm $HERMES_HOME/watcher-state/hn.json
All three scripts use the same template: load watermark, fetch, diff, save, emit. scripts/_watermark.py is the shared helper; import it to get atomic writes + bounded ID set + first-run baseline for free. See any of the three reference scripts for how little boilerplate it takes.
--prime-with-latest N flag in your own script.$HERMES_HOME/watcher-state/ is always writable. Docker/Modal backends may not see arbitrary host paths.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-10 | fail→pass | 10,187 | 1,323 | -87% | 1 | 1 | 0% | 1,719 | 1,349 | -22% | 0 | 0 | — |
case-01 | fail→fail | 7,808 | 5,613 | -28% | 1 | 1 | 0% | 1,488 | 1,568 | +5% | 0 | 0 | — |
case-02 | fail→fail | 6,827 | 6,928 | +1% | 1 | 1 | 0% | 1,291 | 1,757 | +36% | 0 | 0 | — |
case-03 | fail→fail | 8,019 | 4,448 | -45% | 1 | 1 | 0% | 1,565 | 1,427 | -9% | 0 | 0 | — |
case-04 | pass→pass | 11,012 | 13,068 | +19% | 1 | 1 | 0% | 2,310 | 3,832 | +66% | 0 | 0 | — |
case-09 | fail→pass | 8,856 | 2,769 | -69% | 1 | 1 | 0% | 1,858 | 1,615 | -13% | 0 | 0 | — |
case-05 | pass→pass | 9,806 | 4,244 | -57% | 1 | 1 | 0% | 1,801 | 1,826 | +1% | 0 | 0 | — |
case-06 | fail→pass | 10,174 | 7,091 | -30% | 1 | 1 | 0% | 1,858 | 2,474 | +33% | 0 | 0 | — |
case-07 | fail→pass | 11,948 | 4,258 | -64% | 1 | 1 | 0% | 2,230 | 1,805 | -19% | 0 | 0 | — |
case-08 | pass→pass | 10,566 | 3,233 | -69% | 1 | 1 | 0% | 1,740 | 1,817 | +4% | 0 | 0 | — |
case-11 | fail→pass | 7,836 | 1,739 | -78% | 1 | 1 | 0% | 1,245 | 1,430 | +15% | 0 | 0 | — |
case-12 | pass→pass | 8,155 | 2,317 | -72% | 1 | 1 | 0% | 1,497 | 1,542 | +3% | 0 | 0 | — |
case-13 | pass→pass | 6,644 | 2,476 | -63% | 1 | 1 | 0% | 1,045 | 1,530 | +46% | 0 | 0 | — |
case-14 | fail→pass | 12,344 | 2,312 | -81% | 1 | 1 | 0% | 2,065 | 1,522 | -26% | 0 | 0 | — |
case-15 | pass→pass | 8,787 | 1,769 | -80% | 1 | 1 | 0% | 1,424 | 1,450 | +2% | 0 | 0 | — |
case-16 | fail→pass | 11,611 | 2,176 | -81% | 1 | 1 | 0% | 2,057 | 1,491 | -28% | 0 | 0 | — |
case-17 | pass→pass | 11,938 | 2,951 | -75% | 1 | 1 | 0% | 2,038 | 1,629 | -20% | 0 | 0 | — |
case-18 | pass→pass | 12,991 | 7,565 | -42% | 1 | 1 | 0% | 1,988 | 2,235 | +12% | 0 | 0 | — |
case-19 | fail→pass | 7,516 | 2,601 | -65% | 1 | 1 | 0% | 1,320 | 1,555 | +18% | 0 | 0 | — |
case-20 | pass→pass | 5,968 | 1,833 | -69% | 1 | 1 | 0% | 1,010 | 1,501 | +49% | 0 | 0 | — |
case-21 | fail→pass | 5,365 | 3,307 | -38% | 1 | 1 | 0% | 796 | 1,436 | +80% | 0 | 0 | — |
case-22 | pass→pass | 16,266 | 11,301 | -31% | 1 | 1 | 0% | 3,140 | 3,333 | +6% | 0 | 0 | — |
case-23 | pass→pass | 14,830 | 12,538 | -15% | 1 | 1 | 0% | 2,668 | 3,326 | +25% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted, and 20 counted toward the lift figure. The other 3 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +39 percentage points is the difference between those two pass rates over the 20 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.