hermes-plugin-evaluation

hermes-plugin-evaluation is a skill for Claude Code, Codex from AtlasOmnia/donna-starter. It costs 37 tokens per session (743 once invoked), scanned A, original, MIT.

A review process for deciding whether to install and trust a third-party Hermes plugin or integration.

In plain words
What is it for?
Checking plugin documentation, licenses, dependencies, service charges, data flows, tunnels, webhooks, and setup behavior before deciding whether an integration is suitable.
Why use it?
A plugin may be free to install but still require paid services, expose data, open webhooks, or change gateway behavior. The review separates those costs and risks before installation.

Skill for Claude CodeCodex

Which agent this was written for is unclear — built for hermes-agent. Also seen: built for hermes-agent.

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/atlasomnia/donna-starter/hermes-plugin-evaluation
Any agent
npx skills add AtlasOmnia/donna-starter --skill hermes-plugin-evaluation
Clone the repo
git clone --depth 1 https://github.com/AtlasOmnia/donna-starter

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for hermes-plugin-evaluation

README.md
[![agentmods](https://agentmods.dev/badge/skills/atlasomnia/donna-starter/hermes-plugin-evaluation.svg)](https://agentmods.dev/skills/atlasomnia/donna-starter/hermes-plugin-evaluation)
Your own site
<a href="https://agentmods.dev/skills/atlasomnia/donna-starter/hermes-plugin-evaluation"><img src="https://agentmods.dev/badge/skills/atlasomnia/donna-starter/hermes-plugin-evaluation.svg" alt="Measured on agentmods" height="20"></a>
Per session 37 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 743 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00037 $0.00743
Opus 5 $0.00018 $0.00371
Sonnet 5 $0.00007 $0.00149
Haiku 4.5 $0.00004 $0.00074

Measured 6d ago against content hash b7343b474243, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

hermes-plugin-evaluation scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

Copies of this mod

1 near-identical copy found in the catalogue:

skills/autonomous-ai-agents/hermes-plugin-evaluation/SKILL.md · 79 lines

How it starts

The opening of the file, as written. The whole thing — 79 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Hermes Plugin Evaluation

Use this when the user asks about installing, trusting, pricing, or using a third-party Hermes Agent plugin or integration.

Goal

Give a practical go/no-go read before touching the live Hermes gateway. Separate:

  • Plugin code cost/license — whether the repository itself is public/open-source/free to install.
  • Service cost — required SaaS account, usage billing, phone/SMS/voice costs, model API costs, tunnels, hosted routing, or paid feature gates.
  • Operational risk — what data leaves Hermes, whether the plugin opens public webhooks/tunnels, and whether it modifies gateway behavior.

Fast evaluation workflow

  1. Load authoritative Hermes context first when the task involves Hermes plugins:
  • Load hermes-agent if available.
  • Prefer official Hermes docs for CLI syntax and plugin lifecycle.
  1. Inspect the repository without installing it:
  • README / docs
  • plugin.yaml
  • pyproject.toml, package.json, lockfiles
  • LICENSE, NOTICE, or equivalent
  • setup wizard files and after-install notes
  • tool definitions, platform adapter files, webhook/tunnel code
  1. Answer cost precisely:
  • Say “plugin appears free/public” only for the repo/installable code.
  • Do not infer the hosted service is free just because the plugin is public.
  • Identify paid dependencies: phone numbers, SMS/MMS, voice minutes, hosted tunnels, OpenAI/Anthropic/etc. APIs, storage, or managed accounts.
  1. Check for a public pricing page, but treat absence as unknown, not free:
  • If pricing is missing/404/gated, say “pricing not publicly obvious; confirm with vendor.”
  1. Report required credentials and data paths:
  • Required env vars/API keys.
  • Optional credentials that change cost or data flow.
  • Whether inbound messages/calls pass through vendor infrastructure.
  1. Recommend a safe rollout:
  • Test in a non-critical Hermes profile or disabled gateway first.
  • Avoid putting it on the main gateway until pricing, credentials, and data flow are understood.
  • Run plugin-specific doctor/diagnostics before enabling public channels.

Read the full file on GitHub · 79 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 79 lines · 37 tokens per session scan A b7343b474243

Subscribe to this mod's changes

hermes-plugin-evaluation is a skill published in the GitHub repository AtlasOmnia/donna-starter (104 stars, last pushed 7d ago), licensed MIT. It adds 37 tokens to every session and 743 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

npm-downloads-to-leads

Takes a list of npm package names (yours or competitors'), fetches 12 weeks of daily download data from the npm API, computes a breakout velocity score per package to identify hockey-stick growth, fetches maintainer profiles from the npm registry and GitHub API, and outputs a ranked lead brief for each breakout…

Varnan-Tech/opendirectory · 172 tokens

sdk-adoption-tracker

Given your SDK or library name, searches GitHub code search for public repos that import or require it, classifies each repo as company org, affiliated developer, solo developer, or tutorial noise, scores by adoption signal strength, detects new adopters by date, and outputs a ranked list of who is building on you…

Varnan-Tech/opendirectory · 170 tokens

domain-expired-opportunity-finder

Evaluates expired domain candidates against a target niche, scores them by topical relevance, historical activity level, and history cleanliness, then outputs a ranked shortlist with explainable reasoning and risk flags.

Varnan-Tech/opendirectory · 45 tokens

gh-issue-to-demand-signal

Takes a competitor's public GitHub repo URL, fetches their open issues via the GitHub REST API, filters noise locally, clusters issues into 6 demand categories, computes a demand score per issue and per cluster, and outputs a ranked demand gap report with a GTM messaging brief. Use when asked to scan a competitor's…

Varnan-Tech/opendirectory · 154 tokens

company-radar

Competitive intelligence orchestrator tracking companies across 8+ platforms (GitHub, Twitter, Reddit, HN, PH, YC Jobs) with heat scores and AI briefings.

Varnan-Tech/opendirectory · 39 tokens

linkedin-job-post-to-buyer-pain-map

Takes pasted LinkedIn job posts or hiring descriptions and converts them into a structured buyer pain map with inferred pains, capability gaps, buy-vs-build signal, account priority scores, and suggested outreach angles. Use when asked to analyze hiring posts, decode job descriptions for buyer intent, build a pain map…

Varnan-Tech/opendirectory · 134 tokens