Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add mcp/hidai25/eval-view/evalviewgit clone --depth 1 https://github.com/hidai25/eval-viewGrade A, and why
evalview scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
{
"evalview": {
"command": "evalview",
"args": [
"mcp",
"serve"
],
"env": {
"OPENAI_API_KEY": "${OPENAI_API_KEY}",
"ANTHROPIC_API_KEY": "${ANTHROPIC_API_KEY}"
}
}
}What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 13 lines scan A 2da3b891916b
evalview is an MCP server published in the GitHub repository hidai25/eval-view (132 stars, last pushed 9d ago), licensed Apache-2.0. Its token cost is not measured: an MCP server costs its tool schemas, not its config file. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other mcp servers, from other repositories
portable-agent-memory
MCP server "portable-agent-memory" as configured in skywaller0/Portable-Agent-Memory. Runs locally from the portable-agent-memory Python package.
radmeasure
MCP server "radmeasure" as configured in jianghongcheng/radmeasure-agent. Runs locally from the radmeasure Python package.
agent-eval
Statistical regression testing for LLM agents. Get a p-value on whether behavior actually shifted between versions, not just whether one run looked different. Apache 2.0, self-hostable, no SaaS dependency -- a Promptfoo alternative for the statistical testing gap threshold-based eval tools don't cover. Runs locally…
ai-ticket-triage
MCP server "ai-ticket-triage" as configured in Mohemed-Amine-Chalhy/ai-ticket-triage. Runs locally from the ai-ticket-triage Python package.
engramia
Reusable execution memory and evaluation infrastructure for AI agent frameworks. Runs locally from the engramia Python package. Needs 8 environment variables to run.
lastsearch
Evidence-backed web research for AI agents with citations and confidence scores. Runs locally from the lastsearch npm package. Needs 1 environment variable to run.