Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add mcp/hidai25/eval-view/evalview-mcpgit clone --depth 1 https://github.com/hidai25/eval-viewGrade A, and why
evalview-mcp scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
{
"evalview-mcp": {
"command": "uvx",
"args": [
"evalview"
],
"env": {
"OPENAI_API_KEY": ""
}
}
}What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 11 lines scan A bed0209b457f
evalview-mcp is an MCP server published in the GitHub repository hidai25/eval-view (133 stars, last pushed 10d ago), licensed Apache-2.0. Its token cost is not measured: an MCP server costs its tool schemas, not its config file. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other mcp servers, from other repositories
portable-agent-memory
MCP server "portable-agent-memory" as configured in skywaller0/Portable-Agent-Memory. Runs locally from the portable-agent-memory Python package.
radmeasure
MCP server "radmeasure" as configured in jianghongcheng/radmeasure-agent. Runs locally from the radmeasure Python package.
agent-eval
Statistical regression testing for LLM agents. Get a p-value on whether behavior actually shifted between versions, not just whether one run looked different. Apache 2.0, self-hostable, no SaaS dependency -- a Promptfoo alternative for the statistical testing gap threshold-based eval tools don't cover. Runs locally…
ai-ticket-triage
MCP server "ai-ticket-triage" as configured in Mohemed-Amine-Chalhy/ai-ticket-triage. Runs locally from the ai-ticket-triage Python package.
engramia
Reusable execution memory and evaluation infrastructure for AI agent frameworks. Runs locally from the engramia Python package. Needs 8 environment variables to run.
network-ai
AI agent orchestration framework for TypeScript/Node.js - 32 adapters (LangChain, AutoGen, CrewAI, OpenAI Assistants, OpenAI Responses, LlamaIndex, Semantic Kernel, Haystack, DSPy, Agno, MCP, OpenClaw, A2A, Codex, MiniMax, NemoClaw, APS, Copilot, LangGrap. Runs locally from the network-ai npm package.