eval-view
01Plugin Claude Code
Open-source testing and regression detection framework for AI agents. Golden baseline diffing, CI/CD integration, works with LangGraph, CrewAI, OpenAI, Anthropic Claude, HuggingFace, Ollama, and MCP.
Plugin Claude Code
Open-source testing and regression detection framework for AI agents. Golden baseline diffing, CI/CD integration, works with LangGraph, CrewAI, OpenAI, Anthropic Claude, HuggingFace, Ollama, and MCP.
MCP server Claude CodeCodexCursor +2
MCP server "evalview" as configured in hidai25/eval-view. Launched with evalview mcp serve. Needs 2 environment variables to run.
Instructions file CodexOpenCode
Instructions for hidai25/eval-view, covering evalview agent instructions, what evalview is, core concepts, testcase and evaluationresult.
Skill Claude CodeCodex
Beat procrastination with task breakdown, 2-minute starts, and accountability tracking.
Skill Claude CodeCodex
Performs comprehensive code reviews with security, quality, and best practice checks.
Skill Claude CodeCodex
A simple skill that creates a greeting file.
Skill Claude CodeCodex
A skill that helps review code for best practices, bugs, and security issues.
Skill Claude CodeCodex
Generate EvalView test cases — either from a SKILL.md file using LLM-powered generation, or by capturing real agent interactions through a proxy.
Skill Claude CodeCodex
Run EvalView regression checks against golden baselines to detect regressions in AI agent behavior after code, prompt, or model changes.
Skill Claude CodeCodex
Start EvalView watch mode to automatically re-run regression checks whenever project files change.
MCP server Claude CodeCodexCursor +2
Open-source testing and regression detection framework for AI agents. Golden baseline diffing, CI/CD integration, works with LangGraph, CrewAI, OpenAI, Anthropic Claude, HuggingFace, Ollama, and MCP. Runs locally from the evalview Python package. Needs 1 environment variable to run.