hidai25/eval-view

Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.

133Stars on the repository
11Mods indexed here, across every type
9d agoLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

hidai25/eval-view

Skill Claude CodeCodex

Beat procrastination with task breakdown, 2-minute starts, and accountability tracking.

133 9d ago A 21 tokens original Apache-2.0

code-reviewer

02

hidai25/eval-view

Skill Claude CodeCodex

Performs comprehensive code reviews with security, quality, and best practice checks.

133 9d ago A 18 tokens original Apache-2.0

hello-world

03

hidai25/eval-view

Skill Claude CodeCodex

A simple skill that creates a greeting file.

133 9d ago A 11 tokens original Apache-2.0

code-reviewer

04

hidai25/eval-view

Skill Claude CodeCodex

A skill that helps review code for best practices, bugs, and security issues.

133 9d ago A 19 tokens original Apache-2.0

generate-tests

05

hidai25/eval-view

Skill Claude CodeCodex

Generate EvalView test cases — either from a SKILL.md file using LLM-powered generation, or by capturing real agent interactions through a proxy.

133 9d ago A 32 tokens original Apache-2.0

run-eval

06

hidai25/eval-view

Skill Claude CodeCodex

Run EvalView regression checks against golden baselines to detect regressions in AI agent behavior after code, prompt, or model changes.

133 9d ago A 30 tokens original Apache-2.0

watch

07

hidai25/eval-view

Skill Claude CodeCodex

Start EvalView watch mode to automatically re-run regression checks whenever project files change.

133 9d ago A 18 tokens original Apache-2.0