EntityProcess/agentv

Light-weight AI agent evaluation and optimization framework

15Stars on the repository
27Mods indexed here, across every type
1mo agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

deploy-execute

01

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "execute a deployment", "run the deploy plan", or "deploy services". Reads deploy-plan.md and executes each step with health checks.

15 1mo ago A 40 tokens original MIT

deploy-plan

02

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "plan a deployment", "create a deploy plan", or "prepare release steps". Produces a deployment plan with rollback strategy.

15 1mo ago A 37 tokens original MIT

deploy-rollback

03

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "rollback a deployment", "revert services", or "undo deploy". Reads deploy-plan.md and reverses completed steps.

15 1mo ago A 37 tokens original MIT

acme-deploy

04

EntityProcess/agentv

Skill Claude CodeCodex

Use when the user asks about deploying services, checking deployment status, rollback procedures, or release management at Acme Corp.

15 1mo ago A 28 tokens original MIT

csv-analyzer

05

EntityProcess/agentv

Skill Claude CodeCodex

Use when the user asks to analyze CSV data. MUST use the weighted revenue formula defined below.

15 1mo ago A 23 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Use when designing an AI agent system, selecting agentic design patterns, planning multi-phase workflows, choosing between single-agent and multi-agent architectures, or when asked "what kind of agent should I build", "how should I structure this automation", "design an agent for X", or "which agentic pattern fits…

15 1mo ago A 70 tokens original MIT

agent-plugin-review

07

EntityProcess/agentv

Skill Claude CodeCodex

Use when reviewing an AI plugin pull request, auditing plugin quality before release, or when asked to "review a plugin PR", "review skills in this PR", "check plugin quality", or "review workflow architecture". Covers skill quality, structural linting, and workflow architecture review.

15 1mo ago A 60 tokens original MIT

agentv-dev

08

EntityProcess/agentv

Skill Claude CodeCodex

AgentV CLI skills for evaluating, optimizing, and governing AI agents. Triggers: run evals, benchmark agents, write evals, review evals, analyze traces, optimize prompts, governance linting. Covers: eval running, eval writing, eval review, trace analysis, description optimization, autoresearch, and governance…

15 1mo ago A 70 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Capture, optimize, and publish screenshots to Astro docs. Use when asked to take screenshots for docs, update doc images, compress PNG assets, or add visual documentation to the agentv.dev docs site. Triggers on "add screenshots to docs", "update docs images", "compress screenshots", "optimize PNG", "document with…

15 1mo ago B 75 tokens original MIT

agentv-bench

10

EntityProcess/agentv

Skill Claude CodeCodex

Run AgentV evaluations and optimize agents through eval-driven iteration. Triggers: run evals, benchmark agents, optimize prompts/skills against evals, compare agent outputs across providers, analyze eval results, offline evaluation of recorded sessions, run autoresearch, optimize unattended, run overnight…

15 1mo ago A 101 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Migrate AgentV eval YAML across breaking schema changes, especially workspace contract updates and portable-vs-local runtime binding cleanup.

15 1mo ago A 31 tokens original MIT

agentv-eval-review

12

EntityProcess/agentv

Skill Claude CodeCodex

Use when reviewing eval YAML files for quality issues, linting eval files before committing, checking eval schema compliance, or when asked to "review these evals", "check eval quality", "lint eval files", or "validate eval structure". Do NOT use for writing evals (use agentv-eval-writer) or running evals (use…

15 1mo ago A 81 tokens original MIT

agentv-eval-writer

13

EntityProcess/agentv

Skill Claude CodeCodex

Write, edit, review, and validate AgentV EVAL.yaml / .eval.yaml evaluation files. Use when asked to create new eval files, update or fix existing ones, add or remove test cases, configure graders (llm-rubric, script), review whether an eval is correct or complete, convert between EVAL.yaml and evals.json using agentv…

15 1mo ago A 129 tokens original MIT

agentv-governance

14

EntityProcess/agentv

Skill Claude CodeCodex

Author, edit, and lint governance: blocks in .eval.yaml files. Use when creating or updating evaluation suites that carry AI-governance metadata (OWASP LLM Top 10, OWASP Agentic Top 10, MITRE ATLAS, EU AI Act, ISO 42001). Also use non-interactively (e.g., from a GitHub Action) to lint changed eval files and report…

15 1mo ago A 124 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Analyze AgentV evaluation traces and result JSONL files using agentv inspect and agentv results compare CLI commands. Use when asked to inspect AgentV eval results, find regressions between AgentV evaluation runs, identify failure patterns in AgentV trace data, analyze tool trajectories, or compute cost/latency/score…

15 1mo ago A 113 tokens original MIT