EntityProcess

36 mods across 2 repositories, 26 stars between them.

agentv

01

EntityProcess/agentv

Plugin Claude Code

Evaluate and optimize AI agents.

15 1mo ago A tokens not measured original MIT

agentv AGENTS.md

02

EntityProcess/agentv

Instructions file CodexOpenCode

Instructions for EntityProcess/agentv, covering agentv agent guide, product direction, always-read rules, repo map and routing.

15 1mo ago A 3,184 tokens original MIT

agentv CLAUDE.md

03

EntityProcess/agentv

Instructions file

Instructions for EntityProcess/agentv, a project described as: Light-weight AI agent evaluation and optimization framework.

15 1mo ago A 5 tokens copy · 100% MIT

deploy-execute

05

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "execute a deployment", "run the deploy plan", or "deploy services". Reads deploy-plan.md and executes each step with health checks.

15 1mo ago A 40 tokens original MIT

deploy-plan

06

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "plan a deployment", "create a deploy plan", or "prepare release steps". Produces a deployment plan with rollback strategy.

15 1mo ago A 37 tokens original MIT

deploy-rollback

07

EntityProcess/agentv

Skill Claude CodeCodex

This skill should be used when asked to "rollback a deployment", "revert services", or "undo deploy". Reads deploy-plan.md and reverses completed steps.

15 1mo ago A 37 tokens original MIT

acme-deploy

08

EntityProcess/agentv

Skill Claude CodeCodex

Use when the user asks about deploying services, checking deployment status, rollback procedures, or release management at Acme Corp.

15 1mo ago A 28 tokens original MIT

csv-analyzer

09

EntityProcess/agentv

Skill Claude CodeCodex

Use when the user asks to analyze CSV data. MUST use the weighted revenue formula defined below.

15 1mo ago A 23 tokens original MIT

agentv

10

EntityProcess/agentv

Settings file Claude Code

Agent settings configuring permissions, plugins.

15 1mo ago A tokens not measured original MIT

agentic-engineering

11

EntityProcess/agentv

Plugin Claude Code

Design and review AI agent systems: architecture patterns, workflow design, and plugin quality review.

15 1mo ago A tokens not measured original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Use when designing an AI agent system, selecting agentic design patterns, planning multi-phase workflows, choosing between single-agent and multi-agent architectures, or when asked "what kind of agent should I build", "how should I structure this automation", "design an agent for X", or "which agentic pattern fits…

15 1mo ago A 70 tokens original MIT

agent-plugin-review

13

EntityProcess/agentv

Skill Claude CodeCodex

Use when reviewing an AI plugin pull request, auditing plugin quality before release, or when asked to "review a plugin PR", "review skills in this PR", "check plugin quality", or "review workflow architecture". Covers skill quality, structural linting, and workflow architecture review.

15 1mo ago A 60 tokens original MIT

agentv-dev

14

EntityProcess/agentv

Plugin Claude Code

AgentV CLI skills for evaluating, optimizing, and governing AI agents.

15 1mo ago A tokens not measured original MIT

agentv-dev

15

EntityProcess/agentv

Skill Claude CodeCodex

AgentV CLI skills for evaluating, optimizing, and governing AI agents. Triggers: run evals, benchmark agents, write evals, review evals, analyze traces, optimize prompts, governance linting. Covers: eval running, eval writing, eval review, trace analysis, description optimization, autoresearch, and governance…

15 1mo ago A 70 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Capture, optimize, and publish screenshots to Astro docs. Use when asked to take screenshots for docs, update doc images, compress PNG assets, or add visual documentation to the agentv.dev docs site. Triggers on "add screenshots to docs", "update docs images", "compress screenshots", "optimize PNG", "document with…

15 1mo ago B 75 tokens original MIT

agentv-bench

17

EntityProcess/agentv

Skill Claude CodeCodex

Run AgentV evaluations and optimize agents through eval-driven iteration. Triggers: run evals, benchmark agents, optimize prompts/skills against evals, compare agent outputs across providers, analyze eval results, offline evaluation of recorded sessions, run autoresearch, optimize unattended, run overnight…

15 1mo ago A 101 tokens original MIT

analyzer

18

EntityProcess/agentv

Agent

Analyze AgentV evaluation results to identify weak assertions, suggest deterministic upgrades for LLM-grader graders, flag cost/quality improvements, and surface cross-run benchmark patterns. Use when reviewing eval quality, improving evaluation configs, or triaging flaky/expensive evaluations.

15 1mo ago A 55 tokens original MIT

comparator

19

EntityProcess/agentv

Agent

Perform bias-free blind comparison of evaluation outputs from multiple providers or configurations. Randomizes labeling, generates task-specific rubrics, scores N-way comparisons, then unblinds results and attributes improvements. Dispatch this agent when comparing outputs across targets or iterations.

15 1mo ago A 52 tokens original MIT

executor

20

EntityProcess/agentv

Agent

Execute an AgentV evaluation test case by performing the task described in the input. Reads input.json from the test directory, carries out the task using available tools, and writes response.md with the result. Dispatch one executor subagent per test case, all in parallel.

15 1mo ago A 55 tokens original MIT

grader

21

EntityProcess/agentv

Agent

Grade a candidate response for an AgentV evaluation test case. Evaluates all assertion types natively — deterministic checks via string operations, LLM grading via Claude's own reasoning, script-grader via Bash script execution. Zero CLI dependency. Dispatch this agent after a candidate completes a test case.

15 1mo ago A 60 tokens original MIT

mutator

22

EntityProcess/agentv

Agent

Generate improved versions of the artifact under test (skill, prompt, config, or directory of related files) based on failure analysis. Reads the current best artifact from the working tree, applies targeted mutations to address failing assertions, and writes changes in place. Supports single files and multi-file…

15 1mo ago A 70 tokens original MIT

EntityProcess/agentv

Skill Claude CodeCodex

Migrate AgentV eval YAML across breaking schema changes, especially workspace contract updates and portable-vs-local runtime binding cleanup.

15 1mo ago A 31 tokens original MIT

agentv-eval-review

24

EntityProcess/agentv

Skill Claude CodeCodex

Use when reviewing eval YAML files for quality issues, linting eval files before committing, checking eval schema compliance, or when asked to "review these evals", "check eval quality", "lint eval files", or "validate eval structure". Do NOT use for writing evals (use agentv-eval-writer) or running evals (use…

15 1mo ago A 81 tokens original MIT