code quality agents

732 tagged code quality, measured the same way as everything else here.

Browse within: development-workflow 57architecture 55ai-pipeline 53autonomous-coding 53autofix 38ci-cd 37developer-productivity 37Multi-Agent 34ai-code-review 33AI Safety 30contracts 30agent-plugins 29antigravity 29bug-hunter 29

harness-evaluator

55

stone16/harness-engineering-skills

Agent

Harness Evaluator — independent code evaluation with Tier 1 deterministic checks and Tier 2 deep logic analysis. Use when harness orchestrator needs checkpoint evaluation.

32 1mo ago A 35 tokens original Apache-2.0

harness-retro

56

stone16/harness-engineering-skills

Agent

Harness Retro — post-task retrospective analysis, error pattern detection, and CLAUDE.md rule proposals. Use when harness orchestrator needs task retrospective.

32 1mo ago A 34 tokens original Apache-2.0

stone16/harness-engineering-skills

Agent

Harness Spec Evaluator — reviews spec.md for checkpoint quality, architectural feasibility, and cybernetic completeness. Use when harness orchestrator needs spec evaluation before execution.

32 1mo ago A 37 tokens original Apache-2.0

executor-v3

58

malakhov-dmitrii/forge

Agent

TDD implementation agent for forge pipeline-v3. Spawned per-stream; enforces RED-GREEN-REFACTOR with retry protocol and phase-2 short-circuit.

25 1mo ago A 39 tokens original MIT

planner-v3

59

malakhov-dmitrii/forge

Agent

Creates bite-sized, TDD-embedded, one-shot-executable implementation plans with DAG emission, claim verification fan-out, and overlap-matrix self-check. Produces plans that a fresh Claude session can execute without questions.

25 1mo ago A 48 tokens original MIT

skeptic

60

malakhov-dmitrii/forge

Agent

Mirage detection specialist for beast-plan. Verifies plan claims against codebase reality and external facts. Catches assumptions masquerading as facts.

25 1mo ago A 31 tokens original MIT

conflict-arbiter

61

nguyenthienthanh/aura-frog

Agent

Adjudicates detected conflicts between plan-tree tasks. Decides freeze | sequential | replan | escalate per spec §21.5. Read-only on code; writes only to .claude/plans/conflicts.jsonl + history.jsonl.

24 8d ago A 55 tokens original MIT

epic-summarizer

62

nguyenthienthanh/aura-frog

Agent

Distills a completed Epic (T2 Feature done) into a permanentmemory.md section. Captures architectural decisions, gotchas, anti-patterns, conflicts. Writes ONLY to .claude/memory/. Confidence-scored: items below 0.7 land in a Tentative subsection.

24 8d ago A 64 tokens original MIT

frontend

63

nguyenthienthanh/aura-frog

Agent

Frontend frameworks (React/Vue/Angular/Next.js), design systems, accessibility. Use for UI implementation, component work, and responsive design.

24 8d ago A 31 tokens original MIT

Hulupeep/Specflow

Agent

You are a Jest test generator for YAML contracts. You read docs/contracts/.yml files and generate corresponding test files in src/tests/contracts/ that enforce the contracts through pattern scanning at build time.

24 1mo ago A 0 tokens original MIT

specflow-writer

65

Hulupeep/Specflow

Agent

You are a full-stack specflow architect. You produce production-grade ticket specs that combine BDD scenarios, data contracts, UI behaviour, and acceptance criteria into a single source of truth — so that migration-builder, edge-function-builder, and playwright-from-specflow agents can execute without ambiguity.

24 1mo ago A 0 tokens original MIT

waves-controller

66

Hulupeep/Specflow

Agent

You are a wave execution orchestrator. You take a GitHub project board (or list of issues) and execute them in dependency-ordered waves with full contract compliance, testing, and validation. You coordinate all other Specflow agents through an 8-phase workflow.

24 1mo ago A 0 tokens original MIT

hyhmrright/logic-lens

Agent Claude Code

Analyze Logic-Lens benchmark/eval failures. Use after running content-evals, or when pointed at a skills-workspace/iteration- directory or a benchmarks/runs/ entry, to cluster failing cases by failure mode, map each mode to the specific eval IDs, and propose concrete SKILL.md disambiguation-rule changes. Read-only…

21 3d ago A 90 tokens original MIT

iteration-guard

68

hyhmrright/logic-lens

Agent Claude Code

The verify gate of the Logic-Lens iteration loop. Given a baseline iteration and a candidate iteration, compares their summary.json (overall, logic vs format subscores, per-mode, per-language), accounts for single-run variance, and returns a SHIP / ROLLBACK / RERUN recommendation with evidence. Use after…

21 3d ago A 99 tokens original MIT

skill-editor

69

hyhmrright/logic-lens

Agent Claude Code

Applies a single, minimal, generalized edit to a Logic-Lens skill (SKILL.md / guide / shared file) given a concrete failure diagnosis. Use inside the iteration loop after eval-failure-analyzer has produced a proposal, to turn that proposal into an actual edit. Mutates files; does NOT run evals or sync the cache — it…

21 3d ago A 84 tokens original MIT

cove-executor

70

vertti/se-cove-claude-plugin

Agent

You are the Independent Verification Executor in a Software Engineering Chain of Verification (SE-CoVe) system.

20 7mo ago A 0 tokens original MIT

cove-planner

71

vertti/se-cove-claude-plugin

Agent

You are the Verification Task Planner in a Software Engineering Chain of Verification (SE-CoVe) system.

20 7mo ago A 0 tokens original MIT