llm-as-judge agents

1 tagged llm-as-judge, measured the same way as everything else here.

Browse within: arc42 8claude-code-plugin 8legacy-modernization 8

homemade-software-inc/completion-kit

Agent Claude Code

Assesses a pull request against CompletionKit's merge bar — is it worth merging at all, is it secure, is the code excellent, is it as simple as possible, and does it pass the project's hard gates (CI, 100% coverage, the inline test-schema gotcha, conventions). Use when triaging or reviewing an incoming PR, especially…

not rated 3 23d ago A 90 tokens