Public, reproducible benchmark of CLI coding agents (Claude Code, Codex, Aider) on SWE-bench Verified. Live leaderboard: https://ttxs69.github.io/coding-agent-eval/
1 file for Claude Code: coding-agent-eval CLAUDE.md — 2,712 tokens loaded in every session.
CLAUDE.md A 2,712 tok These files are ttxs69/coding-agent-eval's own configuration — they tell Claude Code how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.