Real-world browser-agent benchmark: 210 tasks across 107 websites, multi-agent/multi-browser evaluation, reproducible leaderboard and result submissions.
These files are lexmount/browseruse-agent-bench's own configuration. They tell GitHub Copilot, Codex, OpenCode and Claude Code how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.
.github/copilot-instructions.md A 3,057 tok AGENTS.md A 386 tok CLAUDE.md A 1,419 tok