Instructions file GitHub Copilot
Guidelines for working with AgentEval Memory module — benchmarks, reporting, HTML reports, and LongMemEval integration.
860 tagged Testing, measured the same way as everything else here.
Browse within: playwright 24github-copilot 23agentic 22api-testing 22copilot 21code-quality 17dotnet 15e2e 15evals 15browser-automation 12coding-agents 12ios 12microsoft 11Evaluation 10
Instructions file GitHub Copilot
Guidelines for working with AgentEval Memory module — benchmarks, reporting, HTML reports, and LongMemEval integration.
Instructions file CodexOpenCode
Instructions for 1Password/SCAM, covering agents.md — ai coding guidelines for scam, project overview, architecture, module responsibilities and directory layout.
Instructions file CodexOpenCode
Instructions for flop-labs/technocore-chat: CI runs exactly these — run them before pushing.
Instructions file CodexOpenCode
Instructions for kbrdn1/gwm-cli, covering gwm-cli — house rules for ai assistants, 🔴 primordial rule — test-driven development is mandatory, the tdd loop (red → green → refactor), what counts as "behaviour" and exceptions (narrow, must be argued in the pr description).
Instructions file
Instructions for kbrdn1/gwm-cli, covering gwm-cli — house rules for ai assistants, 🔴 primordial rule — test-driven development is mandatory, the tdd loop (red → green → refactor), what counts as "behaviour" and exceptions (narrow, must be argued in the pr description).
Instructions file CodexOpenCode
Instructions for hidai25/eval-view, covering evalview agent instructions, what evalview is, core concepts, testcase and evaluationresult.
NikiforovAll/github-copilot-rules
Instructions file GitHub Copilot
This file provides guidelines for writing effective, maintainable tests using xUnit and related tools.
anhtester/antigravity-testing-kit
Instructions file Gemini CLI
Instructions for anhtester/antigravity-testing-kit, covering gemini ai - global automation agent rules, git pull restriction rule, browser rules (mandatory), 🖥️ viewport & mode and 🔄 thứ tự debug bắt buộc (playwright mcp).
Instructions file CodexOpenCode
Instructions for minghinmatthewlam/openbench, covering openbench — agent context, local execution context, what openbench is, execution ownership and product goals (the two things we are building toward).
Instructions file CodexOpenCode
Instructions for ikamensh/kodo, covering rules for writing code in this repo, end-to-end test scenarios (mocked), boundary condition tests (final status), rule a.0 and rule a.1.
Instructions file
Claude Code instructions for litegraphdb/litegraph, covering claude.md, build and development commands, build the entire solution, build specific projects and run tests.
Instructions file
Instructions for benseverndev-oss/goldenmatch, covering golden suite monorepo, typescript: pnpm + turborepo (post-2026-05-02 fold), ci (.github/workflows/ci.yml), ci path filters (post-2026-05-06, pr #89) and merge queue: main serializes merges fifo (since 2026-06-15).
Instructions file CodexOpenCode
Instructions for duane1024/l123, covering agents.md — l123 project guide, canonical docs — read these before making decisions, how we develop: red / green / refactor (strict tdd), 1. red and 2. green.
Instructions file CodexOpenCode ✓ vendor
Instructions for microsoft/eval-guide, covering eval guide — ai agent evaluation toolkit, what this toolkit does, available prompt files, routing guide and methodology summary.
Instructions file CodexOpenCode
Instructions for scorbo2/TalkWithMe, covering talkwithme — agent instructions, run the app, testing — pytest, fully offline, config — three yaml files, cached at startup and architecture.
Instructions file
Instructions for nanasess/setup-chromedriver, covering claude.md, commands, build and development, running a single test and or.
Instructions file CodexOpenCode
Instructions for sstraus/tuicommander, covering tuicommander — project rules, doc sync, tests, web-ui testing with agent-browser (browser mode, not tauri) and visual.
Instructions file
Instructions for UiPath/coder_eval, covering claude.md - ai assistant guide, project overview, directory structure, key architectural patterns and success criteria (15 types).
langchain-ai/skills-benchmarks
Instructions file ✓ vendor
Instructions for langchain-ai/skills-benchmarks, covering skills project guidelines, python/typescript parity, skill markdown file parity, task structure and benchmark validation principles.
Instructions file CodexOpenCode
Instructions for helderberto/dotfiles, covering agents.md, before coding, surgical changes, verification and code principles.
Tahanima/playwright-java-test-automation-architecture
Instructions file Gemini CLI
Instructions for Tahanima/playwright-java-test-automation-architecture, covering project: playwright java test automation architecture, context & tech stack, project mapping, development rules and locator strategy (priority order).
Instructions file CodexOpenCode
Instructions for spences10/sveltest, covering agents.md, unbreakable rules, essential commands, development and testing.
Instructions file
Instructions for spences10/sveltest, a project described as: Comprehensive Svelte 5 testing examples and patterns using Vitest Browser Mode, Playwright, and vitest-browser-svelte.
prime-radiant-inc/superpowers-evals
Instructions file CodexOpenCode
AGENTS.md instructions for prime-radiant-inc/superpowers-evals, a project described as: Behavioral eval lab (Quorum) for the superpowers project that drives real coding-agent CLIs (Claude, Codex, Gemini, Kimi, and more) through a QA agent and grades them on workflow compliance against scenario criteria and…
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: