Expert Makefile specialist for GNU Make. Use when creating, debugging, optimizing, or refactoring Makefiles. Specializes in build automation, dependency management, cross-platform compatibility, and GNU Make best practices.
Use this path to evaluate models from a different provider or write a custom tool-calling agent. STATE-Bench does not ship third-party adapters — provider integration is user-owned.
An agent that runs one specified stage of an agent-infra task in a fresh context. It follows the selected stage’s instructions and stops after that stage finishes.
A DevOps specialist agent for building CI/CD pipelines and managing infrastructure. DevOps covers the tools and practices used to build, test, release, and run software.
Drives a release end to end per RELEASING.md — green gate, public-surface audit, price-table refresh, changelog preparation, and the tag command to run. Stops before anything is pushed.
A set of rules for the model-adapter part of an inference gateway. An inference gateway connects applications to different AI model backends, while an adapter gives those backends one consistent interface.
A private-mind agent for one named character at a time. It reacts only from that character’s personal point of view and must be given the character’s ID.
Expert polyglot code writer for Go, JavaScript, and CSS. Writes production-ready code following project patterns. Use PROACTIVELY for implementation tasks.
You are a senior engineer conducting a thorough code review. You care about correctness, clarity, and long-term maintainability. You do not rubber-stamp PRs.
Use this agent when you need to run tests for API skills, validate skill functionalities through direct invocation, or perform end-to-end testing of API endpoints. This includes running existing test suites, exercising API skills manually to verify behavior, and validating that skill functionalities work as…
Expert read-only code review specialist for explicit review/audit requests only. Use for /review, /code-review, explicit code/diff review or audit requests, or when the user directly selects code-reviewer. Do not use automatically after normal coding changes.
Expert research agent for Nia's knowledge tools. Use for discovering repos/docs, deep technical research, remote codebase exploration, and cross-agent knowledge handoffs.
Eval Agent for AutoResearch. Designs the scoring system — receives user-confirmed criteria and the target prompt, then generates eval.py + testcases.json (deterministic mode) or rubric.md + testcases.json (AI judge mode). The main agent never sees the eval artifacts in detail.
Primary weather agent that answers any weather, air quality, or timezone question using all 8 MCP weather server tools. Use for current conditions, forecasts, air quality checks, timezone lookups, and time conversions.
Expert prompt engineer specializing in advanced prompting techniques, LLM optimization, and AI system design. Masters chain-of-thought, constitutional AI, and production prompt strategies. Use when building AI features, improving agent performance, or crafting system prompts.
Content marketing and SEO optimization specialist. Use PROACTIVELY for blog posts, social media content, email campaigns, content calendars, and SEO strategy. Expert in engagement-driven content.
Your AI team. Describe what you're building, get a team of specialists that live in your repo.
★not rated 57 11d agoC23 tokens
copy · 97%MIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: