Agent
These notes track the planned Promptfoo integration for the Codex app-server protocol. They are intentionally implementation-facing: keep them current as the provider, docs, examples, and verification expand.
57 tagged llmops, measured the same way as everything else here.
Browse within: Autonomous Agents 23ai-engineering 23compilers 9gpu-computing 9inference 9llm-inference 9prompt-engineering 9prompt-testing 9prompts 9api 5large-language-models 5llms 5
Agent
These notes track the planned Promptfoo integration for the Codex app-server protocol. They are intentionally implementation-facing: keep them current as the provider, docs, examples, and verification expand.
Agent
This codebase uses Drizzle ORM with SQLite. All database queries must use parameterized SQL so user-controlled input never changes query structure.
Agent
PR titles follow Conventional Commits format. They become squash-merge commit messages and changelog entries.
GoogleCloudPlatform/agent-starter-pack
Agent
The Agent Starter Pack follows a "bring your own agent" approach. It provides several production-ready agent templates designed to accelerate your development while offering the flexibility to use your preferred agent framework or pattern.
Agent Claude Code
Read-only code review agent. Catches type issues, pattern violations, missing tests. Cannot modify files.
Agent Claude Code
Run tests, analyze failures, suggest fixes. Use after code changes.
Agent Claude Code
Generates changelogs from git commit history and conversation context. Must be used immediately when the user asks to bump the version.
Agent Claude Code
name: integration-test-engineer description: Use this agent when you need to create, modify, or debug integration tests in the crates/integration-tests directory. This includes writing new test scenarios, updating existing tests, working with Docker Compose configurations for test environments, handling authentication…
Agent Claude Code
name: mcp-crate-engineer description: Use this agent when editing any files within the crates/mcp directory. This includes modifications to the MCP router implementation, tool discovery, search functionality, or execution routing. The agent should be automatically triggered for any file changes in this directory to…
Agent Claude Code
Senior Architect reviewer. Reviews plans and diffs for simplicity, duplication, encapsulation and abstraction. Read-only — never edits code. Use before opening a PR, or when a design decision needs a second opinion.
Agent
Refresh the Emmy model lifecycle through bounded, read-only research.
Agent
Qualify one model on the workflow-owned GPU and produce reviewed Emmy artifacts.
Agent
Generates monthly customer newsletters via a skills-based pipeline. Orchestrates 6 phases from URL discovery through final assembly and validation.
Agent
Analyzes published newsletters and benchmark intermediates to extract editorial patterns, selection decisions, and thematic intelligence. Use for editorial intelligence mining.
Agent
Builds individual newsletter pipeline skills from agent logic, prompt files, and benchmark examples. Use with /fleet for parallel skill construction.
pillaiharish/opencode-ollama-steroids
Agent
You are the strict reviewer agent for a receipt-first OpenCode + Ollama workflow.
pillaiharish/opencode-ollama-steroids
Agent
You are the builder agent for a receipt-first OpenCode + Ollama workflow.
pillaiharish/opencode-ollama-steroids
Agent
You are a controlled tool-call compatibility probe.
Agent Claude Code
This directory contains the 12 Claude Code agents that implement the SpecRoute skeleton. They are project-internal - they build SpecRoute itself, not consumer-facing templates that ship as artifacts.
Agent Claude Code
Use when checking whether SpecRoute's claims about the outside world are still true - vendor config surfaces, hook event taxonomies, frontmatter contracts, version anchors, transition dates, and counts that must match disk. Owns the currency cycle, not the prose. Runs before releases, after any vendor ships a breaking…
Agent Claude Code
Use proactively before any commit and whenever new content is added. Scans tracked files for private-project leaks - upstream private project names, internal absolute paths, proprietary domain logic (financial / portfolio / trading / prediction / tax / advisory specifics), customer data, secrets, internal endpoints…