LLM agents

2,162 tagged LLM, measured the same way as everything else here.

Browse within: Multi-Agent 233Autonomous Agents 94claude-code-plugin 82prompt-engineering 73human-in-the-loop 63multi-agent-systems 60agentic 59drone 56gazebo 56copilot 52knowledge-graph 51local-first 51ai-security-tool 50offensive-security 50

scout

145

EXXETA/exxperts

Agent

Fast codebase recon that returns compressed context for handoff to other agents.

not rated 353 4d ago A 17 tokens copy · 100% Apache-2.0

audit-deep

146

VasiHemanth/tokentelemetry

Agent Claude Code

Deep single-subsystem audit that reasons about state over time (caches, upserts, migrations, concurrent scans) rather than pattern-matching lines. Used by /bug-audit on the riskiest subsystems; runs on Opus for reasoning depth.

not rated 343 +5 today A 56 tokens original MIT

audit-scanner

147

VasiHemanth/tokentelemetry

Agent Claude Code

Fast, wide sweep of one audit dimension across the codebase. Returns candidate findings with file:line evidence for the verifier to confirm. Used by /bug-audit; runs on Sonnet for breadth per token.

not rated 343 +5 today A 47 tokens original MIT

audit-verifier

148

VasiHemanth/tokentelemetry

Agent Claude Code

Adversarially verifies one candidate finding from /bug-audit — tries to REFUTE it by reading the code and, where cheap, reproducing it with a throwaway script. Kills false positives before they reach the report. Runs on Opus.

not rated 343 +5 today A 56 tokens original MIT

CHEATSHEET

149

cognesy/instructor-php

Agent

Agent loop, state model, tools, context management, hooks, and subagent orchestration.

not rated 326 2d ago A 18 tokens original MIT

default

150

cognesy/instructor-php

Agent

A lightweight coding agent for one non-interactive turn.

not rated 326 2d ago A 12 tokens original MIT

hermes

151

zorost/AI-Engineering-Lab

Agent

"Hermes" is overloaded, and getting it wrong causes real confusion. In an agents context it refers to two related-but-distinct things from Nous Research.

not rated 304 +71 17d ago A 0 tokens original MIT

openclaw

152

zorost/AI-Engineering-Lab

Agent

OpenClaw is the personal-AI-assistant track of Week 17: an open-source assistant that runs on your devices and meets you in the messaging channels you already use. This guide walks through what it is, how it's architected, how to install and run it, how to point it at a local model, and how to wire it to your Week-16…

not rated 304 +71 17d ago C 0 tokens original MIT

contract-review

153

Intelligent-Internet/zenith

Agent

Read-only adversarial contract reviewer. Reviews the full contract set against user scope, inventory, playbook rules, evidence feasibility, shortcut risk, and old-harness-style atomic assertion coverage before tasks are trusted.

not rated 283 +1 28d ago A 44 tokens original Apache-2.0

feature-reviewer

154

Intelligent-Internet/zenith

Agent

Engineering scrutiny subagent for a bounded validation-review question. Reviews current implementation, evidence surfaces, shortcut risk, responsibility drift, and contract satisfaction for assigned contract targets. Parent validator decides.

not rated 283 +1 28d ago A 40 tokens original Apache-2.0

flow-validator

155

Intelligent-Internet/zenith

Agent

Leaf real-surface validation lane for a bounded subset of engineering assertions. Exercises assigned behavior through a parent-specified browser, API, CLI, background, artifact, data, library, parity, or caller-provided tool surface; writes evidence only to assigned paths.

not rated 283 +1 28d ago A 55 tokens original Apache-2.0

securityclaw

156

SecurityClaw/SecurityClaw

Agent

Use for SecurityClaw orchestration, routing, skill-manifest, and investigation workflow changes.

not rated 270 +1 yesterday A 23 tokens original MIT

code-reviewer

157

pgEdge/pgedge-postgres-mcp

Agent Claude Code

Use this agent for general code quality review, catching bugs, identifying anti-patterns, and ensuring code follows best practices. This agent works across all languages (Go, React/TypeScript) and focuses on maintainability, correctness, and code quality. Examples:\n\n \nContext: Developer has written new…

not rated 221 +1 9d ago A 0 tokens PostgreSQL

documentation-writer

158

pgEdge/pgedge-postgres-mcp

Agent Claude Code

Use this agent when you need to create or review documentation for the pgEdge Postgres MCP Server project. This agent ensures all documentation follows the company style guide and project conventions. Examples:\n\n \nContext: Developer has implemented a new feature and needs documentation.\nuser: "I've added a new…

not rated 221 +1 9d ago A 0 tokens PostgreSQL

pgEdge/pgedge-postgres-mcp

Agent Claude Code

Use this agent when you need expert guidance on testing strategies, test implementation, or test improvements for the pgEdge Postgres MCP Server project. Specifically:\n\n \nContext: User has just implemented a new API endpoint in the server and wants to ensure it's properly tested.\nUser: "I've added a new endpoint…

not rated 221 +1 9d ago A 0 tokens PostgreSQL

implement-agent

160

dilolabs/nosia

Agent Claude Code

Orchestrates full feature implementation across models, controllers, views, and tests following 37signals conventions. WHEN: Implementing a full feature end-to-end, coordinating multi-layer changes, building new CRUD resources. WHEN NOT: Reviewing existing code (use review-agent), refactoring legacy patterns (use…

not rated 212 26d ago A 74 tokens original MIT

refactoring-agent

161

dilolabs/nosia

Agent Claude Code

Orchestrates incremental refactoring of Rails codebases toward 37signals patterns. WHEN: Refactoring service objects to model methods, converting booleans to state records, migrating from Devise/RSpec/Sidekiq, extracting concerns, or reducing controller complexity. WHEN NOT: Building new features (use…

not rated 212 26d ago A 75 tokens original MIT

review-agent

162

dilolabs/nosia

Agent Claude Code

Reviews code for adherence to 37signals Rails conventions. Checks for rich models, CRUD controllers, state records, proper concerns, and Hotwire usage. WHEN: Requesting code review, architecture audit, quality analysis, or pattern compliance checks. WHEN NOT: Implementing features (use implement-agent), refactoring…

not rated 212 26d ago A 71 tokens original MIT

analyzer

163

spytensor/openmozi

Agent

Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.

not rated 220 +18 28d ago A 0 tokens copy · 100% MIT

comparator

164

spytensor/openmozi

Agent

Compare two outputs WITHOUT knowing which skill produced them.

not rated 220 +18 28d ago A 0 tokens copy · 100% MIT

grader

165

spytensor/openmozi

Agent

Evaluate expectations against an execution transcript and outputs.

not rated 220 +18 28d ago A 0 tokens copy · 100% MIT

agents

166

damianvtran/local-operator

Agent

Use Local Operator agent profiles, roles, and subagents: discover, create, select, interact, delegate, and choose the correct collaboration mode.

not rated 209 4d ago A 31 tokens original MIT

planner

167

weidu12123/Liyuan

Agent

Creates implementation plans from context and requirements.

not rated 209 +3 5d ago A 9 tokens

reviewer

168

weidu12123/Liyuan

Agent

Code review specialist for quality and security analysis.

not rated 209 +3 5d ago A 11 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: