LLM agents

2,270 tagged LLM, measured the same way as everything else here.

Browse within: Multi-Agent 250Autonomous Agents 94prompt-engineering 81human-in-the-loop 63multi-agent-systems 60swarm-intelligence 60drone 56gazebo 56agentic 54ai-security-tool 50offensive-security 50code 49local-first 48copilot 47

vuln_review

49

Tencent/AI-Infra-Guard

Agent ✓ vendor

A security-review agent that checks vulnerability reports for real, reproducible threats and filters out false positives. A false positive is an alleged problem that is not actually exploitable or harmful in the stated environment.

6.1k +24 2d ago A 0 tokens original Apache-2.0

algorithm-expert

51

areal-project/AReaL

Agent Claude Code

RL algorithm expert. Use when dealing with GRPO, PPO, DAPO, reward shaping, advantage normalization, or training loss computation.

5.7k +4 yesterday A 32 tokens original Apache-2.0

code-verifier

52

areal-project/AReaL

Agent Claude Code

Code verification agent. Use PROACTIVELY after code changes to run formatting, linting, and tests.

5.7k +4 yesterday A 26 tokens original Apache-2.0

areal-project/AReaL

Agent Claude Code

Expert on cluster launching and resource scheduling (Slurm/Ray/Kubernetes). Use when user modifies launcher/scheduler code, configures cluster resources, or troubleshoots deployment issues.

5.7k +4 yesterday A 43 tokens original Apache-2.0

custom-component

54

homeassistant-ai/ha-mcp

Agent

Read this document before changing customcomponents/hamcptools/ or a server feature that depends on the component. The HACS component and the ha-mcp server ship through separate installation paths, so compatibility must hold in both update directions.

4.6k +29 yesterday A 0 tokens original MIT

development

55

homeassistant-ai/ha-mcp

Agent

Read this document when setting up the repository, running the server, choosing verification commands, researching Home Assistant APIs, or orienting yourself in the architecture. Behavioral rules about when testing is required remain in AGENTS.md; test implementation details live in tests/AGENTS.md.

4.6k +29 yesterday A 0 tokens original MIT

github-workflow

56

homeassistant-ai/ha-mcp

Agent

Read this document before triaging issues, changing GitHub automation, managing a pull request, or preparing a release. Repository-wide behavioral rules remain in AGENTS.md; this file owns the detailed commands, labels, bot behavior, and workflow inventory.

4.6k +29 yesterday A 0 tokens original MIT

nanoclaw

57

openagents-org/openagents

Agent

NanoClaw lets a NanoClaw Agent Group act as an OpenAgents Workspace agent. Unlike most OpenAgents agents, NanoClaw is not a stdin/stdout CLI and not a direct LLM API — it is an independent containerized agent runtime. Each Agent Group runs in its own Docker container (Apple Container on macOS; WSL2 on Windows), with…

4.0k +9 2d ago A 0 tokens original Apache-2.0

pi

58

openagents-org/openagents

Agent

OpenAgents supports Pi, Earendil's coding agent for the terminal — executable pi, distributed as the npm package @earendil-works/pi-coding-agent.

4.0k +9 2d ago A 0 tokens original Apache-2.0

scout

62

claraverse-space/ClaraVerse

Agent

Fast codebase recon that returns compressed context for handoff to other agents.

3.9k 1mo ago A 17 tokens

claude-code

63

raullenchai/Rapid-MLX

Agent

Point Anthropic's Claude Code at a local rapid-mlx server. Claude Code speaks the Anthropic Messages API (POST /v1/messages); rapid-mlx implements that route natively, so you can drive Claude Code with any local model.

3.6k +45 yesterday A 0 tokens

hermes-agent

64

raullenchai/Rapid-MLX

Agent

Point Nous Research's Hermes Agent at a local rapid-mlx server. Hermes is a tool-heavy CLI agent (it injects up to 62 tools per request) that speaks the OpenAI-compatible chat completions API (POST /v1/chat/completions).

3.6k +45 yesterday A 0 tokens

matrix

65

raullenchai/Rapid-MLX

Agent

This page renders the Tier-1 agent × model-family integration matrix truthfully from the authoritative test suite in tests/integrations/ — the matrix cells (testagentsmatrix.py), the family aliases and strict-xfail rules (conftest.py), and the pilot run recorded in tests/integrations/README.md.

3.6k +45 yesterday A 0 tokens

AUDIT_MANIFEST

66

langwatch/langwatch

Agent

Total unimplemented-tagged scenarios: 76 Classified: 76.

3.5k +5 changed yesterday A 0 tokens original Apache-2.0

SenteLabsAI/OpenExecutive

Agent Claude Code

Hostile logic reviewer. Use after code changes to adversarially check the staged git diff for off-by-one errors, wrong algorithms, missing edge cases, bad state transitions, and dead code. Reports a minimal triggering input per finding.

3.4k +437 yesterday A 54 tokens

SenteLabsAI/OpenExecutive

Agent Claude Code

Hostile maintainability reviewer. Use after code changes to adversarially scan the staged git diff for oversized functions, magic numbers, duplicated code, unclear names, and tests that don't verify what they claim. Quotes the exact offending lines.

3.4k +437 yesterday A 53 tokens

SenteLabsAI/OpenExecutive

Agent Claude Code

Hostile security reviewer. Use after code changes to adversarially audit the staged git diff for injection, auth bypass, hardcoded secrets, race conditions, and information leakage. Reports severity and a concrete exploit per finding.

3.4k +437 yesterday A 50 tokens

release-validation

70

Mesh-LLM/mesh-llm

Agent Codex

Validates the next MeshLLM release against the last GitHub release on approved real hosts and produces a formal evidence-backed readiness report.

3.4k +20 yesterday A 30 tokens original Apache-2.0

context

71

langchain-ai/langgraphjs

Agent ✓ vendor

Agents often require more than a list of messages to function effectively. They need context.

3.2k 7d ago A 0 tokens original MIT

memory

72

langchain-ai/langgraphjs

Agent ✓ vendor

LangGraph supports two types of memory essential for building conversational agents.

3.2k 7d ago A 0 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: