agent-security instructions

715 tagged agent-security, measured the same way as everything else here.

Browse within: ai-security 24authorization 18llm-security 12oauth2 12zero-trust 12ai-governance 8AI Safety 7formal-verification 5

reins

49

pegasi-ai/reins

Plugin Claude Code

Runtime security for AI agents. Blocks destructive actions before execution, routes high-risk operations through human approval, and maintains an immutable audit trail. Covers OWASP MCP Top 10, ASI Top 10, and Agentic Skills Top 10.

408 3mo ago A tokens not measured original Apache-2.0

reins

50

pegasi-ai/reins

Skill Claude CodeCodex

Use whenever security, policies, governance, guardrails, compliance, or safety are relevant — including blocked commands, audit trails, dangerous operations, deletions, file modifications, shell commands, MCP access, API calls, network requests, credentials, or any action that could be irreversible or destructive.

408 3mo ago C 58 tokens original Apache-2.0

reins

51

pegasi-ai/reins

Skill Claude CodeCodex

Use this skill whenever security, policies, governance, guardrails, compliance, or safety are relevant — including blocked commands, audit trails, dangerous operations, deletions, file modifications, shell commands, MCP access, API calls, network requests, credentials, or any action that could be irreversible or…

408 3mo ago C 90 tokens original Apache-2.0

claude-code

52

pegasi-ai/reins

Agent

Claude Code is Anthropic's CLI-based coding agent. ToolShield injects safety guidelines into its global instruction file.

408 3mo ago A 0 tokens original Apache-2.0

cursor

53

pegasi-ai/reins

Agent

Cursor is an AI-powered code editor. ToolShield injects safety guidelines into Cursor's global user rules stored in its internal SQLite database.

408 3mo ago A 0 tokens original Apache-2.0

openhands

54

pegasi-ai/reins

Agent

OpenHands is an open-source AI-driven development platform. ToolShield injects safety guidelines as a microagent that is always loaded into the agent's system prompt.

408 3mo ago A 0 tokens original Apache-2.0

scan-daily

56

Agent-Threat-Rule/agent-threat-rules

Command Claude Code

Run incremental ecosystem scan: crawl registries → audit new packages → merge results → generate posts.

376 3d ago A 0 tokens original MIT

api-caller

59

NVIDIA/SkillEvaluator

Skill Claude CodeCodex ✓ vendor

Call any REST API dynamically. Make GET, POST, PUT, DELETE requests to any endpoint with custom headers and JSON body.

357 3d ago A 29 tokens original Apache-2.0

calculator

60

NVIDIA/SkillEvaluator

Skill Claude CodeCodex ✓ vendor

Evaluate mathematical expressions and unit conversions. Handles arithmetic, percentages, exponents, and common unit conversions (temperature, distance, weight). No external dependencies.

357 3d ago A 32 tokens original Apache-2.0

NVIDIA/SkillEvaluator

Skill Claude CodeCodex ✓ vendor

Use when converting an existing benchmark, rubric, verifier, task YAML/JSON, or domain check into SkillEvaluator BYOG/BYOT custom evaluation.

357 3d ago A 35 tokens original Apache-2.0

prismor AGENTS.md

62

PrismorSec/prismor

Instructions file CodexOpenCode

Instructions for PrismorSec/prismor, covering agents.md, primary objectives, start here, how to work in this repo and 1. treat security guidance as product logic.

336 3d ago C 4,307 tokens original Apache-2.0

prismor CLAUDE.md

63

PrismorSec/prismor

Instructions file

Instructions for PrismorSec/prismor, covering prismor security — claude.md, prismor runtime protection, cloaking (secret prevention) and working in this repo.

336 3d ago A 359 tokens original Apache-2.0

prismor

64

PrismorSec/prismor

Skill Claude CodeCodex

Runtime security for AI coding agents. Use when about to install a package, paste a secret, run a destructive command, reach an unfamiliar host, govern MCP servers, set up a new workspace, or recover from a Prismor block.

336 3d ago D 51 tokens original Apache-2.0

Asymptote-Labs/agent-beacon

Skill Claude CodeCodex

Verify a Beacon change end to end by running a real Claude Code session inside a disposable Linux cloud sandbox and checking that Beacon captured what the agent actually did. Use when asked to verify, validate, test, or prove that a Beacon change works for real rather than just compiling; when asked whether telemetry…

323 +4 2d ago A 114 tokens original MIT

Asymptote-Labs/agent-beacon

Instructions file CodexOpenCode

AGENTS.md instructions for Asymptote-Labs/agent-beacon, covering agent instructions for agent-beacon, verifying a change end to end and running the standard checks.

323 +4 2d ago A 580 tokens original MIT

Asymptote-Labs/agent-beacon

Instructions file

Claude Code instructions for Asymptote-Labs/agent-beacon, covering claude.md, project scope, product posture, telemetry scope and common commands.

323 +4 changed today D 6,805 tokens original MIT

cursorrules

68

0xNyk/awesome-agent-cortex

Cursor rule Cursor

Cursor rule "cursorrules" from 0xNyk/awesome-agent-cortex, covering .cursorrules — cursor ide rules file, project context, tech stack, code style and development workflow.

211 +2 5d ago A 335 tokens original CC0-1.0

Garudex-Labs/caracal

Agent Claude Code

Invoke ONLY when explicitly asked for an outside-in usability and adoption assessment of Caracal from a real external engineer's perspective. This agent evaluates: Console UX, SDK ergonomics, CLI experience, documentation clarity, deployment workflows, provider/gateway setup, policy model, app/resource modeling…

205 3d ago A 153 tokens original Apache-2.0

Garudex-Labs/caracal

Agent Claude Code

Invoke ONLY when explicitly asked for a production-readiness audit or hardening of the Caracal OSS platform: reliability, stability, recoverability, scalability, observability, performance, operational simplicity, deployment, upgrade safety, and maintainability. Finds production gaps, explains root causes, designs…

205 3d ago A 125 tokens original Apache-2.0