AI Safety skills

369 tagged AI Safety, measured the same way as everything else here.

Browse within: ai-security 267ai-behavior-analysis 149ai-safety-research 149ai-skill-safety 148ai-governance 139agent-governance 125policy-as-code 107credentials 106Guardrails 77claude-code-plugin 68ai-tools 64ai-coding 63compliance 59presentation 56

zugashield

49

Zuga-Technologies/ZugaShield

Skill Claude CodeCodex needs its repo

7-layer AI security + ML detection for OpenClaw. Covers all 10 OWASP Agentic AI risks — prompt injection, tool misuse, memory poisoning, data exfiltration, and more — across ALL channels (Signal, Telegram, Discord, WhatsApp, web) simultaneously.

not rated 1 12d ago A 61 tokens original MIT

rmanish2000-del/warrant-mcp

Skill Claude Code

Write a warrant-mcp policy in plain English, shaped for the closed rule set so review has the best chance of accepting it first time — only warrant-mcp review can say whether it compiles. Use when the user wants to create or edit .warrant/policy.md, write rules for what an AI agent may do in a project, or asks what…

not rated 1 18d ago A 126 tokens original MIT

truthgate

51

ay7627514-dotcom/truthgate

Skill Claude CodeCodex

Gate high-stakes, factual, or multi-step Hermes tasks before action and verify completion against evidence.

not rated 0 1mo ago A 23 tokens original MIT

mettle

52

Creed-Space/METTLE

Skill Claude CodeCodex

Use when a user wants to complete METTLE reverse-CAPTCHA challenges, obtain a signed result, or verify a METTLE credential.

not rated 0 changed 8d ago A 32 tokens original Apache-2.0

blog-writing

53

mizukaizen/hive-doctrine-mcp

Skill Claude CodeCodex

Use when creating a blog post from scratch. Guides the full process from topic research through to a publish-ready draft with proper SEO structure and heading hierarchy.

not rated 0 2mo ago A 0 tokens original MIT

Saenai/microsoft-todo-safe-mcp

Skill Codex

Inspect, validate, operate, and develop a self-contained Microsoft To Do Safe MCP skill with bundled server. Use for Microsoft To Do task/list management, safe backup/plan/preview/apply workflows, MCP startup checks, token-safe diagnostics, and deciding whether to run doctor/test/typecheck/build.

not rated 0 1mo ago A 67 tokens

contextlock

55

LutaElbert/contextlock

Skill Claude CodeCodex

Safely inspect a repository through the ContextLock MCP server. Use when exploring, searching, reviewing, or auditing repository files with ContextLock, especially before reading unfamiliar code or when sensitive files and secrets may be present.

not rated 0 2mo ago A 46 tokens original Apache-2.0

joy7758/titmas-agent-action-gate

Skill Claude CodeCodex

Normalize an ambiguous external-action proposal into a versioned request while preserving uncertainty and granting no authority.

not rated 0 20d ago A 24 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: