guardrail-design
121Skill Claude CodeCodex
Defining behavioral boundaries — what the AI should and shouldn't do.
Skill Claude CodeCodex
Defining behavioral boundaries — what the AI should and shouldn't do.
Skill Claude CodeCodex
Proactively identifying failure modes, misuse, and unintended consequences.
Skill Claude CodeCodex
Showing users what the AI knows, doesn't know, and how confident it is.
Skill Claude CodeCodex
Helping users form warranted trust in the AI — neither overtrust nor undertrust — through deliberate confidence and source signalling.
Skill Claude CodeCodex
Translating organisational values and user expectations into system constraints.
Plugin Claude Code
Design multi-agent systems, handoffs between AI agents, and human-in-the-loop workflows.
Command
Create a human oversight plan for an agentic system.
Command
Design a complete multi-agent workflow with roles, handoffs, and fallbacks.
Command
Map out agent responsibilities, boundaries, and communication patterns.
Skill Claude CodeCodex
Defining what each agent does, knows, and owns in a multi-agent system.
Skill Claude CodeCodex
What happens when an agent fails — retry, fallback, escalate, or graceful degradation.
Skill Claude CodeCodex
Designing smooth transitions between agents and between AI and humans.
Skill Claude CodeCodex
Designing intervention points where humans review, approve, or redirect agent work.
Skill Claude CodeCodex
Making multi-agent workflows visible and debuggable for designers and developers.
Skill Claude CodeCodex
Managing shared context, memory, and state across multiple agents.
Skill Claude CodeCodex
Breaking complex user goals into subtasks that agents can handle.
Plugin Claude Code
Measure AI output quality, user satisfaction, task success, and design effectiveness.
Command
Build a scoring rubric for evaluating AI output quality.
Command
Design a benchmark suite to measure AI product performance over time.
Command
Execute a structured evaluation of an AI feature against defined criteria.
Skill Claude CodeCodex
A/B testing, side-by-side comparison, and preference ranking for AI outputs.
Skill Claude CodeCodex
Classifying AI failures — hallucination, refusal, irrelevance, tone mismatch, latency.
Skill Claude CodeCodex
Adapting Nielsen's heuristics and new AI-specific heuristics for AI interfaces.
Skill Claude CodeCodex
Tracking AI product quality over time — drift, degradation, and improvement.