browoz

4 mods across 1 repository, 3 stars between them.

agentic-evals

01

browoz/agentic-sdlc-skills

Skill Claude CodeCodex

Design evaluation contracts and test plans for agentic systems. Create deterministic tests, trajectory evals, quality dimensions, gold-set criteria, and CI gates before or after implementation. Use when asked for tests first, an eval plan, success criteria, non-deterministic testing, LLM-as-judge setup, or…

3 2mo ago A 95 tokens original MIT

browoz/agentic-sdlc-skills

Skill Claude CodeCodex

Prepare an AI agent system for production operation. Cover SHIELD controls, sandbox/canary/production rollout, OpenTelemetry observability with GenAI semantic conventions, Agent Card drafting, governance, and post-deploy monitoring. Use for production readiness checks, go-live checklists, agent monitoring, agent…

3 2mo ago A 97 tokens original MIT

browoz/agentic-sdlc-skills

Skill Claude CodeCodex

Run a security and dependency audit for agent systems, tool-using AI apps, MCP/A2A integrations, or security-sensitive AI-generated code. Check slopsquatting risk, tool shadowing, rug pulls, memory/context poisoning, secrets, unsafe permissions, and common CWE patterns. Use when asked for security review, dependency…

3 2mo ago A 106 tokens original MIT

agentic-spec

04

browoz/agentic-sdlc-skills

Skill Claude CodeCodex

Create a structured specification before agentic coding work. Assemble the six context types, scale rigor for prototype/internal/production tasks, produce SPEC.md, and configure focused AGENTS.md boundaries. Use when asked to write a spec, plan a feature, design an agent/system, define architecture, or create…

3 2mo ago A 91 tokens original MIT