battle
25Skill Claude Code
Red vs Blue team security competition orchestrator. Runs long-running overnight battles with 1000s of interactions, scoring, and insight generation.
Shared skills for AI agents (Claude Code, Codex, Gemini)
Skill Claude Code
Red vs Blue team security competition orchestrator. Runs long-running overnight battles with 1000s of interactions, scoring, and insight generation.
Skill Claude Code
Standardized compliance QRA benchmarks against candidate LLMs.
Skill Claude CodeCodex
Non-negotiable agent behavior rules. Covers: no silent failures, no error bypassing, no raw AQL, no direct imports, no parallel infrastructure, no swallowed exceptions, use existing skills, fix errors don't dodge them, transparency in verdicts, no simulated reviews, and evidence-gated decisions.
Skill Claude CodeCodex
Repo-specific ArangoDB best practices: leverage texten analyzer (stop words, stemming, BM25), use AQL functions (LEVENSHTEINDISTANCE, TOKENS, NGRAMSIMILARITY, COSINESIMILARITY), store domain knowledge in collections not Python code, and never duplicate DB capabilities.
Skill Claude CodeCodex
Advisory-first, evidence-grounded art direction and review rules for creating digital experiences that feel genuinely custom to one brand instead of template-derived. Use when a user asks for bespoke web design, a distinctive visual world, personality-led art direction, an analysis of what makes a designer's work…
Skill Claude CodeCodex
Best practices for designing, reviewing, and implementing operator chat, evidence chat, run-card chat, artifact-inspector chat, and compliance-review chat surfaces. Use when users ask for chat UX, operator console UX, agent run UX, evidence receipts, trace cards, artifact drawers, progressive disclosure, or…
Skill Claude CodeCodex
Keep Sparta Chat usable as a modern chat interface while preserving evidence-gated compliance semantics. Use when designing, reviewing, or implementing ChatWell, InlineEvidenceCase, EvidenceWorkspace, ArtifactPanel, distance modes, voice/qid interactions, evidence receipts, artifact previews, or assistant answer…
Skill Codex
Best practices for Chatterbox or Chatterbox-Turbo voice agents, especially interruptible Embry-style agents that run memory, search, LLM, or other long-running skills concurrently. Use when designing, reviewing, or coding a voice coordinator, async task batch, JSON event stream, cancellable TTS queue, Chatterbox…
Skill Codex
Best practices for leading Ask compete and bakeoff workflows. Use when a user asks for competing models, isolated candidate implementations, winner selection, feature harvesting, approach comparison, model bakeoffs, or a creator competition where $ask should route browser and API handlers through Tau and the project…
Skill Claude CodeCodex
Best practices for conversational response behavior in voice-first agents: conversation tone, emotional steering, paralinguistic cue injection such as [laughter], wait/delay handling, interruption handling, identity-aware memory grounding, and Chatterbox-ready utterance policy. Use when designing, reviewing, or coding…
Skill Claude CodeCodex
Automated COTS defense UX compliance scanner. Tests against WCAG 2.1 AA, Section 508, MIL-STD-1472H, and NIST 800-53 UI controls via CDP interaction + VLM visual analysis.
Skill Claude CodeCodex
D3.js visualization best practices for performant, responsive, accessible data visualizations. Covers data joins, scales, axes, transitions, responsive SVG, interaction patterns, and accessibility. Use when writing, reviewing, or refactoring D3 visualizations.
Skill Claude CodeCodex
Delivery-proof discipline for agents driving external effects: browser submits, pane messages, file writes, pushes, API calls. Use when an agent is about to claim something was sent, submitted, delivered, landed, or running; when a transport reports success but the destination shows nothing; when an agent is retrying…
Skill Codex
Product UX and design-practice guardrails for classifying design work, selecting applicable best-practices- skills, preventing dashboard theater, defining mockup-first acceptance criteria, and deciding when to involve memory, dogpile, ask/scillm reviewers, interview, ux-lab, review-design, D3, React, infographic, or…
Skill Codex
Evidence-first typography and font-system guidance for digital products, websites, portfolios, dashboards, and design systems. Use when choosing or changing fonts, pairing typefaces, auditing overused font warnings, creating meaningful typographic hierarchy, validating font loading/provenance, or mapping typography to…
Skill Claude CodeCodex
Best practices for agent-resolved GitHub tickets, including bugs, feature requests, optimizations, maintenance, questions, and triage: filing contracts, route and subagent metadata, resolver leases, deterministic verification, review evidence, WebGPT escalation, and proof-based closure.
Skill Claude CodeCodex
Use for greenfield collaboration where the final product, architecture, workflow, schema, prompt, memory representation, visual direction, evaluation method, or implementation path is not fully known. Forces one named artifact, one visible contract, one candidate, one inspection, one status, and one next legal move…
Skill Claude CodeCodex
Repo-specific KDE/QML best practices for agentic coding: singleton design systems, property ordering, accessibility, performance (binding loops, delegate recycling), Plasma integration, and D-Bus patterns.
Skill Claude CodeCodex
Create, audit, and package AI-video-ready reference packs for Kling-style element binding. Use when users ask for Kling contact sheets, Kling-ready assets, element reference packs, character reference sheets, prop sheets, scene sheets, or consistent AI-video references.
Skill Claude CodeCodex
Best practices for Ask one-shot runs: the same question to N seats concurrently, answers returned per seat with no consensus, no judge, and no quorum. Use when a user asks several models one question and wants to read each answer, when partial answers are still useful, or when deciding whether a request is a one-shot…
Skill Claude CodeCodex
Canonical rubric for evaluating, scoring, ranking, and gating career and consulting opportunities for Graham Anderson, so the "top opportunities" selection is principled and repeatable rather than bespoke per run. Use when discovering, evaluating, reviewing, ranking, or deciding which opportunities to apply to in…
Skill Claude CodeCodex
Best practices for building, reviewing, and shipping Pi TypeScript extensions and extension packages. Use when authoring Pi extensions, custom tools, lifecycle hooks, intercom bridges, status guards, provider adapters, TUI components, package manifests, or retained extension evals.
Skill Claude CodeCodex
Evidence-backed standards for creating, reviewing, and hardening Pi extensions. Use when building or changing /.pi/agent/extensions or .pi/extensions code, package-style Pi extensions, event handlers, custom tools, TypeBox schemas, input/message/tool hooks, retry guards, final-report guards, or humorous but serious…
Skill Claude CodeCodex
Deck ARCHITECTURE rules for building pitch decks from source material — measured from a real 263-slide corpus, not invented. Use when planning a deck's sections and slide sequence, when deciding which source images become slides, or when reviewing whether a generated deck is house-shaped. Exists to prevent bespoking…
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: