Instructions file
Instructions for sohaibt/product-mode, covering claude.md product-led collaboration guidelines, pre-flight checklist, the seven principles, 1. frame the problem before the solution and 2. make assumptions & unknowns visible.
Instructions file
Instructions for sohaibt/product-mode, covering claude.md product-led collaboration guidelines, pre-flight checklist, the seven principles, 1. frame the problem before the solution and 2. make assumptions & unknowns visible.
Plugin Claude Code
The PM playbook for building AI agent products. 12 slash commands covering the full agent product lifecycle — from 'should this be an agent?' through eval design to production safety. Grounded in primary sources from Anthropic, OpenAI, Karpathy, Hamel Husain, Linus Lee, and real production deployments.
Skill Claude CodeCodex
Generate a complete PM spec (PRD-style document) for an AI agent product. Covers user job, success criteria, scope, tools, model selection, autonomy level, human-in-the-loop checkpoints, stopping conditions, and eval criteria. Use after architecture-pattern to lock down the product definition.
Skill Claude CodeCodex
Review the UX of an AI agent product against modern patterns beyond chat. Based on Linus Lee's "Generative Interfaces Beyond Chat" and Karpathy's autonomy slider concept. Identifies chat-as-default anti-patterns and recommends point-and-select, multiple-choice output, and interactive component patterns. Use when…
Skill Claude CodeCodex
Recommend the right agent architecture pattern from Anthropic's 5 (prompt chaining, routing, parallelization, orchestrator-workers, evaluator-optimizer) plus OpenAI's manager and decentralized patterns. Use after qualify-agent confirms an agent is needed.
Skill Claude CodeCodex
Position a feature on the autonomy spectrum (suggestion → execution → full autonomy) and design the verification loop. Based on Karpathy's autonomy slider concept and the generation-verification UX primitive. Use when deciding how much human oversight an AI feature needs.
Skill Claude CodeCodex
Model the per-task and monthly cost of an agent product. Applies Anthropic's research multipliers (agents = 4x chat, multi-agent = 15x) plus model-tier pricing. Identifies cost optimization opportunities. Use BEFORE building to avoid surprise infrastructure bills.
Skill Claude CodeCodex
Design a complete evaluation plan for an AI agent product. Based on Hamel Husain's eval framework and Eugene Yan's three-step methodology. Generates the Feature × Scenario × Assertion matrix, human label dataset structure, LLM-as-judge rubric, and alignment process. Use BEFORE shipping any AI feature.
Skill Claude CodeCodex
Catalog the likely failure modes for a specific agent product and design mitigations for each. Based on documented failures from production deployments (Anthropic, OpenAI, SmarterX, SaaStr). Use before launch to anticipate what will go wrong.
Skill Claude CodeCodex
Design a layered guardrail stack for an AI agent product. Based on OpenAI's 7-layer guardrail framework plus mandatory human-in-the-loop triggers. Determines which guardrails to build, in what priority, with what tripwires. Use BEFORE shipping to avoid the SmarterX-class failures.
Skill Claude CodeCodex
Pre-launch checklist for shipping an AI agent product to production. Covers training loop design, ops ownership, infrastructure (rainbow deployments, checkpointing), observability, rollback plans, and on-call alerting. Based on Anthropic's multi-agent production lessons and Lenny/Lemkin's deployment experience.
Skill Claude CodeCodex
Determine whether a feature should be a deterministic system, a single LLM call, a workflow, an agent, or a multi-agent system. Based on Anthropic's "bias toward simplicity" and OpenAI's three agent triggers. Use BEFORE designing any AI feature to avoid over-engineering.
Skill Claude CodeCodex
Audit an agent system for the specific failure modes that caused real production disasters — permission scoping, environment separation, destructive action gates, and backup isolation. Based on the SmarterX database deletion case study and Anthropic's production lessons. Use BEFORE granting an agent production access.
Skill Claude CodeCodex
Review a proposed tool definition (name, description, schema) for Agent-Computer Interface (ACI) quality. Based on Anthropic's principle that tool design is UX design for agents. Use when defining tools for any AI agent system.
Skill Claude CodeCodex
Force 10x thinking on any goal using Brian Chesky's "Add a Zero" exercise. Breaks you out of incremental thinking by asking what would need to be true to achieve 10x the result. Use when a team is thinking too small or optimizing within existing constraints.
Skill Claude CodeCodex
Score any process, workflow, or meeting for bureaucrat-mode creep. Based on Paul Graham's bureaucrat mode anti-patterns and Chesky's war on fake work. Use when something feels slow and you want to know if it's necessary complexity or unnecessary bureaucracy.
Skill Claude CodeCodex
Reframe a crisis, constraint, or setback as a forcing function for founder-mode decisions that were previously politically blocked. Based on Chesky's pandemic transformation and Andy Grove's "great companies are defined by their crises." Use when facing a serious challenge and want to use it as fuel, not just survive…
Skill Claude CodeCodex
Determine whether a decision is a founder-only move (requires institutional memory, passion, and permission) or a manageable decision that can be delegated. Based on Chesky's framework of what only a founder can do. Use when deciding whether to stay involved in or hand off a decision.
Skill Claude CodeCodex
Score whether a delegation decision is healthy (amplifying what you love) or harmful (abdicating what you don't understand). Based on Chesky's principles on what to keep vs. what to let go. Use when deciding whether to delegate something or stay involved.
Skill Claude CodeCodex
Diagnose where you sit on the founder mode vs. manager mode spectrum. Based on Brian Chesky's operating system and Paul Graham's framework. Use when a founder or CEO wants to assess whether they're leading like a founder or drifting into manager mode.
Skill Claude CodeCodex
Take a behavioral self-assessment to discover where you sit on the founder mode spectrum. 15 questions that reveal your actual operating mode — not what you think you do, but what you actually do. Use for honest self-reflection.
Skill Claude CodeCodex
Score an executive or senior hire using Brian Chesky's 'guilty until proven innocent' framework. Generates a reference check script, identifies red flags, and assesses builder vs. manager fit. Use when evaluating a candidate for a senior role.
Skill Claude CodeCodex
Build the product narrative BEFORE building the product. Based on Chesky's principle that "the story will often dictate the product." Use when planning a launch, feature, or product and you need the story to drive the design.
Skill Claude CodeCodex
Diagnose organizational health against Brian Chesky's framework. Detects the division → politics → bureaucracy → complacency arc that kills founder-led companies. Use when a founder suspects their org structure is slowing them down.