sohaibt

26 mods across 4 repositories, 85 stars between them.

sohaibt/product-mode

Instructions file

Instructions for sohaibt/product-mode, covering claude.md product-led collaboration guidelines, pre-flight checklist, the seven principles, 1. frame the problem before the solution and 2. make assumptions & unknowns visible.

48 4mo ago A 1,792 tokens original MIT

agent-pm

02

sohaibt/agent-pm

Plugin Claude Code

The PM playbook for building AI agent products. 12 slash commands covering the full agent product lifecycle — from 'should this be an agent?' through eval design to production safety. Grounded in primary sources from Anthropic, OpenAI, Karpathy, Hamel Husain, Linus Lee, and real production deployments.

13 3mo ago A tokens not measured original MIT

agent-spec

03

sohaibt/agent-pm

Skill Claude CodeCodex

Generate a complete PM spec (PRD-style document) for an AI agent product. Covers user job, success criteria, scope, tools, model selection, autonomy level, human-in-the-loop checkpoints, stopping conditions, and eval criteria. Use after architecture-pattern to lock down the product definition.

13 3mo ago A 61 tokens original MIT

agent-ux-review

04

sohaibt/agent-pm

Skill Claude CodeCodex

Review the UX of an AI agent product against modern patterns beyond chat. Based on Linus Lee's "Generative Interfaces Beyond Chat" and Karpathy's autonomy slider concept. Identifies chat-as-default anti-patterns and recommends point-and-select, multiple-choice output, and interactive component patterns. Use when…

13 3mo ago A 72 tokens original MIT

sohaibt/agent-pm

Skill Claude CodeCodex

Recommend the right agent architecture pattern from Anthropic's 5 (prompt chaining, routing, parallelization, orchestrator-workers, evaluator-optimizer) plus OpenAI's manager and decentralized patterns. Use after qualify-agent confirms an agent is needed.

13 3mo ago A 51 tokens original MIT

autonomy-slider

06

sohaibt/agent-pm

Skill Claude CodeCodex

Position a feature on the autonomy spectrum (suggestion → execution → full autonomy) and design the verification loop. Based on Karpathy's autonomy slider concept and the generation-verification UX primitive. Use when deciding how much human oversight an AI feature needs.

13 3mo ago A 54 tokens original MIT

cost-model

07

sohaibt/agent-pm

Skill Claude CodeCodex

Model the per-task and monthly cost of an agent product. Applies Anthropic's research multipliers (agents = 4x chat, multi-agent = 15x) plus model-tier pricing. Identifies cost optimization opportunities. Use BEFORE building to avoid surprise infrastructure bills.

13 3mo ago A 57 tokens original MIT

eval-design

08

sohaibt/agent-pm

Skill Claude CodeCodex

Design a complete evaluation plan for an AI agent product. Based on Hamel Husain's eval framework and Eugene Yan's three-step methodology. Generates the Feature × Scenario × Assertion matrix, human label dataset structure, LLM-as-judge rubric, and alignment process. Use BEFORE shipping any AI feature.

13 3mo ago A 63 tokens original MIT

failure-mode-map

09

sohaibt/agent-pm

Skill Claude CodeCodex

Catalog the likely failure modes for a specific agent product and design mitigations for each. Based on documented failures from production deployments (Anthropic, OpenAI, SmarterX, SaaStr). Use before launch to anticipate what will go wrong.

13 3mo ago A 52 tokens original MIT

guardrails-plan

10

sohaibt/agent-pm

Skill Claude CodeCodex

Design a layered guardrail stack for an AI agent product. Based on OpenAI's 7-layer guardrail framework plus mandatory human-in-the-loop triggers. Determines which guardrails to build, in what priority, with what tripwires. Use BEFORE shipping to avoid the SmarterX-class failures.

13 3mo ago A 64 tokens original MIT

prod-readiness

11

sohaibt/agent-pm

Skill Claude CodeCodex

Pre-launch checklist for shipping an AI agent product to production. Covers training loop design, ops ownership, infrastructure (rainbow deployments, checkpointing), observability, rollback plans, and on-call alerting. Based on Anthropic's multi-agent production lessons and Lenny/Lemkin's deployment experience.

13 3mo ago A 64 tokens original MIT

qualify-agent

12

sohaibt/agent-pm

Skill Claude CodeCodex

Determine whether a feature should be a deterministic system, a single LLM call, a workflow, an agent, or a multi-agent system. Based on Anthropic's "bias toward simplicity" and OpenAI's three agent triggers. Use BEFORE designing any AI feature to avoid over-engineering.

13 3mo ago A 62 tokens original MIT

risk-audit

13

sohaibt/agent-pm

Skill Claude CodeCodex

Audit an agent system for the specific failure modes that caused real production disasters — permission scoping, environment separation, destructive action gates, and backup isolation. Based on the SmarterX database deletion case study and Anthropic's production lessons. Use BEFORE granting an agent production access.

13 3mo ago A 59 tokens original MIT

tool-design-review

14

sohaibt/agent-pm

Skill Claude CodeCodex

Review a proposed tool definition (name, description, schema) for Agent-Computer Interface (ACI) quality. Based on Anthropic's principle that tool design is UX design for agents. Use when defining tools for any AI agent system.

13 3mo ago A 51 tokens original MIT

add-a-zero

15

sohaibt/founder-mode

Skill Claude CodeCodex

Force 10x thinking on any goal using Brian Chesky's "Add a Zero" exercise. Breaks you out of incremental thinking by asking what would need to be true to achieve 10x the result. Use when a team is thinking too small or optimizing within existing constraints.

3 4mo ago A 61 tokens original MIT

sohaibt/founder-mode

Skill Claude CodeCodex

Score any process, workflow, or meeting for bureaucrat-mode creep. Based on Paul Graham's bureaucrat mode anti-patterns and Chesky's war on fake work. Use when something feels slow and you want to know if it's necessary complexity or unnecessary bureaucracy.

3 4mo ago A 58 tokens original MIT

crisis-catalyst

17

sohaibt/founder-mode

Skill Claude CodeCodex

Reframe a crisis, constraint, or setback as a forcing function for founder-mode decisions that were previously politically blocked. Based on Chesky's pandemic transformation and Andy Grove's "great companies are defined by their crises." Use when facing a serious challenge and want to use it as fuel, not just survive…

3 4mo ago A 68 tokens original MIT

decision-check

18

sohaibt/founder-mode

Skill Claude CodeCodex

Determine whether a decision is a founder-only move (requires institutional memory, passion, and permission) or a manageable decision that can be delegated. Based on Chesky's framework of what only a founder can do. Use when deciding whether to stay involved in or hand off a decision.

3 4mo ago A 59 tokens original MIT

delegation-scorer

19

sohaibt/founder-mode

Skill Claude CodeCodex

Score whether a delegation decision is healthy (amplifying what you love) or harmful (abdicating what you don't understand). Based on Chesky's principles on what to keep vs. what to let go. Use when deciding whether to delegate something or stay involved.

3 4mo ago A 60 tokens original MIT

founder-audit

20

sohaibt/founder-mode

Skill Claude CodeCodex

Diagnose where you sit on the founder mode vs. manager mode spectrum. Based on Brian Chesky's operating system and Paul Graham's framework. Use when a founder or CEO wants to assess whether they're leading like a founder or drifting into manager mode.

3 4mo ago A 55 tokens original MIT

founder-quiz

21

sohaibt/founder-mode

Skill Claude CodeCodex

Take a behavioral self-assessment to discover where you sit on the founder mode spectrum. 15 questions that reveal your actual operating mode — not what you think you do, but what you actually do. Use for honest self-reflection.

3 4mo ago A 52 tokens original MIT

hiring-scorecard

22

sohaibt/founder-mode

Skill Claude CodeCodex

Score an executive or senior hire using Brian Chesky's 'guilty until proven innocent' framework. Generates a reference check script, identifies red flags, and assesses builder vs. manager fit. Use when evaluating a candidate for a senior role.

3 4mo ago A 54 tokens original MIT

launch-story

23

sohaibt/founder-mode

Skill Claude CodeCodex

Build the product narrative BEFORE building the product. Based on Chesky's principle that "the story will often dictate the product." Use when planning a launch, feature, or product and you need the story to drive the design.

3 4mo ago A 48 tokens original MIT

org-health

24

sohaibt/founder-mode

Skill Claude CodeCodex

Diagnose organizational health against Brian Chesky's framework. Detects the division → politics → bureaucracy → complacency arc that kills founder-led companies. Use when a founder suspects their org structure is slowing them down.

3 4mo ago A 45 tokens original MIT