HomenShum

61 mods across 1 repository, 14 stars between them.

deep-diligence

01

HomenShum/NodeBenchAI

Agent Claude Code

Full-stack deep diligence agent — structural QA, design coherence against target personas, narrative alignment, and competitive positioning audit. Covers code, UI, UX, content, performance, accessibility, and product-market fit in one pass.

14 18d ago A 49 tokens original MIT

deploy-and-launch

02

HomenShum/NodeBenchAI

Agent Claude Code

Full deployment, production verification, and launch readiness agent. Deploys Convex backend, Vercel frontend, tests voice server, runs production smoke tests, and produces a launch checklist.

14 18d ago A 41 tokens original MIT

dogfood-loop

03

HomenShum/NodeBenchAI

Agent Claude Code

Self-dogfood NodeBench by using the deployed app and MCP tools, scoring quality, and filing findings.

14 18d ago A 25 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Runs benchmark suites, compares against baselines, and produces machine-readable and human-readable regression reports.

14 18d ago A 26 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Produces structured strategic analysis for diligence, GTM, strategy, and intervention questions. Use when variables, scenarios, trust nodes, or ranked interventions matter.

14 18d ago A 40 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Tightens NodeBench UI for demo quality, executive clarity, and traversal reliability. Use for landing page, decision workbench, and operator-facing polishing.

14 18d ago A 38 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Supervises bounded continuous improvement loops for NodeBench. Use for roadmap-aligned iteration, benchmark review, and safe recurring maintenance.

14 18d ago A 31 tokens original MIT

qa-dogfood

09

HomenShum/NodeBenchAI

Agent Claude Code

Full QA dogfood agent — traverses every surface, clicks every interactive element, checks contrast, performance, accessibility, behavioral correctness, and files findings as actionable P0/P1/P2 issues.

14 18d ago A 44 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Universal full-stack product diligence agent. Drop into any repo — audits UX, design, code quality, accessibility, performance, content, and competitive positioning. No app-specific knowledge needed.

14 18d ago A 43 tokens original MIT

nodebench-qa

12

HomenShum/NodeBenchAI

Command Claude Code

Full QA loop: crawl → findings → fix suggestions → re-crawl → savings report.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Command Claude Code

Master multi-agent orchestration using Claude Code's TeammateTool and Task system. Use when coordinating multiple agents, running parallel code reviews, creating pipeline workflows with dependencies, building self-organizing task queues, or any task benefiting from divide-and-conquer patterns.

14 18d ago B 59 tokens original MIT

scenario-testing

15

HomenShum/NodeBenchAI

Command Claude Code

Review the tests in the current file or feature against the scenario-based testing mandate.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Autonomous self-improving pipeline for creating top-market demo videos using Remotion + Gemini video understanding.

14 18d ago A 0 tokens original MIT

dep-convex-agent

18

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Migration runbook for @convex-dev/agent. Currently pinned to 0.2.10 because v0.6 requires a coordinated AI SDK v5 → v6 bump across many files. Use when bumping the agent or when CI shows the "args was removed in v0.6.0" error pattern.

14 18d ago A 72 tokens original MIT

dep-skills-howto

19

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Schema and authoring guide for dep- skills — project-local migration runbooks for high-touch dependencies. Use when a dep major-bump breaks CI and you need to either apply an existing recipe or write a new one.

14 18d ago A 50 tokens original MIT

dep-tiptap-pm

20

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Pin runbook for @tiptap/pm. Pinned to 3.22.3 because 3.22.4 dropped the ./collab package export, which some transitive consumer (likely @blocknote or @tiptap/extensions) still relies on. Use when bumping tiptap or when build fails with "./collab is not exported".

14 18d ago A 81 tokens original MIT

dep-vega

21

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Coordinated bump runbook for the vega + vega-lite + vega-embed ecosystem. The 5 → 6 jump fixed CVE-2025-59840 (XSS) and required updating the spec schema URL from v5 to v6. Use when bumping any of these three packages or when the vega XSS CVE alert fires.

14 18d ago A 80 tokens original MIT

dep-xlsx

22

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Remediation runbook for the SheetJS xlsx package. Pinned to the SheetJS CDN release because the npm-published xlsx has 2 unfixed HIGH CVEs (ReDoS + Prototype Pollution) and SheetJS only ships fixes via their own CDN. Use when bumping xlsx, when CVE-2024-22363 or CVE-2023-30533 alerts fire, or when considering a…

14 18d ago A 95 tokens original MIT

HomenShum/NodeBenchAI

Skill Claude CodeCodex

Run the repeatable operational loop on the NodeBench real-time chat pipeline or the report generator pipeline. Every pipeline change flows through: instrument → judge → persist → surface → measure → regress. Triggers: "pipeline", "operational loop", "run the standard", "measure the pipeline", "judge the traces", "chat…

14 18d ago A 99 tokens original MIT