Skill Claude CodeCodex
Autonomous self-improving pipeline for creating top-market demo videos using Remotion + Gemini video understanding.
Entity intelligence for any company, market, or question — not a chatbot that answers once, but a system that synthesizes with sources, turns each run into a reusable artifact, and watches for change later. Five-surface app + deep-work Workspace, Convex state, and a hosted public research MCP (npx nodebench-mcp).
Skill Claude CodeCodex
Autonomous self-improving pipeline for creating top-market demo videos using Remotion + Gemini video understanding.
Skill Claude CodeCodex
Migration runbook for @convex-dev/agent. Currently pinned to 0.2.10 because v0.6 requires a coordinated AI SDK v5 → v6 bump across many files. Use when bumping the agent or when CI shows the "args was removed in v0.6.0" error pattern.
Skill Claude CodeCodex
Schema and authoring guide for dep- skills — project-local migration runbooks for high-touch dependencies. Use when a dep major-bump breaks CI and you need to either apply an existing recipe or write a new one.
Skill Claude CodeCodex
Pin runbook for @tiptap/pm. Pinned to 3.22.3 because 3.22.4 dropped the ./collab package export, which some transitive consumer (likely @blocknote or @tiptap/extensions) still relies on. Use when bumping tiptap or when build fails with "./collab is not exported".
Skill Claude CodeCodex
Coordinated bump runbook for the vega + vega-lite + vega-embed ecosystem. The 5 → 6 jump fixed CVE-2025-59840 (XSS) and required updating the spec schema URL from v5 to v6. Use when bumping any of these three packages or when the vega XSS CVE alert fires.
Skill Claude CodeCodex
Remediation runbook for the SheetJS xlsx package. Pinned to the SheetJS CDN release because the npm-published xlsx has 2 unfixed HIGH CVEs (ReDoS + Prototype Pollution) and SheetJS only ships fixes via their own CDN. Use when bumping xlsx, when CVE-2024-22363 or CVE-2023-30533 alerts fire, or when considering a…
Skill Claude CodeCodex
Gemini 3 Pro-powered self-evaluating QA judge loop for live app surfaces.
Skill Claude CodeCodex
Run the repeatable operational loop on the NodeBench real-time chat pipeline or the report generator pipeline. Every pipeline change flows through: instrument → judge → persist → surface → measure → regress. Triggers: "pipeline", "operational loop", "run the standard", "measure the pipeline", "judge the traces", "chat…
Skill Claude CodeCodex
Turn "my app + an agent that demos" into "benchmarked, browser-verified, evidence-backed, prod-proven, and looping." Use when a (solo) founder wants to prove an AI agent works IN their real app — across all its UI surfaces — without cheating. Triggers: "set up proofloop", "proof-loop my app", "benchmark my agent's…