HomenShum/NodeBenchAI

Entity intelligence for any company, market, or question — not a chatbot that answers once, but a system that synthesizes with sources, turns each run into a reusable artifact, and watches for change later. Five-surface app + deep-work Workspace, Convex state, and a hosted public research MCP (npx nodebench-mcp).

14Stars on the repository
61Mods indexed here, across every type
18d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

deep-diligence

01

HomenShum/NodeBenchAI

Agent Claude Code

Full-stack deep diligence agent — structural QA, design coherence against target personas, narrative alignment, and competitive positioning audit. Covers code, UI, UX, content, performance, accessibility, and product-market fit in one pass.

14 18d ago A 49 tokens original MIT

deploy-and-launch

02

HomenShum/NodeBenchAI

Agent Claude Code

Full deployment, production verification, and launch readiness agent. Deploys Convex backend, Vercel frontend, tests voice server, runs production smoke tests, and produces a launch checklist.

14 18d ago A 41 tokens original MIT

dogfood-loop

03

HomenShum/NodeBenchAI

Agent Claude Code

Self-dogfood NodeBench by using the deployed app and MCP tools, scoring quality, and filing findings.

14 18d ago A 25 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Runs benchmark suites, compares against baselines, and produces machine-readable and human-readable regression reports.

14 18d ago A 26 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Produces structured strategic analysis for diligence, GTM, strategy, and intervention questions. Use when variables, scenarios, trust nodes, or ranked interventions matter.

14 18d ago A 40 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Tightens NodeBench UI for demo quality, executive clarity, and traversal reliability. Use for landing page, decision workbench, and operator-facing polishing.

14 18d ago A 38 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Supervises bounded continuous improvement loops for NodeBench. Use for roadmap-aligned iteration, benchmark review, and safe recurring maintenance.

14 18d ago A 31 tokens original MIT

qa-dogfood

09

HomenShum/NodeBenchAI

Agent Claude Code

Full QA dogfood agent — traverses every surface, clicks every interactive element, checks contrast, performance, accessibility, behavioral correctness, and files findings as actionable P0/P1/P2 issues.

14 18d ago A 44 tokens original MIT

HomenShum/NodeBenchAI

Agent Claude Code

Universal full-stack product diligence agent. Drop into any repo — audits UX, design, code quality, accessibility, performance, content, and competitive positioning. No app-specific knowledge needed.

14 18d ago A 43 tokens original MIT

HomenShum/NodeBenchAI

Agent

Before: The Morning Dossier page rendered raw RSS-style log lines ("Trending on Hacker News with 35 points and 6 comments") instead of editorial prose synthesis. Backend failures ("Pending — run this follow-up") leaked into the UI. No structured brief generation existed.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Agent

Make every in-app agent run capable of producing a final operator-facing verdict with open-source citations, trace-backed evidence, and explicit next actions, all surfaced through the existing UI.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Agent

A unified NodeBench subsystem enabling agents to self-heal from both semantic errors (rollback + lessons) and infrastructure failures (capability-aware model failover + budget gates), so long-running research jobs keep making progress when the user steps away.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Agent

Implemented comprehensive streaming UI optimization patterns to ensure smooth, animated streaming with per-step updates at 30-60fps without layout thrashing.

14 18d ago A 0 tokens original MIT

QUICK_START

19

HomenShum/NodeBenchAI

Agent

A complete streaming UI optimization system for FastAgentPanel that ensures smooth 30-60fps rendering with per-step updates and no layout thrashing.

14 18d ago A 0 tokens original MIT

HomenShum/NodeBenchAI

Agent

This document describes the streaming UI optimization patterns implemented in FastAgentPanel to ensure smooth, animated streaming with per-step updates at 30-60fps without layout thrashing.

14 18d ago A 0 tokens original MIT