build-with-dhiraj

60 mods across 1 repository, 4 stars between them.

benchmark-agents

51

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

Advanced AI agent benchmark scenarios that push Vercel's cutting-edge platform features — Workflow DevKit, AI Gateway, MCP, Chat SDK, Queues, Flags, Sandbox, and multi-agent orchestration. Designed to stress-test skill injection for complex, multi-system builds.

4 20d ago B 58 tokens original MIT

benchmark-e2e

52

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

End-to-end benchmark suite for vercel-plugin. Runs realistic projects through skill injection, launches dev servers, verifies everything works, analyzes conversation logs, and produces an improvement report for overnight self-improvement loops.

4 20d ago B 46 tokens original MIT

benchmark-sandbox

53

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

Run vercel-plugin eval scenarios in Vercel Sandboxes instead of local WezTerm panels. Provisions ephemeral microVMs with Claude Code + plugin pre-installed, runs benchmark prompts, extracts hook artifacts, and produces coverage reports.

4 20d ago B 51 tokens original MIT

benchmark-testing

54

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

Create and launch benchmark test projects to exercise vercel-plugin skill injection across realistic scenarios. Sets up isolated directories, installs the plugin, and spawns WezTerm panes running Claude Code with crafted prompts.

4 20d ago B 43 tokens original MIT

plugin-audit

55

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

Audit vercel-plugin performance on real-world projects. Extracts tool calls from Claude Code conversation logs, tests hook matching against actual inputs, identifies pattern coverage gaps, and checks plugin cache staleness. Use when asked to audit, test, or investigate plugin skill injection on a real project.

4 20d ago A 62 tokens original MIT

vercel-plugin-eval

57

build-with-dhiraj/ai-workflow-framework-portability-kit

Skill Claude CodeCodex

Run live eval sessions against the vercel-plugin to verify hook behavior, skill injection, dedup correctness, and coverage. Launches real Claude Code sessions via WezTerm, monitors debug logs, and produces a structured coverage report.

4 20d ago B 52 tokens original MIT

ai-architect

59

build-with-dhiraj/ai-workflow-framework-portability-kit

Agent

Specializes in architecting AI-powered applications on Vercel — choosing between AI SDK patterns, configuring providers, building agents, setting up durable workflows, and integrating MCP servers. Use when designing AI features, building chatbots, or creating agentic applications.

4 20d ago A 55 tokens original MIT

deployment-expert

60

build-with-dhiraj/ai-workflow-framework-portability-kit

Agent

Specializes in Vercel deployment strategies, CI/CD pipelines, preview URLs, production promotions, rollbacks, environment variables, and domain configuration. Use when troubleshooting deployments, setting up CI/CD, or optimizing the deploy pipeline.

4 20d ago A 49 tokens original MIT