/eval
01Command Claude Code
Run eval suite and report regressions vs. main baseline.
not rated 5 1mo ago A 14 tokens
Multi-agent (LangGraph + Claude) app that turns an earnings call into a source-attributed analyst brief: ingest - FinBERT tone + KPI-vs-consensus - SurpriseSignal - grounding check - delivery.
Command Claude Code
Run eval suite and report regressions vs. main baseline.
Command Claude Code
Scaffold a new worker agent with module, prompt v1, schema, tests, and eval stub.
Command Claude Code
Staff-engineer code review: security, regressions, prompt diffs.