Run bounded, evidence-driven training research through W&B Launch: assess project readiness, establish launchable code and queue capacity, smoke-test real jobs, execute serial trials, compare metrics, and persist resumable research state. Use when a coding agent is asked to autonomously test training hypotheses or…
Convert W&B Table artifacts into non-destructive EvalTable previews with scan-first planning, typed input/output/score columns, bounded batches, verification, and safe removal. Use when a coding agent needs to create, inspect, compare, verify, or remove W&B EvalTable previews.
Primary W&B skill for broad or mixed Weights & Biases work: project overviews, W&B runs and artifacts, Weave traces and evaluations, Reports, and Launch workflows. Use when the task spans multiple W&B surfaces or the user asks generally what is happening in a W&B project.
AGENTS.md instructions for wandb/weave-claude-code, covering agents.md, settled decisions: do not "fix" these, verifying traces, transcript lifecycle and config.
Claude Code instructions for wandb/weave-claude-code, a project described as: Claude Code plugin that traces sessions, tool calls, and subagents to W&B Weave for observability and debugging.