Instructions file CodexOpenCode
Instructions for dmmdea/offload-harness, covering agents.md, what this repo is, documentation map, installing the stack and operating the stack.
Instructions file CodexOpenCode
Instructions for dmmdea/offload-harness, covering agents.md, what this repo is, documentation map, installing the stack and operating the stack.
Instructions file
Instructions for dmmdea/offload-harness, covering claude.md — agent orientation map for offload-harness, components & ports, model tiers (served by llama-swap on :11436), golden commands (all verified on this machine) and install / verify the stack (windows).
Skill Claude CodeCodex
Use when setting up the "local-offload" harness on a Windows machine — a free local Gemma-4 cascade that lets a coding agent delegate short-context grunt work (summarize / classify / extract / triage) so those tokens never hit the cloud context. Cross-vendor: NVIDIA (CUDA, ≥8GB), AMD Radeon incl. RDNA3 iGPUs like the…
Skill Claude CodeCodex
Drive a local ComfyUI render server from the shell, with a durable record of every run the server itself forgets. Trigger phrases: submit a comfyui graph, why can't comfyui see my model, how long did that render take, what made this output file, check what this loader accepts, use comfyui, run comfyui.
Skill Claude CodeCodex
The llama-swap operations console — durable history, drain-aware control, and measurement commands with three specific guarantees: keep-set unloads are refused statically by id AND alias (never from server ttl), --drain fails closed when slot state is unreadable, and fit/ctx refuse to answer inside their uncertainty…