vlm
25graph-robots/open-robot-skills
Skill Claude CodeCodex
Free-form and yes/no visual question answering against a hosted vision-language model (OpenRouter API by default; Vertex AI Gemini selectable by config). Use when a workflow needs scene descriptions, semantic checks ("is the drawer open?"), or LLM-judged verification of a camera frame — no GPU required.
39 7d ago A 64 tokens
original Apache-2.0