Use for image-heavy Codex work when generated or inspected media should stay behind a bounded MCP boundary and the controlling task should receive only durable Job IDs, hashes, relative references, and compact text handoffs.
Photo Decode / PhotoDecode / 解图 (photo-decode) — Analyze an uploaded image, reconstruct its visual logic into a source-adaptive background-free flat composition, then derive a palette and independently reconstructed key visual elements. Use when the user asks to 解图, Photo Decode, PhotoDecode, photo-decode, decode an…
Give the agent vision when the model itself is text-only (no image input). Use this skill whenever the user asks you to look at, analyze, describe, read, or OCR any image — a screenshot, a photo, a diagram, a UI, a chart, an error dialog, a pasted picture, or anything visual. Also triggers when the user pastes an…