Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/Oriolshhh/runware-image-mcpWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/commands/oriolshhh/runware-image-mcp/polish-with-images)<a href="https://agentmods.dev/commands/oriolshhh/runware-image-mcp/polish-with-images"><img src="https://agentmods.dev/badge/commands/oriolshhh/runware-image-mcp/polish-with-images/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/commands/oriolshhh/runware-image-mcp/polish-with-images"><img src="https://agentmods.dev/badge/commands/oriolshhh/runware-image-mcp/polish-with-images.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00046 | $0.01662 |
| Opus 5 | $0.00023 | $0.00831 |
| Sonnet 5 | $0.00009 | $0.00332 |
| Haiku 4.5 | $0.00005 | $0.00166 |
Grade A, and why
polish-with-images scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 160 lines — stays where its author put it; the contents beside it link to each section on GitHub.
/polish-with-images — Plan a coherent image system for an existing page
Purpose
Find the exact places where illustrations or images would materially improve an existing webpage, define one professional art direction, and produce a generation-ready specification with few, medium, and many-image options.
This command plans only. It does not generate images or change production code before the user chooses a tier and approves the spec.
Invocation
/polish-with-images <route/page and desired outcome>
Examples:
/polish-with-images / Make the landing page feel more credible and memorable
/polish-with-images /features Explain the workflow visually without making it busy
/polish-with-images /about Add a coherent editorial illustration system
Accepted input
An existing route/page, screenshots or references, desired mood or medium, brand constraints, audience, product claims, and optional asset-count or performance limits.
Prerequisites
A HarnessKit workspace and an existing webpage. A runnable preview is strongly preferred. Static source or screenshots may be used with lower-confidence findings.
Procedure
- Apply
context-discovery. Resolve the exact page, audience, primary task, intended outcome, preview command, stable content/data, supported themes, relevant viewports, existing assets, and asset-delivery conventions. - Inspect repository and page evidence before asking questions. Ask one grouped
batch of at most three decision-changing questions only when necessary:
- illustration, photography, diagram, product capture, or mixed medium;
- immutable brand/reference constraints;
- acceptable performance/licensing/generation-provider constraints. Propose evidence-backed defaults for unanswered preferences.
- Have
visual-qa-testercapture the current page at representative narrow, medium, and wide viewports. Record hierarchy, long text runs, abstract claims, trust gaps, empty states, page rhythm, existing imagery, LCP candidates, themes, and responsive behavior. - Have
visual-assets-directorapplyimage-art-direction,visual-craft, anddesign-system-integration. Produce a candidate-slot inventory. For every possible location record:- exact route, section/landmark, component, and surrounding copy;
- the user or communication problem;
- proposed image role and why image is better than copy/layout alone;
- expected value, confidence, accessibility/performance risk, and rejection reason if omitted.
- Define one page-wide style lock grounded in existing brand evidence: medium, perspective, composition, palette, lighting, texture, geometry, line/edge treatment, depth, detail, background, whitespace, representation, crop behavior, and forbidden motifs. Include an exact shared prompt prefix and suffix. Every asset prompt must repeat them verbatim so independent generation calls remain stylistically consistent.
- Build three cumulative tiers:
- Tier 1 — Few / essential: normally 1–3 highest-value assets.
- Tier 2 — Medium / balanced: normally 4–7 total assets and includes Tier 1.
- Tier 3 — Many / editorial: normally 8–12 total assets and includes Tier 2, but may use fewer when additional images would become filler. For each tier state the outcome, included asset IDs, page-weight budget, implementation effort, strengths, risks, and why it is meaningfully better than the tier below.
- Create a complete asset registry. Every asset must specify:
- stable ID and tiers containing it;
- exact placement and visual purpose;
- asset type, unique subject/action, composition, mood, and focal point;
- one self-contained generation prompt containing the exact shared style lock;
- negative constraints and prohibited content;
- aspect ratio and slot-specific source resolution(s);
- rendered dimensions by breakpoint, safe crop, responsive
srcset/art direction, and optional theme variants; - alt text or explicit decorative rationale, caption if required;
- format, compression target, loading priority, LCP implications, dimensions, fallback, and maximum file size;
- content/licensing dependencies and visual acceptance criteria. Resolutions should differ when placements differ; do not normalize a hero, portrait, card, and inline diagram to one canvas size.
- Have
visual-craft-specialistreview the three tiers and prompts for product specificity, stylistic consistency, hierarchy, visual repetition, generic AI motifs, content integrity, and over-decoration. Resolve findings before persisting the spec. - Write
.agent/specs/<kebab-case-name>.mdusing the required frontmatter below. Include findings, candidate/rejected slots, style lock, all three tiers, asset registry, implementation tasks, asset-generation/selection QA, integration plan, accessibility, performance budgets, browser scenarios, risks, non-goals, and open decisions. - Present a concise comparison table and recommend one tier based on user
value and cost. Then provide a
/sum-spec-style summary and stop for the user to choose a tier and explicitly approve the spec.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 11d ago First seen · 160 lines · 46 tokens per session scan A 68712797c440
polish-with-images is a command published in the GitHub repository Oriolshhh/runware-image-mcp (0 stars, last pushed 2mo ago), licensed MIT. It adds 46 tokens to every session and 1,662 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other commands, from other repositories
rive-interactive-viewmodel_builder
Command "rive-interactive-viewmodel_builder" from freshtechbro/claudedesignskills, covering /rive-interactive-viewmodelbuilder, description, usage, implementation and notes.
video
Create a design based on video.
checklist
Generate a custom checklist for the current feature based on user requirements.
clarify
Identify underspecified areas in the current feature spec by asking up to 5 highly targeted clarification questions and encoding answers back into the spec.
specify
Create or update the feature specification from a natural language feature description.
analyze
Perform a non-destructive cross-artifact consistency and quality analysis across spec.md, plan.md, and tasks.md after task generation.