image-generation

image-generation is a skill for Claude Code, Codex from recomposesh/recompose. It costs 32 tokens per session (2,101 once invoked), scanned C, original, MIT.

A guide for writing prompts for AI image tools such as DALL-E, Midjourney, and Stable Diffusion.

In plain words
What is it for?
Creating image prompts, negative prompts, platform-specific variations, styles, aspect ratios, and alternative concepts.
Why use it?
It helps turn a visual idea into clearer instructions and reduce unwanted results by specifying style, purpose, dimensions, and exclusions.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/recomposesh/recompose/image-generation
Any agent
npx skills add recomposesh/recompose --skill image-generation
Clone the repo
git clone --depth 1 https://github.com/recomposesh/recompose

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for image-generation

README.md
[![agentmods](https://agentmods.dev/badge/skills/recomposesh/recompose/image-generation.svg)](https://agentmods.dev/skills/recomposesh/recompose/image-generation)
Your own site
<a href="https://agentmods.dev/skills/recomposesh/recompose/image-generation"><img src="https://agentmods.dev/badge/skills/recomposesh/recompose/image-generation.svg" alt="Measured on agentmods" height="20"></a>
Per session 32 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,101 The whole file, excluding the scripts and references it only reads on demand.
Security scan C 1 finding. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00032 $0.02101
Opus 5 $0.00016 $0.01051
Sonnet 5 $0.00006 $0.00420
Haiku 4.5 $0.00003 $0.00210

Measured 4d ago against content hash 3a0a93b23716, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade C, and why

image-generation scanned grade C with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Instruction-override phrasinghighPrompt injection

Text telling the model to disregard its earlier instructions or safety rules is the shape of a prompt injection, whoever wrote it.

- Bypass content policies of AI tools
.agents/skills/image-generation/SKILL.md · 375 lines

How it starts

The opening of the file, as written. The whole thing — 375 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Image Generation Skill

Overview

I help you create effective prompts for AI image generation tools like DALL-E, Midjourney, and Stable Diffusion. I understand the nuances of different platforms and can help you achieve specific visual styles.

What I can do:

  • Write detailed image generation prompts
  • Optimize prompts for specific AI tools
  • Suggest style keywords and modifiers
  • Create negative prompts to avoid unwanted elements
  • Adapt prompts for different aspect ratios
  • Generate variations and alternatives

What I cannot do:

  • Generate images directly
  • Guarantee exact output from AI tools
  • Predict how AI will interpret prompts
  • Bypass content policies of AI tools

How to Use Me

Step 1: Describe Your Vision

Tell me:

  • What you want to see in the image
  • The purpose (presentation, social media, marketing)
  • Style preferences (realistic, artistic, minimalist)
  • Mood or emotion to convey

Step 2: Choose the Platform

  • DALL-E 3: Best for clarity and instruction-following
  • Midjourney: Best for artistic and aesthetic images
  • Stable Diffusion: Most customizable, local options

Step 3: Specify Parameters

  • Aspect ratio (1:1, 16:9, 4:3, etc.)
  • Quality level
  • Style references
  • Things to avoid

Prompt Engineering Framework

Basic Prompt Structure

[Subject] + [Action/State] + [Environment] + [Style] + [Technical Parameters]

Detailed Template

[Main Subject]
- Who/what is the focus?
- What are they doing?

[Environment/Setting]
- Where is this taking place?
- Time of day? Weather? Season?

[Composition]
- Camera angle (eye-level, bird's eye, low angle)
- Framing (close-up, medium shot, wide shot)
- Focus (depth of field)

[Style]
- Art style (photorealistic, watercolor, oil painting, etc.)
- Artist reference (optional)
- Era/period

[Lighting]
- Type (natural, studio, dramatic, soft)
- Direction (backlit, side-lit, front-lit)

[Color]
- Palette (warm, cool, monochrome)
- Specific colors to include

[Mood/Atmosphere]
- Emotion to evoke
- Overall feeling

[Technical]
- Quality modifiers
- Aspect ratio
- Negative prompts

Read the full file on GitHub · 375 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 375 lines · 32 tokens per session scan C 3a0a93b23716

Subscribe to this mod's changes

image-generation is a skill published in the GitHub repository recomposesh/recompose (26 stars, last pushed 7d ago), licensed MIT. It adds 32 tokens to every session and 2,101 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it C with 1 finding (instruction-override phrasing). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

image-generation

Create effective AI image generation prompts for DALL-E, Midjourney, and Stable Diffusion. Generate prompts for various styles and use cases.

claude-office-skills/skills · 32 tokens

model-aware-image-prompt-engineer

Model-aware image prompt engineering for any agent or image generation workflow. Use when writing, improving, translating, debugging, or evaluating prompts for OpenAI image models, Gemini Nano Banana, Midjourney, FLUX, Qwen-Image, Z-Image, Stable Diffusion, SDXL, Pony, Illustrious, NoobAI, Animagine, HunyuanImage…

Emily2040/nano-banana-image-skill · 210 tokens

cinematic-frame-director

Turn a user's idea, scene description, or mood into ONE production-grade cinematic still-image prompt in English for text-to-image models (Seedream, Midjourney, Flux, DALL-E, or any image generator). Use this skill whenever the user asks for an image prompt, a cinematic frame, a "movie still" look, a photorealistic…

artnebo/cinematic-frame-director · 156 tokens

nano-banana-image-skill

Model-aware image prompting for Nano Banana Pro and Nano Banana 2. Covers generation, editing, continuity, text-in-image, style routing, and structured JSON output.

Emily2040/nano-banana-image-skill · 40 tokens

vynly-post

A skill that lets an agent publish images it generated to Vynly (vynly.co) - a public, AI-only social feed where every post carries verified provenance (C2PA / SynthID / generator metadata). No signup; two HTTP calls. Use after you create an image the user wants shared publicly.

Vovala14/vynly-mcp · 69 tokens

aigc-prompt-optimizer

把口语化或模糊的 AIGC 创作需求,优化成适合具体工具的专业提示词;支持 prompt battle / 比赛主题发散、Midjourney 出图反馈、二选一选图、视觉问题诊断与迭代改写,并补全构图意图层。支持图片工具(Midjourney、gpt-image、DALL-E、SD、SeaDream)和视频工具(Seedance、Sora、Runway、Kling 等)。当用户说"帮我优化 prompt"、"把这个需求改成 MJ prompt"、"生成 Seedance 提示词"、"prompt battle"、"这张图哪里不好"、"二选一"、"我想做一张/一个视频..."时触发。.

Mr-Salticidae/knowledge-base · 177 tokens