shot-specifier

shot-specifier is a skill for Claude Code, Codex from leynos/visual-storytelling-skills. It costs 153 tokens per session (6,333 once invoked), scanned A, original, ISC.

A workflow for breaking a completed scene plan into numbered video shots with camera, actor, lighting, timing, and effects directions. The result includes storyboard images, prompts, model choices, and asset instructions.

In plain words
What is it for?
Use it to prepare generation-ready shot specifications from a scene inventory, reference images, prompt keywords, and continuity information.
Why use it?
It turns broad scene descriptions into specific production instructions and helps maintain visual continuity across shots.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit Use it to prepare generation-ready shot specifications from a scene inventory, reference images, prompt keywords, and continuity information.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/leynos/visual-storytelling-skills/shot-specifier
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add leynos/visual-storytelling-skills --skill shot-specifier
Clone the repo
git clone --depth 1 https://github.com/leynos/visual-storytelling-skills

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for shot-specifier

README.md
[![agentmods](https://agentmods.dev/badge/skills/leynos/visual-storytelling-skills/shot-specifier/github.svg)](https://agentmods.dev/skills/leynos/visual-storytelling-skills/shot-specifier)
Your own site
<a href="https://agentmods.dev/skills/leynos/visual-storytelling-skills/shot-specifier"><img src="https://agentmods.dev/badge/skills/leynos/visual-storytelling-skills/shot-specifier/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for shot-specifier

Your own site · 80×15
<a href="https://agentmods.dev/skills/leynos/visual-storytelling-skills/shot-specifier"><img src="https://agentmods.dev/badge/skills/leynos/visual-storytelling-skills/shot-specifier.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 153 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 6,333 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00153 $0.06333
Opus 5 $0.00077 $0.03166
Sonnet 5 $0.00031 $0.01267
Haiku 4.5 $0.00015 $0.00633

Measured 11d ago against content hash e992a36c8719, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-11, from the pricing page.

Security

Grade A, and why

shot-specifier scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/shot-specifier/SKILL.md · 601 lines

How it starts

The opening of the file, as written. The whole thing — 601 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Shot Specifier

Takes a completed scene inventory and reference image set as input and produces generation-ready shot specifications: numbered shots with full directorial direction, storyboard keyframes, video prompts, model routing, and an asset pipeline.

This skill is explicitly expensive. Generating storyboard images for every shot in every scene requires many image generation calls. Per-shot video generation is more expensive still. This cost is the price of consistency and production quality — do not skip shots or reduce keyframe coverage to save cost unless the user explicitly requests it.

Input Requirements

Before beginning, verify the following inputs are available:

Input Location Required?
Scene inventory document {project}/scene-pack/{project}_scene_inventory.md Required
Reference images {project}/image_out/ or {project}/refs/ Required
Prompt keyword library {project}/{project}_prompt_keywords.md Required
Continuity inventory {project}/{project}_continuity_inventory.md Strongly recommended
Video role manifest Section in scene inventory Required

If the prompt keyword library or video role manifest are absent, run scene-inventory-extractor-v2 phases 2.4 and 11.6 first before proceeding.

Execution Context

This skill requires:

  • nanobanana image generation Model Context Protocol (MCP) (see references/storyboard-generation.md), with every storyboard image call explicitly using model: gemini-3-pro-image-preview
  • Vision capabilities (for storyboard consistency verification)
  • File system access (structured output directories)

Generation runs silently — no user confirmation gates during storyboard or prompt phases. Halt only on consistency failures that require human judgement. If gemini-3-pro-image-preview is unavailable through nanobanana, or if it cannot accept the required reference images or character-consistency images for the current storyboard operation, STOP and report the blocker. Do not continue with a fallback image model.

Read the full file on GitHub · 601 lines

Files

What ships with it

5 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 11d ago First seen · 601 lines · 153 tokens per session scan A e992a36c8719

Subscribe to this mod's changes

shot-specifier is a skill published in the GitHub repository leynos/visual-storytelling-skills (6 stars, last pushed 1mo ago), licensed ISC. It adds 153 tokens to every session and 6,333 once invoked, about $0.0008 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

wavespeed

Generate or edit AI media (image, video, audio, 3D) by calling the wavespeed CLI on the user's machine. Use whenever the user asks to create, edit, animate, upscale, or transform a visual asset, generate audio/TTS/music, or produce marketing creatives. Every model on the WaveSpeed platform is one wavespeed run call.

WaveSpeedAI/wavespeed-cli · 79 tokens

1688-ecommerce-video-generation-editing

A workflow for generating and editing 1688 e-commerce videos, including videos made from text, images, existing video, or audio.

wubin1836/ai-hive-agent-skills · 463 tokens

1688-ecommerce-image-generation-editing

A tool for creating and editing product images for 1688, Alibaba's Chinese wholesale marketplace, and similar sales channels. It can work from text or reference images to produce product pages, advertisements, posters, and social-media visuals.

wubin1836/ai-hive-agent-skills · 457 tokens

ai-model-expert-1688-ecommerce-image-generation-editing

An AI image-generation and editing workflow for 1688 e-commerce content, using text or reference images to guide the result. 1688 is a Chinese online wholesale marketplace.

wubin1836/ai-hive-agent-skills · 326 tokens

ai-model-expert-1688-ecommerce-video-generation-editing

An AI tool workflow for creating and editing 1688 e-commerce videos from text, images, videos, or audio. 1688 is a Chinese online wholesale marketplace.

wubin1836/ai-hive-agent-skills · 330 tokens

ai-model-expert-ai-ad-video

AI video production support for turning text, images, videos, or audio into advertising and other commercial videos.

wubin1836/ai-hive-agent-skills · 352 tokens