Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/skillmedev/creator-studio/thumbnail-conceptnpx skills add SkillMedev/creator-studio --skill thumbnail-conceptgit clone --depth 1 https://github.com/SkillMedev/creator-studioWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/skillmedev/creator-studio/thumbnail-concept)<a href="https://agentmods.dev/skills/skillmedev/creator-studio/thumbnail-concept"><img src="https://agentmods.dev/badge/skills/skillmedev/creator-studio/thumbnail-concept.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00146 | $0.01512 |
| Opus 5 | $0.00073 | $0.00756 |
| Sonnet 5 | $0.00029 | $0.00302 |
| Haiku 4.5 | $0.00015 | $0.00151 |
Grade A, and why
Thumbnail Concept scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 79 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Thumbnail Concept
A thumbnail is not decoration; it is the first ad for the video, and on YouTube the click package (thumbnail plus title) decides whether the video gets a chance at all - a strong video with a weak thumbnail loses to a mediocre video with a strong one. This skill produces concrete thumbnail concepts before filming or design, so the creator enters production knowing exactly what click signal they are building toward. The costly mistake it prevents is designing the thumbnail last, from whatever footage happens to exist.
Operating procedure
Step 1: gather inputs
- The video's core promise and emotional payoff (curiosity satisfied? problem fixed? disbelief resolved?).
- Whether the creator appears on camera and is willing to shoot dedicated thumbnail expressions. Default: yes if they are a face-forward channel.
- The channel's existing thumbnail style, so concepts either fit the brand or deliberately break pattern - name which.
- Any concrete visual assets: a result, a before/after, a prop, a screenshot. If none are stated, propose them and label them as suggestions.
Step 2: build each concept on the CTR triangle
Every high-performing thumbnail has three elements working together, concepted together, never in isolation:
- A clear focal point - a face with a strong expression, a dramatic object, or a before/after split.
- A text overlay of 2-5 words that adds information the image alone does not carry. Three words is the working target; five is the hard ceiling.
- Color contrast that pops against both YouTube's white and dark UI backgrounds.
If any leg is weak, CTR drops. A great face with redundant text, or sharp copy on a muddy image, both fail.
Step 3: direct the facial expression precisely
When a face is in frame, the expression does most of the emotional signaling. Never write "surprised" - write "mouth open, eyes wide, leaning toward camera." The expression must match the video's emotional payoff; a shocked face on a calm tutorial is a bait signal. Curiosity, disbelief, and genuine excitement are the three highest-performing emotion categories. Ban neutral expressions and posed smiles - they read as stock photography and get skipped.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 3d ago First seen · 79 lines · 146 tokens per session scan A 03bc77bc8485
Thumbnail Concept is a skill published in the GitHub repository SkillMedev/creator-studio (5 stars, last pushed 2mo ago), licensed MIT. It adds 146 tokens to every session and 1,512 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
remocn
Build Remotion videos with remocn — copy-paste animation components and timeline-driven UI primitives from a shadcn registry. Use when composing a video or scene in a Remotion project, adding a single animation, transition, background, or UI-block sim, or reaching for a video-ready UI primitive (button, dialog…
openstoryline-use
Use this skill when OpenStoryline is already installed and the user wants to start the local MCP/Web services, create or continue a session, send editing instructions, perform multi-turn re-editing, and verify rendered video outputs, as well as Chinese requests like “启动 OpenStoryline”, “把 OpenStoryline 跑起来”, “用…
create_profile_style_skill
【META SKILL】分析当前剪辑逻辑与风格,总结并生成一个新的可复用 Skill 文件,存入剪辑技能库。Analyze the current editing logic and style, summarize and generate a new reusable Skill file, and store it in the editing skill library.
subtitle_imitation_skill
【CAPABILITY SKILL】基于用户提供的参考文案样本,对视频素材内容进行深度文风仿写,生成风格化脚本。Based on user-provided reference text samples, the video material is deeply rewritten in terms of writing style to generate a stylized script.
videoagent-audio-studio
Tired of juggling multiple audio APIs? This skill gives you one-command access to TTS, music generation, sound effects, and voice cloning. Use when you want to generate any audio without managing multiple API keys.
videoagent-image-studio
Tired of juggling 8 API keys? This skill gives you one-command access to Midjourney, Flux, Ideogram, and more, with zero setup. Use when you want to generate any image without worrying about API keys.