captions-overlay

captions-overlay is a skill for Claude Code from skillmds/skillmd. It costs 136 tokens per session (1,264 once invoked), scanned A, a copy of captions-overlay, MIT.

A set of rules for placing captions on talking-head or launch videos. Captions are text shown over the video, and this guide treats them as an overlay rather than reserving a separate empty area.

In plain words
What is it for?
Use it when adding subtitles to video to choose between dropping filler, showing ordinary speech in a lower rail, or embedding occasional text within the scene.
Why use it?
It prevents captions from pushing the composition upward or creating a dead band, while defining when spoken words should be shown or omitted.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin.

Part of the data-ml plugin — 17 skills shipped together

Good fit Use it when adding subtitles to video to choose between dropping filler, showing ordinary speech in a lower rail, or embedding occasional text within the scene.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/skillmds/skillmd/captions-overlay
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add skillmds/skillmd --skill captions-overlay
Clone the repo
git clone --depth 1 https://github.com/skillmds/skillmd

Made for: Claude Code.

Or install data-ml, the plugin that ships this one along with the rest of its 17 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for captions-overlay

README.md
[![agentmods](https://agentmods.dev/badge/skills/skillmds/skillmd/captions-overlay/github.svg)](https://agentmods.dev/skills/skillmds/skillmd/captions-overlay)
Your own site
<a href="https://agentmods.dev/skills/skillmds/skillmd/captions-overlay"><img src="https://agentmods.dev/badge/skills/skillmds/skillmd/captions-overlay/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for captions-overlay

Your own site · 80×15
<a href="https://agentmods.dev/skills/skillmds/skillmd/captions-overlay"><img src="https://agentmods.dev/badge/skills/skillmds/skillmd/captions-overlay.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 136 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,264 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin 100% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00136 $0.01264
Opus 5.5 $0.00054 $0.00506
Sonnet 5 $0.00027 $0.00253
Haiku 4.5 $0.00014 $0.00126

Measured 4d ago against content hash c43fd6695fd5, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-23, from the pricing page.

Security

Grade A, and why

captions-overlay scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

100% identical to captions-overlay — 4 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

plugins/data-ml/skills/captions-overlay/SKILL.md · 82 lines

How it starts

The opening of the file, as written. The whole thing — 82 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Captions Overlay Doctrine

Overlay doctrine — supplements the upstream embedded-captions skill. Applies ON TOP of it; do not expect it folded into the upstream skill.

Two ideas combine here. First, the caption model — every spoken phrase is drop, rail, or embed, and embed is the scarce earned peak, not the default. Second, the overlay law — a caption line is composited ON TOP of the film as an overlay; it is NOT a reserved zone, so you never shift content up or leave a dead band to "make room" for it. The two reinforce each other: because captions ride as an overlay (the verbatim rail in front, the occasional embed behind the subject), the composition keeps its full frame and centers on the true vertical center.

The caption model — drop / rail / embed

Every spoken phrase is one of three things (verbatim from embedded-captions):

What How it's shown
drop filler — um/uh, stutters, self-corrections not shown
rail the default — ordinary spoken content (verbatim) clean lower-third subtitle, in front, readable. A punch word can get an inline emphasis highlight (accent colour / active-word pop) — it stays on the rail.
embed a promoted peak — the headline beat one big word composited behind the subject (matte occlusion), designed entrance + exit

Read the full file on GitHub · 82 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 82 lines · 136 tokens per session scan A c43fd6695fd5

Subscribe to this mod's changes

captions-overlay is a skill published in the GitHub repository skillmds/skillmd (1 stars, last pushed yesterday), licensed MIT. It adds 136 tokens to every session and 1,264 once invoked, about $0.0005 per session on Opus 5.5. A static security scan graded it A with 0 findings. It is 100% identical to captions-overlay, differing in 4 lines, and is treated as a copy.