talking-head-hyperframes

talking-head-hyperframes is a skill for Codex from bozhouDev/video-skills-toolkit. It costs 150 tokens per session (1,957 once invoked), scanned A, original, MIT.

A production skill for preparing and checking fixed talking-head video templates for HyperFrames, including a picture-in-picture presenter area. It archives inputs, checks media and subtitles, and creates a handoff for later scene work.

In plain words
What is it for?
Creating or repairing a presenter template, checking audio and timed subtitles, verifying production specifications and motion coverage, and recording readiness or missing inputs.
Why use it?
It confirms that the stage, source files, timings, and required production information are valid before animation and scene implementation begin.

Skill for Codex

Written for Codex: agents/openai.yaml present.

Needs its repository: it runs a file that does not travel with it, so clone the repository first. The line is node scripts/scaffold_talking_head_hyperframes_project.mjs --project-dir /absolute/path/to/project [input flags].

Good fit Creating or repairing a presenter template, checking audio and timed subtitles, verifying production specifications and motion coverage, and recording readiness or missing inputs.

Compare 6 skills from other repositories ↓
Install

Getting it into your agent

It runs from inside its repository, so the clone comes first — what it calls does not travel with the file alone.

Clone the repo
git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit
agentmods
npx agentmods add skills/bozhoudev/video-skills-toolkit/talking-head-hyperframes

Made for: Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for talking-head-hyperframes

README.md
[![agentmods](https://agentmods.dev/badge/skills/bozhoudev/video-skills-toolkit/talking-head-hyperframes.svg)](https://agentmods.dev/skills/bozhoudev/video-skills-toolkit/talking-head-hyperframes)
Your own site
<a href="https://agentmods.dev/skills/bozhoudev/video-skills-toolkit/talking-head-hyperframes"><img src="https://agentmods.dev/badge/skills/bozhoudev/video-skills-toolkit/talking-head-hyperframes.svg" alt="Measured on agentmods" height="20"></a>
Per session 150 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,957 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • NVIDIA SkillSpector pass 7 Sept 2026
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00150 $0.01957
Opus 5 $0.00075 $0.00979
Sonnet 5 $0.00030 $0.00391
Haiku 4.5 $0.00015 $0.00196

Measured 7d ago against content hash 6693ecf8f8f8, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-07, from the pricing page.

Security

Grade A, and why

talking-head-hyperframes scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 7d ago.

The scan reads SKILL.md. This mod also ships 20 executable files (assets/hyperframes-project/scripts/analyze-preview.mjs, assets/hyperframes-project/scripts/build-animation-map.mjs, assets/hyperframes-project/scripts/build-contact-sheet.mjs, …), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/talking-head-hyperframes/SKILL.md · 85 lines

How it starts

The opening of the file, as written. The whole thing — 85 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Talking Head HyperFrames · Template Factory

只产出经过舞台证明的固定模板和可审计交接。内容执行属于 hyperframes-scene-animator

先路由意图

  • 只建或修复固定舞台、归档输入、检查就绪状态:由本 Skill 执行。
  • 用户要求制作镜头,而工程尚无合格模板:先完成本 Skill 的职责,再按交接状态停止或路由。
  • 工程已经 READY_FOR_EXECUTION,且用户只要求镜头实现、语义动效、转场、声音、proof 或整改:直接路由到 hyperframes-scene-animator,不要复制其流程。
  • 用户明确调用本 Skill 做内容场景时,仍按上述边界路由,不能越权代做。

生成或验证工程前,必须读取当前启用的 hyperframeshyperframes-cli skills。不要复制它们的运行时规则;所有 scaffold、校验和舞台 proof 都通过生成项目公开的 package scripts 复现。若这些脚本或实现资源尚未提供,明确报告缺少的能力并停止,不要临时发明替代工程。

唯一职责

本 Skill 负责:

  • 创建或修复固定模板与稳定根 composition;
  • 归档用户已经提供的锁定输入,保留源路径、项目相对路径、大小与 SHA-256;
  • 校验媒体可解码性和时长、字幕 schema/时间、motion contract 覆盖与 source hash;
  • 写 manifest 和 template handoff;
  • 运行只证明固定舞台的结构检查与 stage proof;
  • 根据就绪状态停止或路由下游。

本 Skill 不导演镜头,不写内容场景,不选择语义动效或转场,不做 SFX/BGM 设计,不渲染整片,不审片,也不整改内容实现。不得为了演示而加入标题页、三卡片、流程图、TopBar、通用 enterStyle 或全局 crossfade。

就绪状态

允许缺少正式输入时创建固定模板,但必须把状态写成 TEMPLATE_ONLY。manifest 与 handoff 应逐项列出每个缺失或无效输入,并给出期望路径、实际路径以及可获得的 hash/校验结果;不得伪造占位输入或假装就绪。

只有以下各项全部有效且相互一致时才写 READY_FOR_EXECUTION

  • 可解码且有精确时长的锁定完整人声音频;
  • 非空、时间有效、与音频时长一致的对齐字幕;
  • 用户已确认的导演脚本与制作规格;
  • 素材计划;
  • 覆盖整片、字段完整的 motion contract;
  • 上述锁定输入的源路径、项目相对路径和 SHA-256,以及 contract source hash 比对结果。

口播视频、录屏、截图和字体按实际提供情况归档并记录;缺少可选媒体不冒充必填失败。输入矛盾、导演未确认或 source hash 不一致时保持 TEMPLATE_ONLY 并停止升级状态。

脚手入口

生成前读 references/project-layout.md 确认固定层与下游所有权,再读 references/input-handoff.md 执行输入、hash 和就绪门。

只要挂载数字人或 PIP 占位,还必须完整读取 references/pip-contract.md。PIP 不是一个可随手摆放的视频:模板必须锁定外圈、媒体区、层级、裁切、人物安全区、不透明背景和审查证据。当前 Bozhou Digital Twin 的默认裁切是 object-fit: cover; object-position: 66.7% 50%;更换人物时允许用脚手参数覆盖,但覆盖值必须进入 manifest 并通过静态页与成片放大审查。

node scripts/scaffold_talking_head_hyperframes_project.mjs --project-dir /absolute/path/to/project [input flags]
node scripts/validate_inputs.mjs /absolute/path/to/project
node scripts/validate_fixed_stage.mjs /absolute/path/to/project

生成后必须在项目内运行 npm run check:inputsnpm run check:fixed-stage;需要完整 HyperFrames 静态/运行门时运行 npm run check:template。任何调用 HyperFrames 的 package script 必须保持精确版本 0.7.65

Read the full file on GitHub · 85 lines

Files

What ships with it

60 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 7d ago First seen · 85 lines · 150 tokens per session scan A 6693ecf8f8f8

Subscribe to this mod's changes

talking-head-hyperframes is a skill published in the GitHub repository bozhouDev/video-skills-toolkit (138 stars, last pushed 1mo ago), licensed MIT. It adds 150 tokens to every session and 1,957 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

video-hyperframes

Hyperframes / Remotion-compatible continuous frame animation with autoplay support.

nexu-io/open-design · 18 tokens

video-hyperframes

A web-based sequence of video frames designed for Hyperframes or Remotion, with each frame presenting one visual idea.

nexu-io/html-anything · 22 tokens

remocn

Build Remotion videos with remocn — copy-paste animation components and timeline-driven UI primitives from a shadcn registry. Use when composing a video or scene in a Remotion project, adding a single animation, transition, background, or UI-block sim, or reaching for a video-ready UI primitive (button, dialog…

Remocn/remocn · 88 tokens

minimax-cli

Nested swiss-knife reference for the MiniMax mmx CLI and the canonical MiniMax CLI procedure shipped with the TUI: install mmx-cli, discover the correct TUI-managed MiniMax preset/key slot without leaking secrets, match mainland vs international regions, and route image/video/music/TTS generation or one-shot shell…

Lingtai-AI/lingtai · 72 tokens

videoagent-audio-studio

Tired of juggling multiple audio APIs? This skill gives you one-command access to TTS, music generation, sound effects, and voice cloning. Use when you want to generate any audio without managing multiple API keys.

pexoai/pexo-skills · 50 tokens

multimodal-llm

Vision, audio, video generation, and multimodal LLM integration patterns. Use when processing images, transcribing audio, generating speech, generating AI video (Kling v3, Sora 2, Veo 3.1 std/lite/fast, Runway Gen-4.5 via gen4turbo), or building multimodal AI pipelines.

yonatangross/orchestkit · 82 tokens