media-higgsfield-explainer

media-higgsfield-explainer is a skill for Claude Code from modu-ai/moai-cowork. It costs 236 tokens per session (5,186 once invoked), scanned A, original, Apache-2.0.

A Higgsfield workflow for making non-photorealistic narrated explainer videos. It pairs each narration line with a 10-second animated clip and joins the clips into one finished video.

In plain words
What is it for?
Use it to turn a topic or document into a faceless explanation, animated story, mascot presentation, or narrated documentary-style video with optional subtitles.
Why use it?
It keeps the narration and visuals aligned across a longer explanation, while preserving one visual style throughout the video.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin. Also seen: names the AskUserQuestion tool; mentions Codex.

Part of the moai-media plugin — 14 skills, 2 agents shipped together

Good fit Use it to turn a topic or document into a faceless explanation, animated story, mascot presentation, or narrated documentary-style video with optional subtitles.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/modu-ai/moai-cowork/media-higgsfield-explainer
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add modu-ai/moai-cowork --skill media-higgsfield-explainer
Clone the repo
git clone --depth 1 https://github.com/modu-ai/moai-cowork

Made for: Claude Code.

Or install moai-media, the plugin that ships this one along with the rest of its 14 skills, 2 agents.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for media-higgsfield-explainer

README.md
[![agentmods](https://agentmods.dev/badge/skills/modu-ai/moai-cowork/media-higgsfield-explainer/github.svg)](https://agentmods.dev/skills/modu-ai/moai-cowork/media-higgsfield-explainer)
Your own site
<a href="https://agentmods.dev/skills/modu-ai/moai-cowork/media-higgsfield-explainer"><img src="https://agentmods.dev/badge/skills/modu-ai/moai-cowork/media-higgsfield-explainer/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for media-higgsfield-explainer

Your own site · 80×15
<a href="https://agentmods.dev/skills/modu-ai/moai-cowork/media-higgsfield-explainer"><img src="https://agentmods.dev/badge/skills/modu-ai/moai-cowork/media-higgsfield-explainer.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 236 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 5,186 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • NVIDIA SkillSpector pass 7 Sept 2026
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00236 $0.05186
Opus 5 $0.00118 $0.02593
Sonnet 5 $0.00047 $0.01037
Haiku 4.5 $0.00024 $0.00519

Measured 8d ago against content hash 81e4c14daa59, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-11, from the pricing page.

Security

Grade A, and why

media-higgsfield-explainer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/moai-media/skills/media-higgsfield-explainer/SKILL.md · 253 lines

How it starts

The opening of the file, as written. The whole thing — 253 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Higgsfield 설명 영상 (media-higgsfield-explainer)

moai-media | 블록 조립형 내레이션 영상 (코어: media-higgsfield-core)

개요

설명 영상은 단발 클립 생성과 다르다. 하나의 스타일 키를 전 블록에 고정하고, 블록마다 내레이션 1줄과 10초 클립 1개를 1:1로 짝지은 뒤, 서버 조립기로 순서대로 이어 붙인다. 이 순서를 어기면 스타일이 흔들리고 음성과 화면이 어긋난다.

호출 계약·비용 프리플라이트·namespace 해석은 코어를 따른다:

  • 호출 계약: ../media-higgsfield-core/references/call-schema.md
  • 잡·비용·리드백: ../media-higgsfield-core/references/job-lifecycle.md

프롬프트 템플릿은 references/prompts.md. 1~3단계 진입 전에 반드시 읽는다.

트리거 키워드

설명 영상, 익스플레이너, explainer, 내레이션 영상, 나레이션, 해설 영상, 애니메이션 설명, 마스코트 영상, 얼굴 없는 영상, 스토리 영상, 다큐 스타일 영상

사용 도구 (MCP)

단계 도구
스타일 프리셋 목록 설명영상 프리셋 조회
프리셋 → 스타일 키 미디어 프리셋 해석
커스텀 스타일 키 생성 generate_image (Nano Banana 계열)
보이스 목록 보이스 조회
내레이션 생성 generate_audio (seed_audio)
클립 생성 generate_video (Gemini Omni 계열)
진행 확인 job_status
최종 조립 설명영상 조립 도구

모델 id는 라이브 조회로 확인한다. 조립은 서버가 한다 — 로컬 ffmpeg나 수동 이어붙이기를 쓰지 않는다.

하드 규칙

이 규칙들은 결과 품질이 아니라 성립 여부를 가른다.

  • 모든 화면은 비실사를 유지한다. 같은 STYLE 서술과 사실주의 금지어를 매 블록 프롬프트에 반복한다.
  • 클립에는 말소리가 들어가지 않는다. 클립 오디오는 앰비언스·음악뿐이며 대사·립싱크·내레이션을 넣지 않는다. 목소리는 내레이션 트랙에서만 온다.
  • 블록당 내레이션 1개, 클립 1개. N번 오디오는 반드시 N번 영상에 붙는다.
  • 같은 스타일 키 이미지를 모든 클립에 첨부한다.
  • 이미지·영상 프롬프트는 영어로 쓴다. 내레이션만 사용자가 고른 언어로 쓴다.
  • 실제 주제는 조사한 뒤 대본을 쓴다. 인용·날짜·수치·사건을 지어내지 않는다.
  • 같은 실행 안에서 조립까지 끝낸다. 클립만 흩어놓고 끝내면 실패다.

워크플로우

0단계 — 두 번에 나눠 묻기 (합치지 않는다)

이 스킬은 사용자에게 직접 묻지 않는다. 아래 슬롯을 두 라운드로 나눠 수집하도록 오케스트레이터에 blocker로 요청한다. 한 번에 몰아 묻지 않는 이유는 스타일 선택이 나머지 결정의 전제이기 때문이다.

라운드 1 — 스타일만. 프리셋 목록을 라이브 조회해 이름과 미리보기를 제시하고, 프리셋 선택 / 직접 서술 / 참조 이미지 첨부 중 하나를 받는다. 스타일 선택은 필수이며, 사용자가 명시적으로 위임하지 않는 한 임의로 고르지 않는다.

라운드 2 — 제작 설정. 스타일이 정해진 뒤에만 묻는다.

슬롯 기본
길이 1~10분 정수. 블록 수 N = 분 × 6
내레이션 언어 영어 선택지를 준다
캐릭터 마스코트 / 무인물. 항상 묻는다
화면비 프리셋을 고르면 9:16 16:9 / 9:16 — 아래 주의
자막 켜면 폰트를 고르게 한다(임의 선택 금지). 음성 블록당 추가 비용 발생을 알린다

화면비 주의 (라이브 관측). CMS 프리셋은 저술 시점 기준 전부 9:16 세로다. 따라서 프리셋을 고른 뒤 16:9를 요구하면 프리셋 참조와 충돌한다. 가로형이 꼭 필요하면 프리셋 대신 커스텀 스타일 키로 가는 것이 정상 경로다. 프리셋 목록의 aspect 값은 고정이 아니므로 매번 조회 결과를 확인하고, 프리셋을 고른 경우 그 aspect를 기본값으로 삼는다.

Read the full file on GitHub · 253 lines

Files

What ships with it

1 file beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 8d ago First seen · 253 lines · 236 tokens per session scan A 81e4c14daa59

Subscribe to this mod's changes

media-higgsfield-explainer is a skill published in the GitHub repository modu-ai/moai-cowork (300 stars, last pushed 8d ago), licensed Apache-2.0. It adds 236 tokens to every session and 5,186 once invoked, about $0.0012 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.

Related

Other skills, from other repositories

baoyu-comic

A creator for educational comics, including biographies and tutorials, that turns supplied content or topics into illustrated comic pages.

NousResearch/hermes-agent · 17 tokens

concept-diagrams

Generate flat, minimal educational SVG visuals as HTML.

NousResearch/hermes-agent · 14 tokens

infographic-v2

Generate professional infographics using Nano Banana MCP (Gemini AI image generation). Follows a guided flow - analyze content, suggest visualizable concepts, propose visualization approaches, then generate on-brand images. USE THIS SKILL WHEN user says "create infographic v2", "make a visual v2", "infographic-v2".…

naveedharri/benai-skills · 81 tokens

excalidraw

Create visual presentations, slide decks, and explanatory diagrams in Excalidraw. Use when user asks to create a presentation, slide deck, visual explainer, pitch deck, comparison diagram, process flow, or any multi-slide visual content. Supports two output modes — generating .excalidraw JSON files OR injecting slides…

naveedharri/benai-skills · 99 tokens

marketing-os-carousel

Build an image-first social carousel from an asset already filed in the Marketing OS, export it as a PDF, and record it back as a real channel asset. Brand palette, typography, the logo pointer and the never-black-background rule all resolve from Context/brand/brand-kit.md. Source is a filed newsletter edition…

naveedharri/benai-skills · 221 tokens

youtube-brief

Create a detailed video brief for a new YouTube video through a structured, collaborative process. This is a STEP-BY-STEP, interactive process — never output a complete brief immediately. Each step requires suggestions, user decision, then progression to the next step. USE THIS SKILL WHEN: - User wants to plan a…

naveedharri/benai-skills · 205 tokens