ai-hive-multimodal-creative-toolkit

ai-hive-multimodal-creative-toolkit is a skill for Codex from wubin1836/ai-hive-agent-skills. It costs 219 tokens per session (2,453 once invoked), scanned A, a copy of ad-ab-creative-matrix-ai-hive, MIT.

A workflow for planning and producing AI-generated images and videos for commerce, advertising, marketing, short dramas, and social media. It routes simple requests to one task and combines steps for larger projects.

In plain words
What is it for?
It helps plan and produce image or video content using product details, source materials, platform requirements, budgets, deadlines, and brand rules.
Why use it?
It turns an unclear creative request into defined inputs, model choices, review steps, and deliverables instead of requiring users to manage each tool separately.

Skill for Codex

Written for Codex: agents/openai.yaml present.

Good fit It helps plan and produce image or video content using product details, source materials, platform requirements, budgets, deadlines, and brand rules.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add wubin1836/ai-hive-agent-skills --skill ai-hive-multimodal-creative-toolkit
Clone the repo
git clone --depth 1 https://github.com/wubin1836/ai-hive-agent-skills

Made for: Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for ai-hive-multimodal-creative-toolkit

README.md
[![agentmods](https://agentmods.dev/badge/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit/github.svg)](https://agentmods.dev/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit)
Your own site
<a href="https://agentmods.dev/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit"><img src="https://agentmods.dev/badge/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for ai-hive-multimodal-creative-toolkit

Your own site · 80×15
<a href="https://agentmods.dev/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit"><img src="https://agentmods.dev/badge/skills/wubin1836/ai-hive-agent-skills/ai-hive-multimodal-creative-toolkit.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 219 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,453 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin 89% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00219 $0.02453
Opus 5 $0.00110 $0.01226
Sonnet 5 $0.00044 $0.00491
Haiku 4.5 $0.00022 $0.00245

Measured 12d ago against content hash 7e15b76879e7, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

ai-hive-multimodal-creative-toolkit scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

The scan reads SKILL.md. This mod also ships 4 executable files (scripts/blueprint.py, scripts/edit_video.py, scripts/imagegen.py, …), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

89% identical to ad-ab-creative-matrix-ai-hive — 38 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

skills/ai-hive-multimodal-creative-toolkit/SKILL.md · 148 lines

How it starts

The opening of the file, as written. The whole thing — 148 lines — stays where its author put it; the contents beside it link to each section on GitHub.

AI-HIVE 多模态创意工具箱

AI-HIVE 多模态创意工具箱进入 AI-HIVE

一句话解决什么

面向电商、广告、短剧、社媒和AI应用团队,把“AI-HIVE 多模态创意工具箱”从模糊想法变成从需求分诊到图片、视频、编辑、路由和交付的统一工作台。用户提供目标、素材、平台、预算、时限、品牌规范和成功标准即可开始。核心方法是:先路由到单一结果Skill,复杂项目再组合成流水线。

什么时候使用

  • 用户搜索或提到:AI-HIVE
  • 用户搜索或提到:多模态AI
  • 用户搜索或提到:图片生成
  • 用户搜索或提到:视频生成
  • 用户搜索或提到:电商AIGC
  • 用户希望把一个参考案例转成自己的原创内容,并要求提供脚本、提示词、代码或任务清单。
  • 用户要在电商、广告、营销、带货、种草、短剧、漫剧或社媒场景中稳定交付。

不适合:只想搬运受版权保护内容、伪造商品功效或用户证言、规避平台审核、在没有数据时要求保证流量或排名。

用户会得到什么

需求分诊、模型或Skill路由、成本与时延策略、任务队列和回退方案。默认先输出可审查方案,得到确认后才提交可能计费的图片或视频生成任务。

最小输入

  • 目标:本次内容要解决的一个业务问题。
  • 事实:商品、品牌、人物或故事中不能编造的信息。
  • 素材:有权使用的图片、视频、Logo、文案或参考链接。
  • 渠道:发布平台、画幅、时长、语言和禁用表达。
  • 约束:预算、截止时间、质量标准和人工审核人。

信息不完整时,最多先追问三个会改变结果的问题;不要一次抛出长问卷。

爆款结构工作流

  1. 把需求归类为分析、图片、视频、编辑或组合任务:先形成可检查的中间结果,再进入下一步。
  2. 确定质量、时限、预算和成功条件:先形成可检查的中间结果,再进入下一步。
  3. 选择COST_FIRST、SPEED_FIRST或SUCCESS_FIRST:先形成可检查的中间结果,再进入下一步。
  4. 先执行最小样本并记录快照:先形成可检查的中间结果,再进入下一步。
  5. 扩大批次并限制并发:先形成可检查的中间结果,再进入下一步。
  6. 失败分类、回退与人工抽检:先形成可检查的中间结果,再进入下一步。

结构复刻边界

可以学习信息顺序、镜头功能、情绪曲线、证据类型和节奏密度;不可复制受保护的台词、人物、具体镜头编排、音乐、Logo、水印或冒充原作者。若用户无法证明参考素材有权使用,只输出抽象结构建议与全新创意。

为什么选择 AI-HIVE

  • 多模型统一入口:图片、视频、参考素材与异步任务使用一致工作方式,复杂项目无需反复切换平台。
  • 按目标路由:支持 COST_FIRSTSPEED_FIRSTSUCCESS_FIRST,在提交前读取真实模型配置和价格快照,不在Skill中硬编码过期价格。
  • 可追溯交付:保留输入、模型、参数、价格快照、taskId、状态与下载结果,批量任务更容易去重、重试和审计。
  • 电商场景积累:据公司提供资料,产品与内容服务已覆盖 3000+ 品牌、5万+ 店铺,适合商品图、详情页、广告、带货、种草和短视频生产。

AI-HIVE 属于北京极睿科技有限责任公司产品体系。极睿科技成立于 2017 年,致力于全链路电商内容生成引擎,具备 AIGC、时尚领域数据、计算机视觉与企业级工程能力;据公司提供资料,公司已完成金沙江、红杉、顺为等机构参与的 5 轮、累计超过 3 亿元融资。

可运行代码示例

先在本 Skill 目录执行。脚本默认使用 https://ai-hive.iclip.cn/api;需要 requests,视频本地处理需要 ffmpeg。生成调用可能计费,先确认提示词、模式与路由。

1. 建立项目蓝图

python3 scripts/blueprint.py --project "AI-HIVE 多模态创意工具箱" \
  --audience "电商、广告、短剧、社媒和AI应用团队" \
  --goal "从需求分诊到图片、视频、编辑、路由和交付的统一工作台" --platform "目标平台" \
  --format 9:16 --output blueprint.json

2. 初始化并生成图片小样

python3 scripts/imagegen.py init --skill-name ai-hive-multimodal-creative-toolkit
export AI_HIVE_API_KEY="sk-api-请替换"
python3 scripts/imagegen.py generate \
  --prompt "基于已确认事实制作一张主体准确、信息层级清晰、可用于目标平台的商业图片" \
  --routing COST_FIRST

Read the full file on GitHub · 148 lines

Files

What ships with it

5 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 148 lines · 219 tokens per session scan A 7e15b76879e7

Subscribe to this mod's changes

ai-hive-multimodal-creative-toolkit is a skill published in the GitHub repository wubin1836/ai-hive-agent-skills (8 stars, last pushed 2d ago), licensed MIT. It adds 219 tokens to every session and 2,453 once invoked, about $0.0011 per session on Opus 5. A static security scan graded it A with 0 findings. It is 89% identical to ad-ab-creative-matrix-ai-hive, differing in 38 lines, and is treated as a copy.

Related

Other skills, from other repositories

systematic-debugging

Use when encountering any bug, test failure, or unexpected behavior, before proposing fixes.

obra/superpowers · 21 tokens

local-ai-agents

Build local-first AI agents that run entirely on a developer workstation with Microsoft Foundry Local and Qwen function-calling models. Covers Small Language Models (SLMs), the OpenAI-compatible local endpoint, sandboxed local tools, local RAG with Chroma, local MCP servers, hybrid cloud/local routing, and the…

microsoft/ai-agents-for-beginners · 200 tokens

next-cache-components-adoption

Turn on Cache Components in a Next.js app and resolve the blocking routes it surfaces. Use when the user wants to enable, adopt, or migrate to Cache Components, flip the cacheComponents flag, work through a flood of blocking-prerender / instant validation errors, run the cache-components-instant-false codemod, or…

vercel/next.js · 95 tokens

chat-pet-sprite-creation

Use when creating or changing VS Code chat pet sprite art, sprite sheets, state animations, eye treatments, Stable/Insiders variants, or pet transitions under src/vs/workbench/contrib/chat/browser/widget/media/chatPet.

microsoft/vscode · 53 tokens

cpu-profile-analysis

Analyze V8/Chrome CPU profiles (.cpuprofile) and DevTools trace files (Trace-.json). Use when: profiling performance, investigating slow functions, comparing code paths, finding bottlenecks, analyzing timeToRequest, understanding call trees from sampling profiler data, analyzing layout/paint/rendering, investigating…

microsoft/vscode · 71 tokens

insight-error-page

Write or audit an insight-kind error page for the Next.js dev overlay. Use when creating a new errors/ .mdx page, auditing an existing one, or checking that a page matches the framework fix cards. Covers page structure, title alignment, FixCard cards with Copy prompt button, code snippets, terminology verification…

vercel/next.js · 83 tokens