calesthio/generative-media-skills

Research-backed agent skills and tools for premium image, video, audio, voice, and generative media production across AI coding assistants.

This repository also configures its own agents. See what generative-media-skills tells them →

170Stars on the repository
158Mods indexed here, across every type
2mo agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

editing-montage

49

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent editing and montage direction for AI agents producing generated videos, ads, trailers, social clips, explainers, product videos, documentary-style pieces, and music- or beat-driven cuts. Use when planning, revising, QAing, or handing off an edit: story structure, shot selection, continuity…

not rated 170 +20 2mo ago A SkillSpector: pass 105 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent motion graphics direction for AI agents producing title sequences, lower thirds, explainer graphics, product UI overlays, social ads, kinetic typography, diagrams, data callouts, brand films, and video post. Use when planning, prompting, storyboarding, specifying, implementing, reviewing, or QAing…

not rated 170 +20 2mo ago A SkillSpector: warn 97 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent still-image post-production skill for RAW/rendered intake, nondestructive development, exposure, white balance, tone, color, ICC proofing, masking, repair, compositing, truthful retouching by genre, generative disclosure, sharpening, noise, resampling, variants, metadata, export, QA, rights, and…

not rated 170 +20 2mo ago A SkillSpector: pass 130 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent title design and kinetic typography direction for video, animation, ads, explainers, social hooks, brand films, lyric/text videos, title cards, main titles, lower thirds, captions-as-design, and other motion-led text. Use when an agent must plan, prompt, implement, hand off, iterate, or QA…

not rated 170 +20 2mo ago A SkillSpector: pass 114 tokens original MIT

vfx-compositing

53

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent VFX compositing direction for generated video, ads, trailers, product films, social clips, explainers, mixed-source edits, and image/video composites. Use when an agent must plan, brief, generate, integrate, repair, or quality-check visual effects composites involving plates, alpha/mattes, keying…

not rated 170 +20 2mo ago A SkillSpector: warn 123 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production guidance for translating measured audio features into deterministic video timing and motion. Use for beat-, onset-, phrase-, energy-, silence-, or spectrum-reactive visualizers, edits, typography, and generated compositions; not for music-video concept direction, audio generation, or a…

not rated 170 +20 2mo ago A SkillSpector: pass 66 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production workflow for agents assembling, auditing, executing, and handing off ComfyUI node-graph workflows for image, video, upscale, inpaint, conditioning, and batch media generation; use when a task involves ComfyUI workflow JSON/API graphs, model and custom-node inventories, reproducibility…

not rated 170 +20 2mo ago A SkillSpector: warn 85 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Production guidance for converting sourced data and approved claims into truthful, accessible, deterministic animated visualizations with D3. Use for data-driven charts, maps, networks, hierarchies, transitions, annotations, responsive video variants, frame-by-frame browser rendering, and visualization QA; not for…

not rated 170 +20 2mo ago A SkillSpector: pass 82 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent FFmpeg finishing workflow for AI agents preparing generated or edited media deliverables. Use when finalizing images, image sequences, video, audio, captions, overlays, social/export variants, checksums, manifests, delivery specs, and QA with ffmpeg/ffprobe, including transcode versus stream-copy…

not rated 170 +20 2mo ago A SkillSpector: pass 106 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Production guidance for authoring, integrating, and reviewing GSAP animation in browser-rendered media. Use for deterministic kinetic typography, SVG drawing and morphing, motion paths, FLIP transitions, responsive motion systems, or frame-seekable GSAP timelines inside HTML video composition and React-based…

not rated 170 +20 2mo ago A SkillSpector: pass 76 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production workflow for AI agents assembling generated or source media into HyperFrames HTML/CSS/JS videos. Use for HyperFrames composition planning, scene architecture, media custody, animation/timing, captions/audio, deterministic preview/render QA, accessibility/flashing checks, provenance…

not rated 170 +20 2mo ago A SkillSpector: warn 67 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Production guidance for assessing, exporting, packaging, integrating, capturing, validating, and handing off Lottie vector animations. Use for Bodymovin/Lottie JSON, dotLottie archives, renderer and player compatibility, fonts and glyphs, image assets, markers and segments, deterministic video capture, responsive…

not rated 170 +20 2mo ago A SkillSpector: pass 89 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production workflow for creating Manim-based explainer animations. Use when an agent must plan, code, render, QA, or hand off precise math, science, data, diagram, algorithm, or concept animations with Manim, including storyboard, scene architecture, formulas, coordinate systems, voiceover timing…

not rated 170 +20 2mo ago A SkillSpector: pass 80 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production guidance for deterministic Canvas 2D and p5.js animation. Use for particles, fields, trails, weather, procedural textures, generative geometry, and lightweight 2D simulations that need fixed media dimensions, seeded repeatability, transparent compositing, aspect variants, performance…

not rated 170 +20 2mo ago A SkillSpector: pass 71 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Provider-independent production workflow for assembling generated or sourced media into React/Remotion videos. Use when an agent must plan, build, render, review, variant-render, or hand off Remotion compositions with media custody, captions, audio, animation, accessibility, provenance, QA, and delivery requirements.

not rated 170 +20 2mo ago A SkillSpector: pass 65 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Production guidance for planning, building, animating, capturing, and reviewing complete Three.js scenes for rendered media. Use for browser-native 3D product shots, title sequences, procedural worlds, glTF scene assembly, camera and lighting animation, shaders, post-processing, deterministic frame export, and…

not rated 170 +20 2mo ago A SkillSpector: pass 85 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Plan, troubleshoot, and hand off provider-independent in-camera VFX and LED volume virtual-production work, with Unreal Engine/nDisplay/Live Link awareness. Use for ICVFX suitability, LED stage geometry, frustums, tracking, lens calibration, sync, render nodes, color, lighting/reflections, artifact mitigation…

not rated 170 +20 2mo ago A SkillSpector: pass 107 tokens original MIT

meshy-3d

66

calesthio/generative-media-skills

Skill Claude CodeCodex

Produce 3D assets with Meshy (meshy.ai) through its REST API and web platform — text-to-3D, image-to-3D, multi-image-to-3D, AI texturing/PBR, remesh/topology control, UV, and auto-rigging/animation. Use this skill when an agent must generate a mesh from a prompt or reference image, control polygon count and topology…

not rated 170 +20 2mo ago A SkillSpector: warn 199 tokens original MIT

tencent-hunyuan3d

67

calesthio/generative-media-skills

Skill Claude CodeCodex

Generate 3D assets with Tencent's Hunyuan3D family — open-weight self-hosted models (Hunyuan3D-2.0/2.1 and HunyuanWorld / HY-World for scenes) and the hosted, closed API tiers (2.5, PolyGen, 3.0/3.1 via Tencent Cloud and third-party hosts). Use when a task involves image-to-3D or text-to-3D mesh generation, the…

not rated 170 +20 2mo ago A SkillSpector: pass 203 tokens original MIT

tripo-3d

68

calesthio/generative-media-skills

Skill Claude CodeCodex

Produce 3D assets with Tripo (tripo3d.ai / Tripo AI by VAST) through its OpenAPI and Studio platform — text-to-3D, image-to-3D, and multiview-to-3D generation; PBR texturing; auto-rigging and preset animation; retopology/low-poly and quad remesh; format conversion (GLB/FBX/OBJ/USDZ/STL/3MF) and engine import…

not rated 170 +20 2mo ago A SkillSpector: warn 274 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Use NVIDIA Maxine / NVIDIA AI for Media audio effects for speech cleanup and enhancement in live or offline media pipelines, including Background Noise Removal, denoise+dereverb, room echo removal, acoustic echo cancellation, audio super-resolution, Studio Voice, Speaker Focus, and Voice Font. Use when selecting…

not rated 170 +20 2mo ago A SkillSpector: pass 109 tokens original MIT

d-id-avatar-video

70

calesthio/generative-media-skills

Skill Claude CodeCodex

Use D-ID to plan, generate, stream, localize, and QA avatar/talking-head videos, including V2 Photo Avatar Talks, V3 Pro/Instant Avatar Clips, V4 Expressive Avatar Scenes, D-ID Agents, voices, consent, lifecycle, safety, and artifact custody.

not rated 170 +20 2mo ago A SkillSpector: warn 64 tokens original MIT

calesthio/generative-media-skills

Skill Claude CodeCodex

Produce Hedra character, talking-avatar, lip-sync, motion-avatar, and live avatar work. Use when an AI agent needs to choose Hedra models, prepare image/audio/script inputs, direct character performance, call or plan around the Hedra API or Studio, evaluate avatar/lip-sync results, handle consent/rights/privacy, or…

not rated 170 +20 2mo ago A SkillSpector: warn 78 tokens original MIT

heygen-avatar-video

72

calesthio/generative-media-skills

Skill Claude CodeCodex

Produce HeyGen avatar videos with Direct Video, Video Agent, Digital Twin, Avatar Realtime, and related avatar/voice/asset APIs. Use for provider routing, avatar and voice selection, script-to-video workflows, lip-sync from audio, backgrounds/scenes, callbacks and polling, consent and rights checks, localization…

not rated 170 +20 2mo ago A SkillSpector: pass 83 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: