genmedia-labs

30 mods across 1 repository, 3 stars between them.

ace-step

01

genmedia-labs/skills

Skill Claude CodeCodex

Generate, inpaint, and outpaint music with ACE Step on RunComfy via the runcomfy CLI. ACE Step is StepFun-AI's open-weights music foundation model — tag-driven composition (genre, mood, instruments), multilingual lyrics with section markers, 5 s to 4 min stereo output, $0.0002–0.0003 per second (≈ 27× cheaper than…

3 20d ago A 230 tokens copy · 100% MIT

ai-avatar-video

02

genmedia-labs/skills

Skill Claude CodeCodex

Create AI avatar, talking-head, and lip-sync videos on RunComfy via the runcomfy CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar), Wan-AI Wan 2-7 (audio-driven mouth sync via audiourl on a portrait), HappyHorse 1.0 (Arena #1 t2v / i2v with in-pass audio), and Seedance v2 Pro (multi-modal…

3 20d ago A 238 tokens copy · 100% MIT

ai-image-generation

03

genmedia-labs/skills

Skill Claude CodeCodex

Generate and edit images on RunComfy via the runcomfy CLI — a smart router across the full image-model catalog: FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance Seedream 5 / 4-5 / 4-0 and Dreamina 4-0, Alibaba Qwen Image and Z-Image Turbo, Wan 2-7. Covers…

3 20d ago A 237 tokens copy · 100% MIT

ai-music

04

genmedia-labs/skills

Skill Claude CodeCodex

Generate AI music on RunComfy via the runcomfy CLI — a smart router across the music-model catalog. Routes to ElevenLabs AI Music Generation (premium 44.1 kHz stereo vocal tracks, 5 s–5 min, $0.0083/s) and ACE Step / ACE Step 1.5 (StepFun-AI open-weights, tag-driven composition, multilingual lyrics, $0.0002–0.0003/s…

3 20d ago A 272 tokens copy · 100% MIT

ai-video-generation

05

genmedia-labs/skills

Skill Claude CodeCodex

Generate AI videos on RunComfy via the runcomfy CLI — a smart router across the full video-model catalog: HappyHorse 1.0 (Arena #1, native in-pass audio), Wan-AI Wan 2-7 (open weights, audio-driven lip-sync), ByteDance Seedance v2 / 1-5 / 1-0 (multi-modal cinematic), Kling 3.0 / 2-6, Google Veo 3-1, MiniMax Hailuo…

3 20d ago A 276 tokens copy · 100% MIT

codex-pet

06

genmedia-labs/skills

Skill Claude CodeCodex

Codex Pet generator on RunComfy. Build a Codex-compatible Codex Pet spritesheet.webp + pet.json from a single reference image, drop it into ${CODEXHOME:-$HOME/.codex}/pets/ / and Codex picks it up as a custom Codex Pet next to the 8 built-ins. This skill produces the exact Codex Pet atlas Codex expects (1536x1872…

3 20d ago A 279 tokens copy · 100% MIT

controlnet-pose

07

genmedia-labs/skills

Skill Claude CodeCodex

Pose-conditioned generation on RunComfy via the runcomfy CLI. Routes across Kling 2-6 Motion Control Pro / Standard (transfer the motion / blocking of a reference video onto a target character), community Wan 2-2 Animate (audio-driven character animation with pose conditioning), and Z-Image Turbo ControlNet LoRA…

3 20d ago A 186 tokens copy · 100% MIT

genmedia-labs/skills

Skill Claude CodeCodex

Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the runcomfy CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals…

3 20d ago A 198 tokens copy · 100% MIT

face-swap

09

genmedia-labs/skills

Skill Claude CodeCodex

Swap a face / character into video or images on RunComfy via the runcomfy CLI. Routes across community Wan 2-2 Animate (audio-driven character animation + identity swap), GPT Image 2 Edit (single-shot precise face swap on still images via reference composition), Nano Banana Edit (batch identity-preserving swap), Flux…

3 20d ago A 210 tokens copy · 100% MIT

flux-2-klein

10

genmedia-labs/skills

Skill Claude CodeCodex

Generate images with Flux 2 Klein (Black Forest Labs' distilled fast variant of Flux 2) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents Flux 2 Klein's strengths (sub-second latency, multi-reference brand…

3 20d ago A 200 tokens copy · 100% MIT

flux-kontext

11

genmedia-labs/skills

Skill Claude CodeCodex

Edit images with Flux 1 Kontext Pro (Black Forest Labs' precise local image-edit model) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents Flux Kontext's strengths (single-reference precise local edits, strong…

3 20d ago A 167 tokens copy · 100% MIT

gpt-image-2

12

genmedia-labs/skills

Skill Claude CodeCodex

Generate and edit images with OpenAI GPT Image 2 (ChatGPT Images 2.0) on RunComfy. Documents GPT Image 2's strengths (embedded text, logos, multilingual typography, instruction precision), its 3 fixed sizes, edit-with-preservation language, and when to route to a sibling (Flux 2 / Nano Banana Pro / Seedream) instead.…

3 20d ago A 153 tokens copy · 100% MIT

gpt-image-edit

13

genmedia-labs/skills

Skill Claude CodeCodex

Edit images with OpenAI GPT Image 2 (the /edit endpoint of ChatGPT Images 2.0) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents GPT Image Edit's strengths (preservation language, multilingual in-image text…

3 20d ago A 174 tokens copy · 100% MIT

happyhorse-1-0

14

genmedia-labs/skills

Skill Claude CodeCodex

Generate text-to-video with HappyHorse 1.0 on RunComfy. Documents HappyHorse 1.0's strengths (#1 on Artificial Analysis Video Arena, native 1080p with in-pass synchronized audio, multi-shot character consistency, 6-language prompt support), the duration / aspect-ratio / resolution schema, and when to route to Wan 2.7…

3 20d ago A 156 tokens copy · 100% MIT

image-edit

15

genmedia-labs/skills

Skill Claude CodeCodex

Edit images on RunComfy — this skill is a smart router that matches the user's intent to the right edit model in the RunComfy catalog. Picks Nano Banana Edit (batch up to 20, identity-preserving default), OpenAI GPT Image 2 Edit (multilingual in-image text rewrite, multi-ref composition, layout precision), Flux…

3 20d ago A 186 tokens copy · 100% MIT

image-inpainting

16

genmedia-labs/skills

Skill Claude CodeCodex

Mask-driven image inpainting on RunComfy via the runcomfy CLI. Routes to Tongyi MAI Z-Image Turbo Inpainting (the dedicated inpainting endpoint with mask, strength, and control-scale) and to identity-preserving edit models (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when a mask isn't available and the…

3 20d ago A 178 tokens copy · 100% MIT

image-outpainting

17

genmedia-labs/skills

Skill Claude CodeCodex

Image outpainting on RunComfy via the runcomfy CLI — extend a still beyond its original canvas, fill in what the camera didn't capture, change aspect ratio (square → 16:9, portrait → landscape) while preserving the original content. Routes across Nano Banana 2 Edit (default, spatial-language driven), GPT Image 2 Edit…

3 20d ago A 196 tokens copy · 100% MIT

image-to-video

18

genmedia-labs/skills

Skill Claude CodeCodex

Animate any still image on RunComfy — this skill is a smart router that matches the user's intent to the right i2v model in the RunComfy catalog. Picks HappyHorse 1.0 I2V (Arena #1, native audio, identity preservation) for general animations, Wan 2.7 with audiourl for custom-voiceover lip-sync, or Seedance 2.0 Pro for…

3 20d ago A 190 tokens copy · 100% MIT

kling-3-0

19

genmedia-labs/skills

Skill Claude CodeCodex

Kling 3.0 video generation on RunComfy. Kling 3.0 (also called Kling V3.0) is Kuaishou Technology's third-generation multi-shot video model with native synchronized audio and consistent character identity across shots. This skill covers all six Kling 3.0 endpoints, spanning three rendering tiers (Standard, Pro, 4K)…

3 20d ago A 182 tokens copy · 100% MIT

lipsync

20

genmedia-labs/skills

Skill Claude CodeCodex

Lip-sync a face to a specific audio track on RunComfy via the runcomfy CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar from a portrait + audio), Sync Labs sync v2 / Pro (state-of-the-art mouth sync onto a video), Kling lipsync (audio-to- video and text-to-video with synced speech), and Creatify…

3 20d ago A 182 tokens copy · 100% MIT

nano-banana-2

21

genmedia-labs/skills

Skill Claude CodeCodex

Generate images with Google Nano Banana 2 (Gemini-family flash-tier text-to-image) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents Nano Banana 2's strengths (rapid iteration, in-image typography rendering…

3 20d ago A 174 tokens copy · 100% MIT

nano-banana-edit

22

genmedia-labs/skills

Skill Claude CodeCodex

Edit images with Google Nano Banana 2 (image-to-image edit endpoint) on RunComfy. Documents Nano Banana Edit's strengths (preserve subject identity, swap background, localize edits with spatial language, multi-image batch edits up to 20 inputs), the schema, and when to route to GPT Image 2 edit / Flux Kontext / Nano…

3 20d ago A 136 tokens copy · 100% MIT

relight

23

genmedia-labs/skills

Skill Claude CodeCodex

Relight a still image — change the lighting setup, color temperature, direction, or mood — on RunComfy via the runcomfy CLI. Routes to Qwen Edit 2509's dedicated relight LoRA endpoint for purpose-built relighting, with fallback to identity-preserving edit endpoints (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext…

3 20d ago A 179 tokens copy · 100% MIT

runcomfy-cli

24

genmedia-labs/skills

Skill Claude CodeCodex

Run any model on RunComfy from the command line. The runcomfy CLI is one binary, one auth, hundreds of model endpoints — image generation, image edit, video generation, image-to-video, lip-sync, face swap, video edit, inpainting, outpainting, extend, ControlNet, relight, upscale, LoRA training and more. Submit a…

3 20d ago A 240 tokens copy · 100% MIT