19,672 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
Understand, generate, and edit footage with Diffusion Studio via the dapi CLI: analyze video/audio/images, generate them with AI, and compose video compositions. Use for any media analysis, media generation, or video editing task.
Nano Banana (nano-banana) image generation skill. Use this skill when the user asks to "generate an image", "generate images", "create an image", "make an image", uses "nano banana", or requests multiple images like "generate 5 images". Generates images using Google's Gemini models (Flash, Pro, or Nano Banana 2) for…
Daily pipeline that picks one long video from a folder, transcribes it with Whisper, uses Gemini 3 Flash multimodal to find every viral short-form moment, cuts each candidate with FFmpeg, adds a hook-text overlay, presents the candidates to the user for approval, and publishes the approved clips to TikTok / Instagram…
A workflow that turns a local short video into a structured Chinese prompt for AI video-generation tools. It examines the visuals, movement, camera work, and audio rhythm, while avoiding invented details when frames are unclear.
A photo-editing style that turns one portrait or pet headshot into a square electric-blue poster with a tightly cropped black-and-white head, coloured stars, and a small barcode detail.
A design workflow for creating or revising cover images for Douyin, WeChat Channels, Xiaohongshu, and other short-video content. It focuses on readable titles, clear layouts, and a recognisable creator identity.
Multi-agent pipeline that builds a polished presentation deck from a single topic. Four agents work in sequence — Strategist defines the narrative, Builder creates the deck, Critic reviews it like a McKinsey EM, Fixer applies the top fixes. Use when you need a presentation that survives senior audiences.
A skill for turning research papers and other scientific material into reusable prompts for AI-generated figures. It can also use reference images when available.
Create long-form generative art with Claude: deterministic hash-seeded rendering, resolution-agnostic output, traits and rarity design, preview capture, determinism verification tools, and platform guides for Art Blocks, 256ART, Verse, Highlight, Plottables, bootloader.art and Artpoint.
Create and edit professional motion graphics videos with Remotion (React-based video). Use this skill EVERY time the user wants to create a video, edit a video, animate something, build an intro/outro/logo animation, make a Reel/Short/promo/launch video, add text animations or captions to footage, composite images and…
Guides users through professional filmmaking workflows in Higgsfield Cinema Studio, including creating multi-shot sequences, configuring optical stacks, applying color grading, managing Soul Cast AI actors, and structuring per-scene prompts with Director Panel camera movements. Use when the user mentions Cinema…
GPT Image 2 CLI as a Claude Code skill: OpenAI gpt-image-2 + Codex imagegeneration under one command surface, with masks, transparent backgrounds, custom sizes up to 4K, and structured JSON / JSONL progress output.
★not rated 136▲
+3 2d agoA
tokens not measured
originalMIT
Create a consistent personalized IP avatar, character sheet, digital double, sticker set, or expression pack from user-owned portrait photos and a curated catalog of preset visual styles. Run a strict Plan-Mode-style, one-question-at-a-time wizard before generating. Use when Codex needs to generate or refine a…
Find, compare, and contribute AI presentation generation, PowerPoint automation, PPTX editing, and slide workflow tools using the awesome-ai-ppt repository. Use this skill when the user asks to choose AI PPT tools, compare HTML-first/image-first/PPTX-native/infrastructure approaches, evaluate whether a GitHub project…
Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, generative art, algorithmic art, flow fields, or particle systems. Create original algorithmic art rather than copying existing artists' work to avoid copyright…
Create polished HTML slide decks and PDF-ready documents for consulting deliverables. Uses the RRBC design system with warm light mode, dark mode cover pages, Lora/Inter/Roboto Mono typography, and data visualization palette. Trigger on 'deck', 'slides', 'presentation', 'pitch deck', 'keynote', 'report', or 'PDF'.
Generate or edit raster images by calling the ChatGPT/Codex hosted imagegeneration flow with local Codex or OpenClaw OAuth credentials, then save decoded image files for OpenClaw and other agent workflows.
Designs and builds video-driven parallax hero landing pages with smooth video scrubbing, progress-locked typography, and clean navigation. Powered by Gemini Omni and Nano Banana.
★not rated 132▲
+1 26d agoA42 tokens
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: