Use when one or more photographs must become adaptive photo-plus-abstraction editorial compositions while source facts, spatial relationships, and strict photo preservation remain controlled.
Create beautiful, performant SVG animations and illustrations. Use this skill when the user asks to create SVG graphics, icons, illustrations, animated logos, path animations, morphing shapes, loading spinners, or any animated SVG content. Covers SMIL animations, CSS-driven SVG animation, path drawing effects, shape…
Produce vertical, square, and platform-specific versions of an Adobe Premiere Pro sequence while preserving the master edit. Use for 9:16, 1:1, and social cutdowns, Auto Reframe workflows, safe-zone checks, title repositioning, or multi-platform exports through the Premiere Pro MCP server.
Generate interior design planning artifacts, 2D SVG floorplans, BOM outputs, analysis boards, and renovation image sets from structured requirements. Use for staged interior-design work that moves through plan, 2D drafting, BOM, and image outputs.
Manipulate images locally using Python and PIL/Pillow. Use when the user asks to resize, crop, rotate, flip, filter, enhance, combine, overlay, watermark, add text to, convert, compress, create, or edit images locally. Also use for thumbnails, borders, color adjustments, transparency, animated GIFs, or extracting…
A text-to-speech tool that uses Xiaomi MiMo V2.5 models to turn written text into spoken audio. It supports preset voices, designed voices, cloned voices, and controls for style, emotion, dialect, and singing.
Run and validate AutoVideo-Agent Markdown-to-video workflows, including the v0.1-compatible deterministic offline command and v0.2 provider mode with ComfyUI media, Mock or Command TTS, scene-level SRT, normalized timelines, FFmpeg, and deterministic QA. Use for build, plan, provider, or QA requests; MiniMax, hosted…
A skill for turning source documents such as PDFs, Word files, web pages, or Markdown into designed presentation pages and exporting them as PowerPoint files. It uses a staged workflow with planning, generation, preview, and quality checks.
A guide for creating background music plans for product-promo videos. It reads materials such as a storyboard or EDL (an edit decision list containing shot timings) and produces music direction, timing points, and generation prompts.
Generates a consistent illustrated image for every scene of a story — keeping characters, locations, and visual style coherent across the whole sequence using cascading reference images. Reads a {slug}scenes.json (from scene-splitter), proposes a visual style, ASKS the user for the aspect ratio and image model…
Plan and orchestrate end-to-end video production pipelines in ComfyUI with validation gates and error recovery. Handles img2vid, txt2vid, vid2vid, and multi-shot video production. Produces pipeline plans with correct step ordering (generate, validate, animate, validate, concat), model selection, retry strategies (seed…
A skill that turns a supplied photograph into a minimalist printed poster or zine-style artwork using hand-written HTML, CSS, and SVG, then exports a PNG. It does not use an image-generation model.
Generate clean, playful low-density isometric voxel icon illustrations and explicitly requested four-frame physically credible loops built from large opaque face-lit acrylic blocks on an exact.
Generate and edit images using Wan and Qwen Image models. Supports text-to-image, image editing (style transfer, subject consistency, text rendering), and interleaved text-image output. TRIGGER when: user wants to create illustrations, product images, artistic designs, posters, text-to-image generation, edit/transform…
A tool for managing podcast subscriptions, synchronizing RSS episode details, transcribing audio, creating chapters, searching transcripts, and saving notes as Markdown or Obsidian files. RSS is a standard feed that publishes new episode information.
Generate AI-illustrated comic-style educational guides from documentation, source code, or any technical content. Produces actual comic PNG images using AI image generation (not HTML). Supports 10+ anime styles: Doraemon, Naruto, One Piece, Dragon Ball, Spy x Family, Chibi, Chinese ink, Ghibli, Crayon Shin-chan…
Transcribe audio (voice notes, recordings, meeting audio) to text via NetMind's Whisper model. Use when the user shares an audio file or URL and your model cannot hear audio. Zero config for NetMind-powered users — the API key is injected automatically.
Use when generating, editing, composing, or iterating on images — illustrations for reports/web, posters with Chinese or English typography, pitch-deck slides, UI mockups, infographics, pixel art, game sprites, character reference sheets, app icons, logo concepts, photoreal product shots, or photo edits with precise…
Create a consistent oil-style visual system in two modes: finished explanatory images with short accurate labels generated directly inside the scene, and transparent character illustrations produced with a bundled background-removal script. Use for concepts, mechanisms, comparisons, workflows, tradeoffs, hero artwork…
★not rated 86▲
+4 1mo agoA77 tokens
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: