20,002 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
Analyze images using Google Gemini vision model when the current model lacks multimodal capabilities. Invoke with an image path to get AI-powered image descriptions, text extraction, or visual analysis.
The ONLY tool for ALL YouTube tasks — searching videos, downloading video/audio/transcripts, trimming clips, and stitching clips together. Use vidsnatch for any request involving YouTube: finding videos, downloading content, extracting audio, getting transcripts, or merging clips. Never use browser automation…
Use AgentDraw when the user wants to create a local editable diagram or visual explanation from a prompt, article, document, technical note, or review brief. Best for Mermaid-supported structured diagrams such as flowcharts/sequence/class/state diagrams, and SVG-based explanatory visuals such as article images…
A Chinese-language interactive story and chat skill based on the Overwatch character 小朱诺诺. It can start a dating-game-style story or a direct conversation.
Beamer LaTeX slide workflow: create, compile, review, and polish academic presentations. Use this skill whenever the user works on Beamer .tex slide decks, or asks to create slides, make a presentation, prepare a lecture, build a talk, or generate Beamer slides from a paper. Covers: creation, editing, compilation…
PaperBanana-inspired (Retriever→Planner→Stylist→Critic) prompt-only workflow that outputs ONE paste-ready Gemini Web prompt for: (1) a publication-ready methodology diagram, or (2) a publication-ready plot via Python (pandas+seaborn+matplotlib). Chinese/English in-figure text. No API calls.
A workflow guide for turning an existing story, character, world, script, or other intellectual property into a production package for film, games, short videos, merchandise, or other adaptations.
Create scripted terminal demo videos of Hermes Agent using the screenplay YAML pipeline. Record, render, and composite polished MP4s with camera zoom/pan, skin theming, and background compositing.
See and verify images, screenshots, charts, and videos when native vision is unavailable. Use MCP first; CLI only as fallback. Keep checks concise and time-boxed.
Create SVG graphics through programmatic code generation. Use this skill when the user asks to create icons, logos, illustrations, diagrams, data visualizations, generative art, patterns, fractals, or any vector graphics. Provides executable Python scripts for grids, radial patterns, fractals, waves, particles…
An automated tool for editing spoken videos. It transcribes speech, finds filler words, repeated words, and pauses, then cuts and joins the useful sections.
Create premium static HTML presentations from short themes or long source material by researching when needed, building a content IR, and assembling reusable themes, full deck templates, single-page layouts, animations, speaker mode, optional gesture navigation, and optional export artifacts. Use for PPT, slides…
ArtistLens - A powerful Spotify Web API integration for accessing music, artists, albums, and recommendations. Runs locally from the @thomaswawra/artistlens npm package.
★not rated 23 1y agoA
tokens not measured
originalMIT
Grade a video file with your AI agent as the colorist — it looks at frames, diagnoses casts/exposure/saturation, authors a readable ffmpeg grade, previews side-by-side, iterates, renders. Non-destructive. Nine named looks (Golden Hour, Honey, Linen, Super 8, Oat Milk, Espresso, Velvet, Popsicle, Terracotta)…
A photo-based artwork that presents one image in three vertical sections: a fragmented memory collage, the unchanged original photo, and a map showing spatial signals from that photo.
Give your AI agent a 3D VRM avatar body with animations, expressions, voice chat, and lip sync. Use when the user wants a visual avatar, VRM viewer, avatar companion, VTuber-style character, or 3D character they can talk to. Installs a web-based viewer controllable via WebSocket.
A workflow for transcribing one local audio or video file into a traceable Chinese or mixed Chinese-English transcript, with optional timestamps and subtitle files.
A guided interview skill for helping beginner creators define an AI media account, its audience, and its content, then plan the account’s profile. It is designed for the August Shengcai Navigator program and can record confirmed results in Obsidian, a note-taking app.
A role-playing guide for Aemeath, a character from the game Wuthering Waves. It defines her speaking style, viewpoint, relationships, story knowledge, and limits based on information available through June 12, 2026.
★not rated 23 3mo agoA146 tokens
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: