Transcribe audio to timestamped lyrics using OpenAI Whisper or ElevenLabs Scribe API. Outputs LRC, SRT, or JSON with word-level timestamps. Use when users want to transcribe songs, generate LRC files, or extract lyrics with timestamps from audio.
Music songwriting guide for ACE-Step. Provides professional knowledge on writing captions, lyrics, choosing BPM/key/duration, and structuring songs. Use this skill when users want to create, write, or plan a song before generating it with ACE-Step.
Generate song cover/thumbnail images using Gemini API. Creates artistic images suitable for music video backgrounds. Use when users want to generate album art, song covers, thumbnails, or background images for MVs.
Instructions for timoncool/ACE-Step-Studio, covering ace-step studio — agent guidelines, project overview, architecture, three-process architecture (single terminal) and model switching.
A workflow for turning ideas into prompts for AI-generated images, videos, and music across several media platforms. When needed, it can also send those prompts to the chosen platform through browser automation.
AGENTS.md instructions for calesthio/Resonant: When a user asks to compose, arrange, mix, analyze, or render music, use the resonant MCP tools. Do not edit .resonant JSON directly.
Claude Code instructions for calesthio/Resonant, a project described as: Free, local AI music studio for Windows—generate songs, play instruments, arrange, mix, export WAV, and connect Codex or Claude through MCP.
Instructs agents to control Stratawright DAW via the daw-cli command-line IPC interface. Covers session state, transport, track management, gain staging, VST3/AU plugin hosting, signal routing, clips, MIDI timeline editing, and non-visual DSP analysis & audio intelligence. Requires the DAW application to be running.
AI music production suite using ACE-Step 1.5 via direct Python API. Song generation, cover/style transfer, section editing (repaint), track extraction, multi-track layering, audio completion, songwriting, analysis, platform export, enhancement, and LoRA training. 50+ languages, 10-minute compositions, 48kHz stereo.
Creates cover versions and style transfers of existing songs using ACE-Step 1.5. Takes a reference audio file and generates a new version with different style, genre, or vocal characteristics while preserving musical structure.
Generates music from text descriptions and lyrics using ACE-Step 1.5's Python API. Creates full songs with vocals, instrumentals, or both. Supports 50+ languages, 10-600 second duration, quality presets (draft/standard/high/max), and batch generation.
Music production suite using ACE-Step 1.5 via Python API. Routes /music commands for generation, cover, repaint, compose, analyze, export, enhance, random, and LoRA training. 50+ languages, up to 10-minute tracks, 48kHz stereo.
Composition specialist for ACE-Step 1.5. Turns a rough idea ("a sad song about rain") into a fully-specified generation plan: caption, lyrics with structure tags, BPM, key, duration, and recommended ACE-Step task type + quality preset. Researches genre conventions in references/ and picks values with reasons.
A workflow for laying out promotional posters from finished key images, character reference images, or cover images. It targets platforms such as Bilibili, Xiaohongshu, Kuaishou, and WeChat Moments, while keeping the main character’s face and visual identity unchanged.
A workflow for turning vague or conversational AI image and video requests into prompts suited to particular creative tools. It covers image tools such as Midjourney and DALL-E and video tools such as Sora, Runway, and Kling.
A workflow for generating 16:9 horizontal cover images for sponsored or educational videos with GPT Image 2. It uses a person reference image and a video script to develop title ideas, layouts, and image prompts before producing several covers.