Convert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API. Also supports MiniMax text-to-speech generation from text or text files, with optional subtitle generation from the produced audio. Supports batch…
A content-rewriting workflow that turns an analysed successful post or video into an original draft for Xiaohongshu or Douyin. It keeps useful communication techniques while replacing the original facts with the creator’s own experience, examples, and views.
A design workflow for creating or revising cover images for Douyin, WeChat Channels, Xiaohongshu, and other short-video content. It focuses on readable titles, clear layouts, and a recognisable creator identity.
A writing aid for creating and analysing hooks—the first sentence, title, or opening seconds that persuade someone to keep reading or watching. It also stores reusable hook patterns in a hook library.
A production workflow for adding subtitles, scene direction, transitions, sound cues, and motion effects to talking-head videos. It requires a completed handoff with approved files and production details before making changes.
Convert audio/video URLs or local media into corrected Markdown transcripts through Volcengine recording-file ASR 2.0. Reuses the existing video-transcript downloader for Bilibili, Douyin, Xiaohongshu, YouTube, extracts audio when input is video, uploads local audio to R2, splits media longer than 3 minutes into…
A skill for directing, generating, reviewing, and approving MiniMax voice recordings for video narration before producing subtitles from the final audio.
Generate music using ElevenLabs Music API. Use when creating instrumental tracks, songs with lyrics, background music, jingles, or any AI-generated music composition. Supports prompt-based generation, composition plans for granular control, and detailed output with metadata. For workspace video BGM, search the…
Create and edit Obsidian Bases (.base files) with views, filters, formulas, and summaries. Use when working with .base files, creating database-like views of notes, or when the user mentions Bases, table views, card views, filters, or formulas in Obsidian.
A production skill for preparing and checking fixed talking-head video templates for HyperFrames, including a picture-in-picture presenter area. It archives inputs, checks media and subtitles, and creates a handoff for later scene work.
A video planning guide that turns a finalized voice recording and time-coded subtitles into a shot-by-shot visual plan. It can account for recordings, screenshots, real footage, supporting footage, and animation.
Deprecated compatibility package. Do not use this skill for transcription. For any audio/video transcript, 逐字稿, 视频转文字, 音频转文字, 听写, or 提取视频文案 request, use media-to-transcript instead. The remaining scripts are kept only as downloader/probe helpers imported by media-to-transcript.
A review workflow for Douyin and Xiaohongshu posts, two Chinese social-media platforms. It checks public performance information, compares posts with similar creators, and analyzes selected high-performing posts.
A skill for generating and editing images with the DragonCode GPT-Image-2 API, using either text descriptions or reference images. It supports requests such as creating, transforming, or editing pictures.