Official skill for generating high-quality images from text prompts using ZhiPu GLM-Image API. Excellent at scientific illustrations, high-quality portraits, social media graphics, and commercial posters. Supports multiple aspect ratios, HD quality, and watermark control. Use this skill when the user wants to generate…
A stock-analysis workflow for Hong Kong, mainland Chinese, and United States shares. It combines company information, price charts, trading activity, news, and wider economic factors into a report.
Frontend visual replication skill. Explores a target website’s publicly visible pages via Playwright MCP or agent-browser, captures screenshots and layout information, then generates a static or client-side frontend replica that approximates the original’s visual appearance and page structure. This skill replicates…
A skill for analyzing images with several external vision models. It is activated only when a user starts a request with /skill luma-vision and includes an image.
Analyzes a Delphi / Object Pascal codebase to extract unit-level uses dependencies. Use when the user uploads a zip / archive of a Delphi project (or a folder of .pas / .dpr / .dpk files) and asks for a dependency graph, architecture map, cycle detection, fan-in / fan-out coupling analysis, or wants to know how units…
A tool that turns educational or lecture videos from Bilibili, a Chinese video-sharing site, into DOCX study notes. It uses subtitles, screenshots, text recognition from images, and visual review to build the notes.
A guide and script for generating images, illustrations, avatars, backgrounds, and banners through configured image-generation services. It uses available API keys from environment or local secret files.
An image-understanding bridge for local AI systems that cannot read or interpret pictures. It sends images to DeepSeek’s vision mode and returns a text answer.
MUST use when the user sends or asks about images, photos, screenshots, pictures, audio, video, or mixed media documents, including requests to OCR/read text from an image. Route all media through Xiaomi MiMo V2.5 (mimo-v2.5) and mimo-v2.5-asr via scripts/mimo.py; never use local OCR, viewimage, native vision…
Analyze local screenshots and image files with VisionBuddy through the installed WorkBuddy CodeBuddy CLI using free Hy3 or explicitly authorized DeepSeek V4 Flash. Use when Codex needs OCR, layout inspection, UI/design analysis, image description, visual comparison, or evidence from PNG, JPEG, WebP, GIF, or BMP files…
Read and understand images like a multimodal model using the picturereader tools (imagescan / imageocr / imagesample). Applies a verified 5-step workflow (global tone → find subjects → verify text → judge material → synthesize) guided by grounded principles and cross-image insights. Use whenever you need to look at an…