Command
A command that asks a vision tool to analyse a local image using a natural-language question.
一个 MCP(Model Context Protocol)服务器,基于任意兼容 Anthropic Messages 协议的识图模型,提供 7 个视觉工具——通用图像分析、UI 转代码、OCR、错误诊断、图表理解、数据可视化分析、UI 对比。 兼容 Claude Code 及任何 MCP 客户端。你只需指向自己的模型 API(地址 + Key + 模型 ID),即可获得一套开箱即用的视觉能力。
Command
A command that asks a vision tool to analyse a local image using a natural-language question.
Command
A command for analysing charts and other data visualisations, including trends, unusual values, and business-relevant findings.
Command
A command for interpreting technical diagrams such as architecture drawings, flowcharts, UML diagrams, and database relationship diagrams.
Command
A command for comparing two user-interface screenshots and listing their visual differences by severity.
Command
A command for diagnosing error screenshots with a vision MCP server. It examines stack traces, console output, or error dialogs in an image and suggests fixes.
Command
A command that extracts all visible text from a screenshot while preserving its original layout and structure. OCR, or optical character recognition, is the process of reading text in images.
Command
A command that turns a user-interface screenshot into code or a written design description. A user interface is the visible part of an application that people interact with.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: