CaoMeiYouRen/vision-augment

本地优先的多模态视觉 MCP —— 为无视觉 LLM(DeepSeek、GLM 等)提供可自定义端点的看图 / OCR / 文档解析能力。

0Stars on the repository
2Mods indexed here, across every type
21d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

vision-augment

01

CaoMeiYouRen/vision-augment

MCP server Claude CodeCodexCursor

A local visual tool for AI models that cannot view images, documents, or other visual material directly. It supports configurable image viewing, text extraction from images, and document parsing.

0 21d ago A tokens not measured original MIT