CaoMeiYouRen/vision-augment

本地优先的多模态视觉 MCP —— 为无视觉 LLM(DeepSeek、GLM 等)提供可自定义端点的看图 / OCR / 文档解析能力。

0Stars on the repository
2Mods indexed here, across every type
22d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

vision-augment

01

CaoMeiYouRen/vision-augment

Skill Claude CodeCodex

A visual and document-analysis service for agents that cannot directly understand images. It can inspect pictures, read text with OCR, and parse documents such as PDF, Word, PowerPoint, spreadsheet, and HTML files.

not rated 0 22d ago A 104 tokens original MIT