KuaaMU/mcp-vision-bridge

MCP server that gives text-only LLM coding agents vision — analyze images via any multimodal model (mimo, Claude, Gemini, OpenAI-compatible). Works with Claude Code, Codex, Kimi, opencode, PI.

12Stars on the repository
4Mods indexed here, across every type
24d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

mcp-vision-bridge

01

KuaaMU/mcp-vision-bridge

Plugin Claude Code

Vision for text-only coding agents: analyzeimage MCP tool + vision skill + auto-loop hook (captures pasted images from the session transcript). Bring your own multimodal model (mimo, Claude, Gemini, GPT-4o).

12 24d ago A tokens not measured original MIT

mcp-vision-bridge

02

KuaaMU/mcp-vision-bridge

Plugin Claude Code

Give your text-only coding agent (DeepSeek V4 Flash, Qwen, Kimi) vision. Adds an analyzeimage MCP tool that routes images through any multimodal model (mimo, Claude, Gemini, GPT-4o) and returns a detailed text description. Paste or drag in one or many images — analyzed in a single call, scoped to the current session.

12 24d ago A tokens not measured original MIT