llm inference mcp servers

5 tagged llm inference, measured the same way as everything else here.

automegakernel

01

RightNow-AI/AutoMegaKernel

MCP server Claude CodeCodexCursor +2

MCP server "automegakernel" as configured in RightNow-AI/AutoMegaKernel. Runs locally from the agent Python package.

135 2mo ago A tokens not measured original MIT

kiagent

02

edjafarov/kiagent-core

MCP server Claude CodeCodexCursor +2

Your mail, chats and documents indexed into a local SQLite corpus, served to AI clients over MCP. Remote server at {subdomain}.localkiagent.com.

7 8d ago A tokens not measured original MIT

inference-aiops

04

AIops-tools/Inference-AIops

MCP server Claude CodeCodexCursor

Governed AI-ops for GPU inference clusters (vLLM + Ray Serve/Jobs): latency/utilization RCA, replica scaling, drain, model lifecycle, and destructive-op guardrails with a built-in governance harness (audit, budget, undo, risk tiers). Runs locally from the inference-aiops Python package.

0 6d ago A tokens not measured original MIT

mcp

05

cheapestinference/mcp

MCP server Claude CodeCodexCursor +2

MCP server "mcp", hosted remotely at api.cheapestinference.com, as configured in cheapestinference/mcp.

0 1mo ago A tokens not measured