AIops-tools/Inference-AIops

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 30 MCP tools (preview)

0Stars on the repository
3Mods indexed here, across every type
6d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

inference-aiops

01

AIops-tools/Inference-AIops

MCP server Claude CodeCodexCursor

Governed AI-ops for GPU inference clusters (vLLM + Ray Serve/Jobs): latency/utilization RCA, replica scaling, drain, model lifecycle, and destructive-op guardrails with a built-in governance harness (audit, budget, undo, risk tiers). Runs locally from the inference-aiops Python package.

0 6d ago A tokens not measured original MIT