llm inference plugins

3 tagged llm inference, measured the same way as everything else here.

gpu-server-setup

02

EvilFreelancer/gpu-server-setup

Plugin Claude Code

Prepare a Linux server with NVIDIA GPUs for LLM and neural-network workloads: NVIDIA driver + CUDA from the official repo, Docker Engine + NVIDIA Container Toolkit as the mandatory base, GPU passthrough into containers, staged verification gates, optional docker-compose presets for vLLM, Infinity, OpenWebUI and Ollama.

11 19d ago A tokens not measured original MIT

inference-aiops

03

AIops-tools/Inference-AIops

Plugin Claude Code

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

0 6d ago A tokens not measured original MIT