RudrenduPaul/InferBench

Benchmark local LLM inference engines (llama.cpp, omlx) on your own hardware with real tokens/sec, not borrowed numbers

0Stars on the repository
1Mods indexed here, across every type
5d agoLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

inferbench

01

RudrenduPaul/InferBench

MCP server Claude CodeCodexCursor

Vendor-neutral local-LLM-inference benchmark and hardware-config advisor for omlx and llama.cpp -- measures real tokens/second on your own hardware, live. Runs locally from the inferbench-cli Python package.

0 5d ago A tokens not measured original Apache-2.0