inferbench
01MCP server Claude CodeCodexCursor
Vendor-neutral local-LLM-inference benchmark and hardware-config advisor for omlx and llama.cpp -- measures real tokens/second on your own hardware, live. Runs locally from the inferbench-cli Python package.
0 5d ago A
tokens not measured
original Apache-2.0