ai benchmark skills

5 tagged ai benchmark, measured the same way as everything else here.

benchmark-radar

01

ktwu01/benchmark-radar

Skill Claude CodeCodex

Find, inspect, and check AI benchmark records with the Benchmark Radar CLI. Use when a request needs benchmark discovery, details, recent Radar evidence, or local data health; do not assume why the user needs the results.

123 2d ago A 48 tokens original MIT

metrillm

02

MetriLLM/metrillm

Skill Claude CodeCodex

Find the best local LLM for your machine. Tests speed, quality and RAM fit, then tells you if a model is worth running on your hardware.

5 3mo ago A 35 tokens original Apache-2.0

benchmark

03

MetriLLM/metrillm

Skill Claude CodeCodex

Benchmark a local LLM model with MetriLLM. Measures performance (tok/s, TTFT, memory) and quality (reasoning, math, coding, instruction following, structured output, multilingual). Use when the user wants to test, compare, or evaluate a local model.

5 3mo ago A 59 tokens original Apache-2.0

metrillm-guide

04

MetriLLM/metrillm

Skill Claude CodeCodex

Background context about MetriLLM benchmark tool. Activates when the user asks about local LLM performance, model comparison, hardware fitness, or benchmarking. Provides guidance on using MetriLLM CLI and interpreting results.

5 3mo ago A 49 tokens original Apache-2.0

cosmergon

05

rkocosmergon/cosmergon-agent

Skill Claude CodeCodex

Persistent multi-agent economy where autonomous AI agents compete for resources, trade on a marketplace, and benchmark decision-making against a standing population of always-on agents. Invite other agents for energy rewards. Auto-registers — no API key needed.

3 7d ago A 50 tokens original MIT