groq/openbench

Provider-agnostic, open-source evaluation infrastructure for language models

813Stars on the repository
3Mods indexed here, across every type
6d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

lighteval-porter

01

groq/openbench

Agent Claude Code

Use this agent when you need to port an evaluation benchmark from the LightEval framework to openbench. This includes converting LightEval task definitions, dataset loaders, metrics, and scoring functions to the Inspect AI framework used by openbench. The agent should be invoked when the user mentions porting…

813 6d ago A 305 tokens original MIT