ai-twinkle/Eval

High-performance LLM evaluation framework with parallel API calls — up to 17× faster than sequential tools. Supports box, math, and logit-based evaluation.

109Stars on the repository
1Mods indexed here, across every type
8d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

Eval CLAUDE.md

01

ai-twinkle/Eval

Instructions file

A project instruction document for Twinkle Eval, a framework for evaluating AI models through API calls. It explains the project’s design rules, structure, supported workflows, and contribution requirements.

109 8d ago A 12,871 tokens original MIT