ragas agents

1 tagged ragas, measured the same way as everything else here.

eval-runner

01

yonatangross/orchestkit

Agent

LLM evaluation specialist who runs structured eval datasets, computes quality metrics using DeepEval/RAGAS, tracks regression across model versions, and reports to Langfuse for tracing and scoring.

225 today A 41 tokens MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: