Plugin Claude Code
DeepEval plugins for LLM evaluation, tracing, and testing in Claude Code.
17 tagged llm evaluation, measured the same way as everything else here.
Browse within: python 5
Plugin Claude Code
DeepEval plugins for LLM evaluation, tracing, and testing in Claude Code.
Plugin Claude Code
Skills for adding DeepEval evaluations, tracing, datasets, Confident AI reports, and iterative improvement loops to AI applications.
Plugin Claude Code
Plugin marketplace listing 1 plugin: nuguard.
Plugin Claude Code
AI application security for Claude — generate an AI Bill of Materials (AI-SBOM), run static analysis, behavioral validation, and adversarial red-team testing for AI agents and LLM-powered applications.
Plugin Claude Code
AI Application Security — SBOM generation, static analysis, behavioral testing, and adversarial red-teaming for AI agents and LLM-powered applications.
Plugin Claude Code
Claude Code distribution for the deslop Agent Skill.
Plugin Claude Code
Deletion-first audit and cleanup for accumulated test, verification, and fallback bloat.
stefanobaghino/simple-output-styles
Plugin Claude Code
Output styles that make Claude write clearly, for everyone.
Plugin Claude Code
mcpbr - MCP Benchmark Runner plugin marketplace.
Plugin Claude Code
Expert benchmark runner for MCP servers using mcpbr. Handles Docker checks, config generation, and result parsing.
Plugin Claude Code
Plugin marketplace listing 1 plugin: tunelab.
Plugin Claude Code
Cut your AI bill without losing accuracy. tunelab helps you move repetitive LLM work — classifying, routing, extracting, drafting — onto small models that run for free on your Mac. It decides by experiment (testing on your own data first), trains locally with MLX, evaluates honestly, and explains every step so you…
vassiliylakhonin/agenda-intelligence-md
Plugin Claude Code
Plugin marketplace listing 4 plugins: agenda-intelligence, global-think-tank-analyst, central-asia-caspian, gulf-middle-east.
vassiliylakhonin/agenda-intelligence-md
Plugin Claude Code
Deterministic evidence-packet linter for claim-backed AI output, with compatibility agenda-analysis skills and MCP tools. Reports packet completeness, not factual truth.
Plugin Claude Code
Legacy advisory Stop-hook bundle for the fixed English gate. Current Claude Stop compatibility is not certified in v0.1.7; use the CLI directly. Requires the hermeneutic package.
Plugin Claude Code
AI evaluation strategy design assistant using Evaluation-Driven Development (EDD).
Plugin Claude Code
Run synthetic focus groups and user research panels using AI personas. Define personas in YAML, design survey instruments, and collect structured qualitative feedback — all from Claude Code.