Goodeye-Labs

14 mods across 3 repositories, 13 stars between them.

truesight

02

Goodeye-Labs/truesight-mcp-skills

Plugin Claude Code

MCP server and agent skills for the Truesight AI quality platform. Score inputs, build evaluations, analyze errors, and review results through natural language.

7 5mo ago A tokens not measured original MIT

truesight

03

Goodeye-Labs/truesight-mcp-skills

MCP server Claude CodeCodexCursor +2

MCP server "truesight", hosted remotely at api.truesight.goodeyelabs.com, as configured in Goodeye-Labs/truesight-mcp-skills.

7 5mo ago A tokens not measured original MIT

create-evaluation

06

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Scope what quality should be measured, convert it into one or more actionable binary evaluations, deploy those evaluations through Truesight MCP, and generate a companion skill that applies them correctly. Use when a user wants to create new evals, quality checks, guardrails, or pass/fail criteria for AI outputs.

7 5mo ago A 66 tokens original MIT

error-analysis

07

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Systematically identify and categorize failure modes in evaluated traces using Truesight datasets and error-analysis tools. Use when quality issues are unclear, after major pipeline changes, or when incidents indicate drift.

7 5mo ago A 41 tokens original MIT

eval-audit

08

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Audit an existing evaluation workflow and produce severity-ranked findings with concrete next actions. Use when inheriting an eval setup, diagnosing quality regressions, or checking LLM evaluation process maturity.

7 5mo ago A 40 tokens original MIT

evaluate-trace

09

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Evaluate one or more traces against an existing Truesight live evaluation. Use when a deployed live evaluation already exists and the user wants run outputs with optional handoff to review and promotion.

7 5mo ago A 41 tokens original MIT

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Generate synthetic test data for LLM evaluations using dimension-based tuple expansion. Use when the user needs synthetic traces, test cases, eval datasets, or when create-evaluation needs synthetic fallback data.

7 5mo ago A 43 tokens original MIT

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Judge flagged trace outputs and promote judged items back to datasets. Use when an evaluation run requires human judgment or when review queue items need to be judged for promotion into the dataset.

7 5mo ago A 42 tokens original MIT

truesight-workflows

12

Goodeye-Labs/truesight-mcp-skills

Skill Claude CodeCodex

Orchestrator for Truesight MCP skills. Use this when the user needs help choosing the right Truesight workflow or when intent is ambiguous across LLM evaluate, error analysis, review, templates, or evaluation creation.

7 5mo ago A 51 tokens original MIT

bump-version

13

Goodeye-Labs/goodeye-cli

Skill Claude CodeCodex

Version bump and optional PyPI release for goodeye-cli. Use when bumping the version, cutting a release, or pushing a git tag to trigger PyPI publish.

4 1mo ago A 39 tokens original MIT

goodeye

14

Goodeye-Labs/goodeye-cli

MCP server Claude CodeCodexCursor +2

Design, save, and run outcome-aligned AI workflows and verifiers, with reliable image output. Remote server at mcp.goodeye.dev.

4 1mo ago A tokens not measured original MIT