bootstrap-template-evaluation

bootstrap-template-evaluation is a skill for Claude Code from Goodeye-Labs/truesight-mcp-skills. It costs 37 tokens per session (483 once invoked), scanned A, original, MIT.

A guided workflow for launching a code or content evaluation from a pre-built Truesight template. It finds a suitable template, creates a private dataset, deploys the evaluation, and checks it with sample inputs.

In plain words
What is it for?
Use it to select a template, create its evaluation dataset, deploy a live evaluation, save the returned API key, and run a basic verification.
Why use it?
It avoids building evaluation settings from the beginning when an existing template already fits the task.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin. Also seen: names the AskUserQuestion tool.

Part of the truesight plugin — 9 skills, 1 MCP server shipped together

Good fit Use it to select a template, create its evaluation dataset, deploy a live evaluation, save the returned API key, and run a basic verification.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add Goodeye-Labs/truesight-mcp-skills --skill bootstrap-template-evaluation
Clone the repo
git clone --depth 1 https://github.com/Goodeye-Labs/truesight-mcp-skills

Made for: Claude Code.

Or install truesight, the plugin that ships this one along with the rest of its 9 skills, 1 MCP server.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for bootstrap-template-evaluation

README.md
[![agentmods](https://agentmods.dev/badge/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation/github.svg)](https://agentmods.dev/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation)
Your own site
<a href="https://agentmods.dev/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation"><img src="https://agentmods.dev/badge/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for bootstrap-template-evaluation

Your own site · 80×15
<a href="https://agentmods.dev/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation"><img src="https://agentmods.dev/badge/skills/goodeye-labs/truesight-mcp-skills/bootstrap-template-evaluation.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 37 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 483 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00037 $0.00483
Opus 5 $0.00018 $0.00242
Sonnet 5 $0.00007 $0.00097
Haiku 4.5 $0.00004 $0.00048

Measured 11d ago against content hash e896af7639c3, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-11, from the pricing page.

Security

Grade A, and why

bootstrap-template-evaluation scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/bootstrap-template-evaluation/SKILL.md · 61 lines

What it actually says

Bootstrap Template Evaluation

Use this skill when a pre-built template likely covers the target use case.

Interactive Q&A protocol (mandatory)

If template choice is ambiguous, ask one question at a time using the structured question tool (loaded per the HARD-GATE above).

Example question structure:

Which template family best matches your goal?
A) AI writing detection
B) Code quality
C) Unsure, list all templates first

Rules:

  • Ask one question per message.
  • Use the structured question tool for every question. Structure each with a short header, 2-4 options with labels and descriptions, and place the recommended option first. Do not add "(Recommended)" or similar annotations to option labels.
  • Ask one follow-up only when needed.

Workflow

  1. Discover templates:
    • Call list_templates.
  2. Select template:
    • Match use case to template slug.
  3. Provision private dataset:
    • Call provision_template(slug).
  4. Deploy live evaluation:
    • Call create_and_deploy_evaluation(dataset_id).
    • Capture api_key immediately because it is returned only once.
  5. Verify:
    • Run run_eval with representative inputs.
  6. Return deployment artifacts:
    • dataset_id
    • live_evaluation_id
    • verification result

Guardrails

  • If no template fits, hand off to create-evaluation.
  • Do not skip verification after deployment.

Scopes reference

  • list_templates requires datasets:read
  • provision_template requires datasets:write
  • create_and_deploy_evaluation requires evaluations:write, live-evaluations:write
  • run_eval requires live-evaluations:execute
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 11d ago First seen · 61 lines · 37 tokens per session scan A e896af7639c3

Subscribe to this mod's changes

bootstrap-template-evaluation is a skill published in the GitHub repository Goodeye-Labs/truesight-mcp-skills (7 stars, last pushed 5mo ago), licensed MIT. It adds 37 tokens to every session and 483 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.