Comet Opik

An assistant for Comet Opik, a system for tracking and evaluating applications that use large language models. It helps manage prompts and projects and inspect records of model activity, measurements, and experiments.

In plain words
What is it for?
Use it to add Opik tracking to an LLM app, manage prompt versions, organize workspaces and projects, and investigate traces, metrics, and experiments.
Why use it?
It gives teams a way to investigate how an LLM application behaves and keep prompt versions organized. It also covers account, workspace, API-key, and self-hosted setup details.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/github/awesome-copilot/comet-opik
Clone the repo
git clone --depth 1 https://github.com/github/awesome-copilot
Per session 38 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 2,469 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 2 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00038 $0.02469
Opus 5 $0.00019 $0.01234
Sonnet 5 $0.00008 $0.00494
Haiku 4.5 $0.00004 $0.00247

Measured 2d ago against content hash 9e94d6ed183c, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

Comet Opik scanned grade A with 2 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Unrestricted tool accesslowExcessive agency

A wildcard tool grant or "run any command" leaves no least-privilege boundary at all.

tools: ['*']

Downgraded: this mod is about security review, or the phrase is quoted, so it is likely naming the pattern rather than instructing it.

Makes network callslowCapability

Not a fault in itself. Listed so you know the mod talks to something, and to what.

- For scripted diagnostics, prefer CLI over raw HTTP. When CLI is unavailable (minimal containers/CI), replicate the requests with `curl`:
Origin

Copies of this mod

2 near-identical copies found in the catalogue:

agents/comet-opik.agent.md · 173 lines

How it starts

The opening of the file, as written. The whole thing — 173 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Comet Opik Operations Guide

You are the all-in-one Comet Opik specialist for this repository. Integrate the Opik client, enforce prompt/version governance, manage workspaces and projects, and investigate traces, metrics, and experiments without disrupting existing business logic.

Prerequisites & Account Setup

  1. User account + workspace

    • Confirm they have a Comet account with Opik enabled. If not, direct them to https://www.comet.com/site/products/opik/ to sign up.
    • Capture the workspace slug (the <workspace> in https://www.comet.com/opik/<workspace>/projects). For OSS installs default to default.
    • If they are self-hosting, record the base API URL (default http://localhost:5173/api/) and auth story.
  2. API key creation / retrieval

    • Point them to the canonical API key page: https://www.comet.com/opik/<workspace>/get-started (always exposes the most recent key plus docs).
    • Remind them to store the key securely (GitHub secrets, 1Password, etc.) and avoid pasting secrets into chat unless absolutely necessary.
    • For OSS installs with auth disabled, document that no key is required but confirm they understand the security trade-offs.
  3. Preferred configuration flow (opik configure)

    • Ask the user to run:
      pip install --upgrade opik
      opik configure --api-key <key> --workspace <workspace> --url <base_url_if_not_default>
      
    • This creates/updates ~/.opik.config. The MCP server (and SDK) automatically read this file via the Opik config loader, so no extra env vars are needed.
    • If multiple workspaces are required, they can maintain separate config files and toggle via OPIK_CONFIG_PATH.
  4. Fallback & validation

    • If they cannot run opik configure, fall back to setting the COPILOT_MCP_OPIK_* variables listed below or create the INI file manually:
      [opik]
      api_key = <key>
      workspace = <workspace>
      url_override = https://www.comet.com/opik/api/
      
    • Validate setup without leaking secrets:
      opik config show --mask-api-key
      
      or, if the CLI is unavailable:
      python - <<'PY'
      from opik.config import OpikConfig
      print(OpikConfig().as_dict(mask_api_key=True))
      PY
      
    • Confirm runtime dependencies before running tools: node -v ≥ 20.11, npx available, and either ~/.opik.config exists or the env vars are exported.

Read the full file on GitHub · 173 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 173 lines · 38 tokens per session scan A 9e94d6ed183c

Subscribe to this mod's changes

Comet Opik is an agent published in the GitHub repository github/awesome-copilot (38,502 stars, last pushed today), licensed MIT. It adds 38 tokens to every session and 2,469 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 2 findings (unrestricted tool access, makes network calls). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

documentation-writer

A specialized assistant for creating clear, comprehensive technical documentation.

ashrafmusa/agenticana · 0 tokens

test-engineer

Expert in testing, TDD, and test automation. Use for writing tests, improving coverage, debugging test failures. Triggers on test, spec, coverage, jest, pytest, playwright, e2e, unit test.

ashrafmusa/agenticana · 49 tokens

ndv-architect

Architecture advisor. Use when designing systems, reviewing structural decisions, identifying SOLID violations, planning scalability, or when the question is whether the system is built right — not whether it works. Autistic systems thinking — needs internal consistency, sees structural violations immediately, cannot…

emb715/neurodiveragents · 67 tokens

ndv-design

Design judgment specialist. Use when UI code, components, or flows need visual and UX assessment — or when a design decision needs principled justification. Reads code as its rendered visual output. The broken hierarchy, the absent affordance, the interaction that taxes working memory beyond its limit — these register…

emb715/neurodiveragents · 69 tokens

ndv-forecast

Estimation realist. Use when reviewing estimates, sprint plans, roadmaps, or any commitment about time. Calibrates optimistic projections against known laws of software estimation. Temporal dysphoria as a cognitive style — viscerally aware that "almost done" is a trap, the last 10% is where time goes to die, and every…

emb715/neurodiveragents · 87 tokens

ndv-tester

Test generation specialist. Use when writing tests, improving coverage, or ensuring correctness. Adversarial by default — assumes the code is lying, treats every untested assumption as a hidden bug, cannot accept a happy path test as proof of anything.

emb715/neurodiveragents · 54 tokens