enrich-data

A command for adding metadata that explains a project's data, including event descriptions, property descriptions, and tags. It passes the work to a separate metadata-management skill.

In plain words
What is it for?
Use it to enrich a Mixpanel project's data dictionary, called a Lexicon. It requires a project in scope and the permission to write Lexicon metadata.
Why use it?
It helps an agent understand what the data means before using it for analysis. It also records the resulting readiness score for the session.

Command

Part of the mixpanel plugin — 12 skills, 23 commands shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add commands/mixpanel/ai-plugins/enrich-data
Clone the repo
git clone --depth 1 https://github.com/mixpanel/ai-plugins

Or install mixpanel, the plugin that ships this one along with the rest of its 12 skills, 23 commands.

Per session 0 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 531 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00000 $0.00531
Opus 5 $0.00000 $0.00266
Sonnet 5 $0.00000 $0.00106
Haiku 4.5 $0.00000 $0.00053

Measured 3d ago against content hash 596b31a51577, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

enrich-data scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/mixpanel/skills/prepare-ai-readiness/commands/enrich-data.md · 28 lines

How it starts

The opening of the file, as written. The whole thing — 28 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Command: enrich-data

Set up the Lexicon metadata the agent needs to understand the data itself: event descriptions, property descriptions, and tags. This command delegates to the manage-lexicon skill run inline — it does not reimplement enrichment. Its job is to hand off cleanly and capture the result for the unified readiness status.

Session reads: org_id, project_id, project_name, caller_role Session writes: lexicon_score


Step 1 — Preconditions

  • A project must be in scope (project_id). If only org-level context was being worked on, ask which project to enrich — Lexicon is per-project.
  • The caller needs write permission for Lexicon (see SKILL.md's "Permissions gate writes" constraint for the role matrix). Check caller_role; if missing, name the required role and offer to have an admin run this step.
  • Confirm manage-lexicon is available. If it is not, follow SKILL.md's "manage-lexicon can be unavailable" constraint — additionally, point the user to the manage-lexicon skill in the Mixpanel skills repository, then return.

Step 2 — Hand off to manage-lexicon

Hand this project_id to manage-lexicon and let it do two things, in this order: first measure current metadata health (description coverage on events and properties, tag coverage) and capture that score; then fill empty event and property descriptions and add tags — using whatever entry points that skill exposes. Respect its own guardrails (verify current behavior against that skill) — expect at least fill-only-empty (never overwrite existing metadata), add tags rather than replace, and a preview + CONFIRM gate before writes. Those guardrails are the reason we delegate rather than rebuild — don't bypass them.

Let manage-lexicon own its previews and confirmations. This command does not duplicate or wrap those prompts; the user interacts with manage-lexicon's flow directly.

Step 3 — Capture result

After enrichment, record the post-run coverage into lexicon_score (events described %, properties described %, events tagged %), so status can show both layers in one readout. If manage-lexicon ran a final score, reuse it; otherwise ask it to score coverage once more to capture the after state.

Read the full file on GitHub · 28 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 3d ago First seen · 28 lines · 0 tokens per session scan A 596b31a51577

Subscribe to this mod's changes

enrich-data is a command published in the GitHub repository mixpanel/ai-plugins (15 stars, last pushed 9d ago), licensed Apache-2.0. It costs nothing until one of its globs matches a file; then it loads 531 tokens. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.