Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/treex-x/workflowx/auditxnpx skills add TreeX-X/WorkFlowX --skill auditxgit clone --depth 1 https://github.com/TreeX-X/WorkFlowXWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00014 | $0.00295 |
| Opus 5 | $0.00007 | $0.00148 |
| Sonnet 5 | $0.00003 | $0.00059 |
| Haiku 4.5 | $0.00001 | $0.00030 |
Grade A, and why
auditX scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
auditX
Use only when xflow requests an independent evaluatorX review.
Review
- Read the review task, Child acceptance criteria, changed files, and relevant diff.
- Convert each applicable acceptance criterion into a focused executable test or check.
- Run the smallest useful test set.
- Record passed tests, failed cases, failure causes, and blocking dependencies.
- Inspect code manually only to explain a failed test or a named integration risk; do not perform a full static review by default.
If the project has no runnable test path, report Unevaluable with the reason and the checks attempted. Never claim a test passed unless it was run.
Result Contract
Return a concise Evaluation Result:
- Status: PASS | NEEDS_FIX | UNEVALUABLE
- Tests Run: [commands or checks]
- Passed: [cases]
- Failed: [case, observed result, likely cause, repair scope, regression risk]
- Blockers: [dependency or environment issue, or None]
For NEEDS_FIX, keep each failure compact enough for the Main Agent to turn directly into a Repair Packet. Distinguish local defects from cross-Child integration findings when evidence supports it.
evaluatorX is read-only. The Main Agent owns document updates and any follow-up implementation.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 33 lines · 14 tokens per session scan A 1baa62a9a43e
auditX is a skill published in the GitHub repository TreeX-X/WorkFlowX (55 stars, last pushed 2d ago), licensed MIT. It adds 14 tokens to every session and 295 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
agent-orchestrator
Meta-skill que orquestra todos os agentes do ecossistema. Scan automatico de skills, match por capacidades, coordenacao de workflows multi-skill e registry management.
project-orchestration
Orchestrate multi-agent workflows for feature development using planning agents, context handoff, and stage management.
company-product-context
Compiles comprehensive company product context from PDF documents, web research, and industry knowledge.
codebase-context-extractor
This skill provides a comprehensive context extraction system for large codebases. It intelligently analyzes code structure, dependencies, and relationships to extract relevant context for understanding, debugging, or modifying code.
deep-researcher
Performs comprehensive, multi-layered research on any topic with structured analysis and synthesis of information from multiple sources.
llmtornado-tutorial-generator
Generates comprehensive code tutorials on LlmTornado API formatted for Medium publication with examples, explanations, and best practices.