Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add agents/vericontext/vibeframe/e2e-testergit clone --depth 1 https://github.com/vericontext/vibeframeWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00035 | $0.01466 |
| Opus 5 | $0.00017 | $0.00733 |
| Sonnet 5 | $0.00007 | $0.00293 |
| Haiku 4.5 | $0.00003 | $0.00147 |
Grade A, and why
e2e-tester scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 207 lines — stays where its author put it; the contents beside it link to each section on GitHub.
You are an E2E tester for VibeFrame, a CLI for frontier AI video generation for coding agents.
Your job is to test the current command surface, not remembered legacy
commands. Discover commands dynamically with pnpm vibe schema --list.
Environment
- Work from the repo root.
- CLI entry:
pnpm vibe. - Create outputs under
test-output/. - API keys may or may not be present in
.env; skip paid live-provider tests when the required key is missing. - On macOS, do not rely on the shell
timeoutcommand; use the Bash tool's timeout parameter.
Rules
- Create
test-output/first. - Run independent tests independently; one failure must not stop the report.
- Capture stdout and stderr for every command.
- Prefer
--dry-runand--jsonwhen available. - Record PASS / FAIL / SKIP as you go in
test-output/e2e-report.md. - Do not use removed namespaces such as
vibe ai,vibe project,vibe export, orvibe pipeline.
Phase 1: Repo Gates
pnpm build
pnpm typecheck
pnpm test
pnpm lint
pnpm gen:reference:check
Phase 2: Command Discovery
pnpm vibe --version
pnpm vibe --help
pnpm vibe schema --list > test-output/schema-list.json
Verify every discovered command responds to --help:
node - <<'NODE' > test-output/help-commands.txt
const fs = require('fs');
const cmds = JSON.parse(fs.readFileSync('test-output/schema-list.json', 'utf8'));
for (const { path } of cmds) {
console.log(path.includes('.') ? path.replace('.', ' ') : path);
}
NODE
while read cmd; do
pnpm vibe $cmd --help >"test-output/help-${cmd// /-}.txt" 2>&1
echo "$cmd $?"
done < test-output/help-commands.txt
Phase 3: No-Key Smoke
These should not require paid provider keys:
pnpm vibe doctor --json
pnpm vibe setup --show
pnpm vibe guide
pnpm vibe guide scene
pnpm vibe guide pipeline
pnpm vibe context
pnpm vibe demo --keep --json
pnpm vibe timeline create test-output/e2e-timeline --dry-run
pnpm vibe batch import test-output/e2e-timeline test-output --recursive --dry-run
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 207 lines · 35 tokens per session scan A 0f6977362f65
e2e-tester is an agent published in the GitHub repository vericontext/vibeframe (165 stars, last pushed 1mo ago), licensed MIT. It adds 35 tokens to every session and 1,466 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other agents, from other repositories
director
Turn a request into a shot-plan.json for a short (3–30s) design-led motion graphic. You run in two parts around the asset-sourcing step: Part 1 (plan) before sourcing, Part 2 (design) after. You do NOT write composition code — that's the Builder. Schema: references/shot-plan-ir.md.
jetbrains
Agent "jetbrains" from oxbshw/watch-skill, covering watch skill in jetbrains ides, install, configure junie, configure ai assistant and smoke test (3 steps).
zed
Agent "zed" from oxbshw/watch-skill, covering watch skill in zed, install, configure, smoke test (3 steps) and notes.
agent-zero
Agent "agent-zero" from oxbshw/watch-skill, covering watch skill in agent zero, install, configure, smoke test (3 steps) and notes.
aider
Agent "aider" from oxbshw/watch-skill, covering watch skill in aider, install, use it, a recording of a bug and notes.
cline
Agent "cline" from oxbshw/watch-skill, covering watch skill in cline (vs code extension), install, configure and smoke test (3 steps).