Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add promptadvisers/bench-studio-public --skill bench-studiogit clone --depth 1 https://github.com/promptadvisers/bench-studio-publicWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/promptadvisers/bench-studio-public/bench-studio)<a href="https://agentmods.dev/skills/promptadvisers/bench-studio-public/bench-studio"><img src="https://agentmods.dev/badge/skills/promptadvisers/bench-studio-public/bench-studio/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/promptadvisers/bench-studio-public/bench-studio"><img src="https://agentmods.dev/badge/skills/promptadvisers/bench-studio-public/bench-studio.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector warn
SkillSpector: 1 finding, up to medium
These are SkillSpector’s own severities. On a checked sample its high-severity flags on skills were ~96% false positives — a documented command, a public API, a “never do X” rule — so we show them as a caution to read, not a verdict. Why →
- medium Rogue Agent · line 3 Skill establishes unauthorized persistence across sessions via cron jobs, startup scripts, or state files. Session persistence allows an attacker to maintain access beyond the current interaction.Fix: Remove any persistence mechanisms (cron jobs, startup scripts, state files). Skills should not maintain state across sessions without explicit user consent.
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00074 | $0.00602 |
| Opus 5 | $0.00037 | $0.00301 |
| Sonnet 5 | $0.00015 | $0.00120 |
| Haiku 4.5 | $0.00007 | $0.00060 |
Grade A, and why
bench-studio scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
Copies of this mod
1 near-identical copy found in the catalogue:
- bench-studio — 100% identical, 0 lines differ
How it starts
The opening of the file, as written. The whole thing — 44 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Bench Studio
Use Bench as a local creative operating system. Prefer its MCP tools; fall back to the loopback API at http://localhost:8787 when the MCP server is not connected.
Route the request
- Images or videos: inspect models, inspect the chosen model's inputs, upload any media, then generate.
- Websites: start
create_website, retain the project id, pollget_project, and return the preview plus source artifact. - Documents: start
create_document, pollget_project, and return the PDF plus editable HTML preview. - Existing work: use
list_resultsfor media orlist_projectsfor websites and documents. - Pricing or storage: use
get_usagebefore making assumptions.
Media workflow
- Call
list_modelswith the requested output and required input modality. - Call
get_model_capabilitiesbefore attaching media. Treatschema-supportedas declared support, not proof of visual fidelity. - Call
upload_mediafor every local asset. Keep both its hosted URL and local archive URL. - Map each uploaded asset to an exact capability field. Never invent
image_url; models differ. - Call
create_mediawith the smallest suitable model and explicit parameters. - Return local result URLs first and hosted fal URLs second.
Do not say a reference influenced an output merely because fal accepted the field. Say “submitted successfully” unless an output was actually reviewed.
Website and document workflow
Write a concrete creative brief containing audience, purpose, mood, required sections, and constraints. Pick low depth for routine drafts and medium for consequential work. Builds use the local ChatGPT-authenticated Codex SDK and create inspectable files under Bench's project archive.
Poll until complete, failed, or cancelled. Do not report a queued job as delivered. For PDFs, visually inspect representative rendered pages before calling the document finished when the user asks for a final production artifact.
Safety and truthfulness
What ships with it
2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 10d ago First seen · 44 lines · 74 tokens per session scan A ad1877123686
bench-studio is a skill published in the GitHub repository promptadvisers/bench-studio-public (105 stars, last pushed 26d ago), licensed MIT. It adds 74 tokens to every session and 602 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
html-ppt-zhangzara-retro-zine
A neighborhood zine on the disappearing corner shops — portraits, voices, and what a block loses when they close. Built as a decision-grade story deck for community, local readers.
html-ppt-zhangzara-studio
A photography studio's portfolio-and-rate deck — the signature work, the process, and the packages that win the brief. Built as a decision-grade design craft deck for prospective clients.
motion-frames
A single-frame motion-design composition with looping CSS animations — rotating type ring, animated globe, ticking timer, parallax labels. Renders as a hero video poster you can hand straight to HyperFrames or any keyframe-based exporter. Use when the brief asks for "motion design", "animated hero", "loop", "video…
webgl-halftone-drift
A self-contained WebGL2 hero: a flowing field screened through a rotated halftone dot grid into a duotone print aesthetic; move the cursor to bend the drift.
webgl-holographic-foil
A self-contained WebGL2 hero: thin-film interference over a crushed-foil surface whose palette shifts with the viewing angle; move the cursor to tilt the film.
motion-graphics
A short, design-led motion graphic where motion is the message — kinetic typography, stat count-up, chart/data-viz hit, logo sting / brand lockup, lower-third / callout / social overlay, animated map (highlight regions, connect places, zoom to a location), animated tweet / news-article / headline, webpage / UI…