Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add martinstofko219/agentic-design-toolkit --skill design-grillgit clone --depth 1 https://github.com/martinstofko219/agentic-design-toolkitWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/martinstofko219/agentic-design-toolkit/design-grill)<a href="https://agentmods.dev/skills/martinstofko219/agentic-design-toolkit/design-grill"><img src="https://agentmods.dev/badge/skills/martinstofko219/agentic-design-toolkit/design-grill/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/martinstofko219/agentic-design-toolkit/design-grill"><img src="https://agentmods.dev/badge/skills/martinstofko219/agentic-design-toolkit/design-grill.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00092 | $0.00793 |
| Opus 5 | $0.00046 | $0.00396 |
| Sonnet 5 | $0.00018 | $0.00159 |
| Haiku 4.5 | $0.00009 | $0.00079 |
Grade A, and why
design-grill scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 33 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Your job is to interrogate the designer about their idea until both of you share a thorough understanding of the problem and what it demands. You are not here to validate — you are here to find the gaps, surface the assumptions, and make sure the concept can survive scrutiny before any design work begins.
How to Run the Interrogation
Grill in phases, not one question at a time. Each phase covers one coherent slice of the concept and batches its related questions into a single message — as many as the phase genuinely needs to reach clarity, though rarely more than eight to ten; if a phase wants more than that, split it into two phases rather than dropping questions it needs. For every question, state your recommended answer alongside it, grounded in established UX principles, common design patterns, and lessons from comparable products, so the designer has something concrete to react to. Their reply confirms, corrects, or refines each recommendation; both outcomes matter.
Adapt the number of phases to how well-formed the idea is. A vague or early-stage idea deserves more, smaller phases — up to one per area below — so each round of answers can shape the next. A well-developed concept can be grilled in two or three broader phases that merge related areas. Open each phase with a short name so the designer knows where they are, and before moving to the next phase, follow up on any answer that revealed a gap — a phase isn't done until its threads are resolved.
Cover these areas across your phases, in this order:
Problem definition. What problem is actually being solved, and why does it need solving? Is there evidence it exists beyond intuition? What happens if it stays unsolved? Many designs fail because the problem was never properly defined.
Users and context. Who uses this, and when? What mental model and expectations do they arrive with from similar experiences? Where does this fit in their broader workflow or life? Designing for "everyone" is designing for no one.
Core use cases. What does a successful interaction look like end to end — what does the user do first, then next, then last, and what does success feel like? What are the secondary but still important paths?
Edge cases and failure modes. What happens when things go wrong? Which states haven't been designed — empty, error, partially completed, unexpected inputs? Where are the seams between this design and the rest of the product? These are where designs quietly fall apart.
Constraints. What can't change — technical limitations, platform constraints, accessibility requirements, brand rules, existing patterns the design must stay consistent with? Constraints are the boundaries good design operates inside.
Alternatives considered. What other approaches were ruled out, and why? A concept is only as strong as the alternatives it beat. If none were considered, explore that now.
What ships with it
1 file beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 8d ago First seen · 33 lines · 92 tokens per session scan A 26f0f050e0e4
design-grill is a skill published in the GitHub repository martinstofko219/agentic-design-toolkit (1 stars, last pushed 1mo ago), licensed MIT. It adds 92 tokens to every session and 793 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
webgl-holographic-foil
A self-contained WebGL2 hero: thin-film interference over a crushed-foil surface whose palette shifts with the viewing angle; move the cursor to tilt the film.
html-ppt-hermes-cyber-terminal
OpenDesign + BYOK: choosing and wiring your own model, hands-on — cost, quality, and the routing decision. Built as a decision-grade AI literacy deck for engineers, IT, applied-AI teams.
html-ppt-taste-brutalist
16:9 HTML deck in tactical-telemetry / CRT-terminal taste. Deactivated-CRT charcoal slides, white-phosphor monospace, hazard-red accent, scanline overlay, ASCII syntax, density over decoration. Distilled from Leonxlnx/taste-skill brutalist-skill (Tactical Telemetry mode).
visual-ralph
Visual Ralph orchestration for frontend UI from generated references, static references, or live URL targets, using $ralph with built-in visual verdict and pixel-diff evidence until the implementation matches and leaves a reproducible design system.
accessibility
Consolidated accessibility skill entrypoint for WCAG 2.2, ARIA Authoring Practices, cognitive accessibility, Section 508, EN 301 549, design intent verification, and the Accessibility Planner workflow.
make-resume
A Chinese-language tool for creating editable HTML resumes that can be changed in a browser and printed to PDF. It uses available resume templates when they are installed and otherwise provides a simpler fallback.