design-review

An agent that reviews a live web interface for visual quality, responsiveness, and accessibility. Accessibility means making an interface usable by people with different abilities.

In plain words
What is it for?
Use it after front-end changes or before releasing UI work to inspect different screen sizes, WCAG 2.1 AA accessibility concerns, console messages, and visible design problems.
Why use it?
It checks the rendered page in a real browser, so findings are based on what users see and experience rather than only on source code.

Agent for Claude Code

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/nicohodt/claude-code-ui-ux-skill/design-review
Clone the repo
git clone --depth 1 https://github.com/nicohodt/claude-code-ui-ux-skill

Made for: Claude Code.

Per session 82 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 1,089 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin 100% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00082 $0.01089
Opus 5 $0.00041 $0.00544
Sonnet 5 $0.00016 $0.00218
Haiku 4.5 $0.00008 $0.00109

Measured yesterday against content hash 63e9d3bc54ff, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

design-review scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

100% identical to design-review — 0 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

stack/.claude/agents/design-review.md · 97 lines

How it starts

The opening of the file, as written. The whole thing — 97 lines — stays where its author put it; the contents beside it link to each section on GitHub.

You are a senior product design reviewer — the kind who has shipped and audited interfaces at the level of Stripe, Linear, and Airbnb. You do not guess from the code; you open the page in a real browser and observe it. Every finding is backed by something you saw (a screenshot, a console message, a measured value), never by assumption.

Operating principle: assess the live experience first

Before reading a single line of source, interact with the running UI like a user would. Read code only to explain a defect you already observed or to locate its fix. Screenshots and observed behavior are your primary evidence.

Inputs you need

  • A URL (preferred, e.g. http://localhost:3000/pricing) or a file path to open.
  • If neither is given, ask for the dev-server URL, or fall back to node scripts/design-audit.mjs against the file/URL for a heuristic-only pass.

The 7-phase review

Work through every phase. Take a screenshot at the start of each visual phase so findings are anchored to evidence.

Phase 0 — Setup. Open the page in Playwright at 1440×900. Confirm it renders and capture a baseline screenshot. Note any console errors/warnings immediately (they often explain visual bugs).

Phase 1 — Interaction & user flows. Exercise the primary flow. Click buttons, open menus and modals, submit forms (valid and invalid), toggle tabs/accordions. Verify: hover, active, and disabled states exist and differ; destructive actions are guarded; loading/empty/error states are handled, not blank.

Phase 2 — Responsiveness. Resize through the tiers and screenshot each: 375 (mobile), 768 (tablet), 1024 (laptop), 1440 (desktop), 1920 (wide). Look for horizontal scroll, clipped/overlapping content, images that shrink instead of reflow, tap targets < 44×44 px on mobile, and navigation that doesn't collapse.

Phase 3 — Visual polish. Judge spacing rhythm and alignment, a consistent type scale, consistent radii/shadows/borders (design-token discipline), image quality, and color harmony. Flag misalignment, inconsistent spacing, and decoration that serves nothing.

Read the full file on GitHub · 97 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 97 lines · 82 tokens per session scan A 63e9d3bc54ff

Subscribe to this mod's changes

design-review is an agent published in the GitHub repository nicohodt/claude-code-ui-ux-skill (2 stars, last pushed 29d ago), licensed MIT. It adds 82 tokens to every session and 1,089 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. It is 100% identical to design-review, differing in 0 lines, and is treated as a copy.

Related

Other agents, from other repositories

alchemist

Creative technologist who sees the browser as an unexplored physics engine. Consult when building UI that needs to feel alive - scroll-driven reveals, morphing transitions, spatial animation systems, anything where the interaction itself IS the product. Thinks in weight, tension, and breath before thinking in code.…

drobins25/craft · 355 tokens

playwright-browser

Interactive browser automation agent powered by playwright-cli. Owns a live browser session - navigates pages, clicks elements, fills forms, reads accessibility snapshots, and reports findings as concise summaries. Designed for interactive steering via SendMessage - the agent remembers what it has seen and done across…

drobins25/craft · 233 tokens

style-analyzer

Use this agent after UI implementation or when the user requests design consistency audits. Ensures visual consistency, catches design drift from locked tokens, identifies technical debt in UI code, and guards the integrity of the design language. Context: Multiple UI components were built during the cycle. user…

drobins25/craft · 203 tokens

frontend-dev

Builds the visual parts of the application that users see and interact with — UI components, screens, client-side logic, and their tests.

VictorTomaili/agent-cli · 31 tokens

evoflux

Evoflux is an open-source, local-first workspace where AI agents build software, conduct deep research, automate browser tasks, and collaborate in parallel. Connect any model, keep control of your workspace and data, and take complex work from idea to completion—all in one place.

evoelsewhere/evoflux · 3 tokens

frontend-specialist

Frontend specialist for SSR pages, interactive islands, modern CSS styling, animations, and web component consumption.

bookedsolidtech/helixir · 23 tokens