screenshot-to-code

screenshot-to-code is a skill for Codex from furqanistic/aura-skills. It costs 123 tokens per session (2,333 once invoked), scanned A, original, MIT.

A skill for rebuilding a website or application interface from screenshots or other visual references as frontend code. It covers discovery, layout and behavior decisions, implementation, screenshot comparison, and refinement across screen sizes.

In plain words
What is it for?
Use it to recreate landing pages, dashboards, application screens, components, or responsive interfaces from supplied screenshots and to compare the implementation against those references.
Why use it?
It provides a repeatable way to match visible design details while keeping the resulting interface responsive, accessible, and maintainable.

Skill for Codex

Written for Codex: agents/openai.yaml present.

Good fit Use it to recreate landing pages, dashboards, application screens, components, or responsive interfaces from supplied screenshots and to compare the implementation against those references.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/furqanistic/aura-skills/screenshot-to-code
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add furqanistic/aura-skills --skill screenshot-to-code
Clone the repo
git clone --depth 1 https://github.com/furqanistic/aura-skills

Made for: Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for screenshot-to-code

README.md
[![agentmods](https://agentmods.dev/badge/skills/furqanistic/aura-skills/screenshot-to-code/github.svg)](https://agentmods.dev/skills/furqanistic/aura-skills/screenshot-to-code)
Your own site
<a href="https://agentmods.dev/skills/furqanistic/aura-skills/screenshot-to-code"><img src="https://agentmods.dev/badge/skills/furqanistic/aura-skills/screenshot-to-code/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for screenshot-to-code

Your own site · 80×15
<a href="https://agentmods.dev/skills/furqanistic/aura-skills/screenshot-to-code"><img src="https://agentmods.dev/badge/skills/furqanistic/aura-skills/screenshot-to-code.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 123 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,333 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00123 $0.02333
Opus 5 $0.00062 $0.01167
Sonnet 5 $0.00025 $0.00467
Haiku 4.5 $0.00012 $0.00233

Measured 11d ago against content hash b58ed85cde14, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-10, from the pricing page.

Security

Grade A, and why

screenshot-to-code scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/screenshot-to-code/SKILL.md · 156 lines

How it starts

The opening of the file, as written. The whole thing — 156 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Screenshot to Code

Treat the screenshot as visual evidence, not merely inspiration. Reproduce its design language and visible behavior while keeping the implementation responsive, accessible, and maintainable.

Operating standard

Work in five gated phases. Do not skip directly from seeing the screenshot to polishing CSS:

  1. Discover — understand the screenshots, repository, target route, available assets, and runtime.
  2. Specify — turn visual evidence into layout, token, component, content, state, and responsive decisions.
  3. Construct — build semantic structure and correct macro geometry before fine styling.
  4. Compare — render at controlled viewports and measure differences against the references.
  5. Refine — fix the highest-impact mismatch, recapture, and repeat until the completion gate passes.

Favor visual fidelity over personal design preference, and engineering quality over screenshot-only hacks. A strong result must satisfy both.

Start with context

  1. Open and inspect every attached screenshot at full resolution. Record its pixel dimensions, orientation, visible browser/device chrome, crop, content density, and likely CSS viewport.
  2. Inspect the repository before coding. Identify framework, routing, styling approach, component library, design tokens, fonts, assets, breakpoints, testing tools, and the correct page or component entry point.
  3. If a matching implementation exists, render it and improve it incrementally. Preserve established architecture and working behavior.
  4. If starting from scratch, use the user's requested stack. Otherwise prefer the repository's existing stack; do not replace it just for convenience.
  5. Ask only for information that cannot be safely inferred and would materially change the result, such as an unreadable key asset or unknown target route.

For multiple screenshots, map each image to a route, viewport, UI state, and shared shell. Determine whether they show separate pages, responsive variants, modal/menu states, or steps in one flow. Do not implement each screenshot as an unrelated page when evidence shows a common system.

Read the full file on GitHub · 156 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 11d ago First seen · 156 lines · 123 tokens per session scan A b58ed85cde14

Subscribe to this mod's changes

screenshot-to-code is a skill published in the GitHub repository furqanistic/aura-skills (5 stars, last pushed 22d ago), licensed MIT. It adds 123 tokens to every session and 2,333 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

chakra-ui-builder

Build responsive, accessible UI components and layouts using Chakra UI v3, install or configure Chakra UI in new and existing projects, and design scalable themes using tokens, semantic tokens, recipes, and slot recipes. Use this skill whenever a user asks to build, create, or generate any UI component, page, form…

chakra-ui/chakra-ui · 214 tokens

visual-ralph

Visual Ralph orchestration for frontend UI from generated references, static references, or live URL targets, using $ralph with built-in visual verdict and pixel-diff evidence until the implementation matches and leaves a reproducible design system.

Yeachan-Heo/oh-my-codex · 50 tokens

frontend-visual-qa

Audits already-rendered web, landing-page, HTML deck/slide, browser tool/game, dashboard/admin, design-system, and desktop UIs using real-browser or native-app journeys, inspected screenshots, DOM geometry, responsive or projection viewports, and a bundled Playwright sweep. Use after UI implementation to find…

daymade/claude-code-skills · 145 tokens

prototype-web

A clickable, high-fidelity web product prototype with navigation, a hero section, feature cards, steps, social proof, and optional pricing. It is designed to resemble a finished landing page while remaining a prototype.

nexu-io/html-anything · 24 tokens

animation-principles

Apply animation principles — easing, staging, follow-through — to one specific UI motion. Use when tuning how an animation feels. For product-wide duration and easing tokens use motion-system (design-systems); for a full interaction spec use micro-interaction-spec.

Owl-Listener/designer-skills · 59 tokens

refactoring-ui

Audit and fix visual hierarchy, spacing, color, and depth in web UIs. Use when the user mentions "my UI looks off" (or amateur/unprofessional), "fix the design", "Tailwind styling", "color palette", "visual hierarchy", "design system", "spacing scale", or "component styling". Also trigger when building consistent…

wondelai/skills · 132 tokens