visual-qa

visual-qa is a skill for Claude Code, Codex from smileynet/teach-me. It costs 54 tokens per session (1,756 once invoked), scanned A, original, MIT.

A visual and behaviour check for lesson pages that exercises interactive parts and saves screenshots of the results.

In plain words
What is it for?
Use it after interface changes to test all pages or focus on one feature, inspect the generated manifest, and review screenshots for layout and appearance problems.
Why use it?
It shows whether features such as tooltips, quizzes, trays, and SVGs work in the browser, while separating functional failures from purely visual issues.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/smileynet/teach-me/visual-qa
Any agent
npx skills add smileynet/teach-me --skill visual-qa
Clone the repo
git clone --depth 1 https://github.com/smileynet/teach-me

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for visual-qa

README.md
[![agentmods](https://agentmods.dev/badge/skills/smileynet/teach-me/visual-qa.svg)](https://agentmods.dev/skills/smileynet/teach-me/visual-qa)
Your own site
<a href="https://agentmods.dev/skills/smileynet/teach-me/visual-qa"><img src="https://agentmods.dev/badge/skills/smileynet/teach-me/visual-qa.svg" alt="Measured on agentmods" height="20"></a>
Per session 54 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,756 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00054 $0.01756
Opus 5 $0.00027 $0.00878
Sonnet 5 $0.00011 $0.00351
Haiku 4.5 $0.00005 $0.00176

Measured yesterday against content hash 25fda432bb96, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

visual-qa scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.kiro/skills/visual-qa/SKILL.md · 151 lines

How it starts

The opening of the file, as written. The whole thing — 151 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Visual QA

Verify that UI components render and behave correctly by exercising them and analyzing the evidence.

General Check

Run the automated tool to exercise all components across all pages:

mise run visual-qa

This produces .scratch/visual-qa/manifest.json + screenshots per page. If it exits 0, all behavioral checks pass (tooltips appear, trays open, quizzes give feedback, SVGs render). If it exits 1, something is broken — read the manifest for which checks failed.

The tool is a behavioral check. It answers "does this work?" not "does this look right?"

mise run visual-qa exercises components on a page. To verify the cross-page USER JOURNEY (does clicking through actually navigate?), run the navigation suite:

mise run test:nav

It discovers all library domains from the aggregate index #page-data island (no hardcoded slugs), self-serves the library/ root headless, and for EACH domain walks aggregate → domain map → a lesson → its quiz → breadcrumb back-nav, plus the index resume cue. Navigation is asserted by act-then-verify (click → URL changes → landed <h1>), not by link-presence. Per-domain pass/fail + screenshots land in test-results/ (navigation-report.md

  • screenshots/nav-*). Exit 0 = every domain's journey navigates correctly.

NOT in core mise run verify (slower browser journey) — run it after nav/breadcrumb/map/quiz changes, or when adding a domain. The two-view Tree|Map toggle + tree keyboard model are covered separately by mise run verify's interactive gate (index_two_view_toggle, index_tree_keyboard).

Feature-Specific Visual Review

After building or modifying a specific feature, run the tool with --focus to scope screenshots, then analyze those screenshots against the feature's design intent.

python tools/visual-qa.py --serve --focus glossary

Then load the screenshots and analyze. The analysis prompt should be tailored to what the feature is supposed to look and feel like.

Read the full file on GitHub · 151 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday Changed · +21 lines 25fda432bb96
  2. 5d ago First seen · 130 lines · 54 tokens per session scan A 086a1e12ac3e

Subscribe to this mod's changes

visual-qa is a skill published in the GitHub repository smileynet/teach-me (3 stars, last pushed today), licensed MIT. It adds 54 tokens to every session and 1,756 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

anki

Anki spaced repetition learning system reference. Covers the science of SRS and SM-2 algorithm, card design principles, optimal deck settings, FSRS scheduler, study workflow, essential add-ons, custom templates, filtered decks, and AnkiConnect API.

bytesagain/ai-skills · 53 tokens

reading-metaskill

当用户想养成阅读习惯、问「读什么书/怎么读」「如何学习新领域/怎么入门某学科」时调用。 核心理念: 阅读是终极元技能; 读你所爱直到爱上阅读, 没有读完义务; 读原著与经典优先; 以教促学; 每天1-2小时即可进入极少数人行列。 不适用于: 具体某本书的书评、考试备考资料选择。 Triggers: 阅读/读书/怎么学习/入门/原著/书单/reading/how to learn.

kangarooking/cangjie-skill · 139 tokens

coach

Learning telemetry, strategy, and schedule — retention stats, calibration, grader audit, n-of-1 experiments, HTML dashboard. Use for "how am I doing", weekly check-ins, strategy questions, auditing the grader, or adjusting how Engram teaches.

nagisanzenin/engram · 53 tokens

textbook-distillation

Turn a textbook or long-form source into a self-paced learning track: intake the material, build a chapter map, draft a lesson plan, then generate self-contained HTML lecture notes in a style the human specifies (layout, palette, emphasis), each lesson carrying worked examples, exercises, and checkpoint questions.…

Lingtai-AI/lingtai · 162 tokens

classify-interview-questions

将批量面经或零散面试题逐题去重并分发:Agent/LLM/AI工程题写入 zero2Agent 的 learn-agent-interview,传统后端八股写入相邻 zero2Leetcode 的夏季八股。大批量输入使用 gpt-5.6-luna API 逐篇并发抽题和语义召回,再审查、去重和写答案;不新建面经实录文章。.

ranxi2001/zero2Agent · 109 tokens

new-article

在 zero2Agent 项目中创建新的学习文章。当用户说"写一篇新文章"、"创建文章"、"新建文章"、"在某模块下添加一篇关于X的文章"、"帮我起草一篇讲XX的内容"、"整理面经"时触发。适用于所有模块下新建内容,包括面试维度拆解文章和面经实录。即使用户没有明确说"文章",只要涉及给 zero2Agent 项目增加教学内容,也应当触发此技能。.

ranxi2001/zero2Agent · 124 tokens