usability-testing

usability-testing is a skill for Claude Code, Codex from liqiongyu/lenny_skills_plus. It costs 25 tokens per session (1,976 once invoked), scanned A, original, Apache-2.0.

A method for testing how easily people can complete specific tasks in a product, prototype, or simulated flow. It includes the study plan, participant tasks, session script, findings, and recommended fixes.

In plain words
What is it for?
Use it to plan moderated tests, evaluate live products or prototypes, test an idea with a fake door or Wizard of Oz setup, and synthesize findings.
Why use it?
It reveals where users become confused or stuck based on observed behavior, rather than guesses from the team. The results can turn usability problems into a prioritized list of changes.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit Use it to plan moderated tests, evaluate live products or prototypes, test an idea with a fake door or Wizard of Oz setup, and synthesize findings.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/liqiongyu/lenny_skills_plus/usability-testing
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add liqiongyu/lenny_skills_plus --skill usability-testing
Clone the repo
git clone --depth 1 https://github.com/liqiongyu/lenny_skills_plus

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for usability-testing

README.md
[![agentmods](https://agentmods.dev/badge/skills/liqiongyu/lenny_skills_plus/usability-testing/github.svg)](https://agentmods.dev/skills/liqiongyu/lenny_skills_plus/usability-testing)
Your own site
<a href="https://agentmods.dev/skills/liqiongyu/lenny_skills_plus/usability-testing"><img src="https://agentmods.dev/badge/skills/liqiongyu/lenny_skills_plus/usability-testing/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for usability-testing

Your own site · 80×15
<a href="https://agentmods.dev/skills/liqiongyu/lenny_skills_plus/usability-testing"><img src="https://agentmods.dev/badge/skills/liqiongyu/lenny_skills_plus/usability-testing.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 25 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,976 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00025 $0.01976
Opus 5 $0.00013 $0.00988
Sonnet 5 $0.00005 $0.00395
Haiku 4.5 $0.00003 $0.00198

Measured 9d ago against content hash f50fb736102d, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

usability-testing scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/usability-testing/SKILL.md · 138 lines

How it starts

The opening of the file, as written. The whole thing — 138 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Usability Testing

Scope

Covers

  • Designing task-based usability studies tied to a specific product decision
  • Testing live flows, prototypes, and “faked” implementations (fake door, Wizard of Oz)
  • Running moderated sessions (remote or in-person) and capturing high-quality evidence
  • Turning findings into a prioritized fix list (including high-ROI microcopy/CTA improvements)

When to use

  • “Create a usability test plan and script for .”
  • “We need to test a prototype with 5–8 users next week.”
  • “Validate a value proposition before building (fake door / Wizard of Oz).”
  • “Help me synthesize usability findings into a prioritized backlog.”

When NOT to use

  • You need statistically reliable estimates or causal impact (use analytics/experimentation)
  • You need open-ended discovery (“what problems do users have?”) without a specific flow to evaluate (use conducting-user-interviews)
  • You need a design critique or heuristic review without live user sessions (use running-design-reviews)
  • You need to write specs or design docs for a feature, not test an existing flow (use writing-specs-designs)
  • You need to apply behavioral/persuasion design patterns to a flow (use behavioral-product-design); this skill evaluates usability, not designs behavioral nudges
  • You’re working with high-risk populations or sensitive topics (medical, legal, minors) without appropriate approvals/training
  • You don’t have a concrete scenario/flow to evaluate (clarify the decision first)

Inputs

Minimum required

  • Product + target user segment (who, context of use)
  • The decision this test should inform (what will change) + timeline
  • What you’re testing (flow/feature) + prototype/build link (or “recommend stimulus”)
  • Platform + environment (web/mobile/desktop; remote/in-person)
  • Constraints: session type, number of participants, incentives, recording policy, privacy constraints

Missing-info strategy

  • Ask up to 5 questions from references/INTAKE.md.
  • If still unknown, proceed with explicit assumptions and list Open questions that would change the plan.

Read the full file on GitHub · 138 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 9d ago First seen · 138 lines · 25 tokens per session scan A f50fb736102d

Subscribe to this mod's changes

usability-testing is a skill published in the GitHub repository liqiongyu/lenny_skills_plus (52 stars, last pushed 3mo ago), licensed Apache-2.0. It adds 25 tokens to every session and 1,976 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.

Related

Other skills, from other repositories