running-tests

A guide for running the project's automated tests through its provided test command, choosing checks that match the changes made.

In plain words
What is it for?
Use it after code changes to run affected tests, a single module's tests, image tests, all tests, or the complete test suite.
Why use it?
It avoids using the wrong test command or running more checks than necessary. It also helps diagnose simulator launch problems and review image comparison failures.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/kyleve/stuff/running-tests
Any agent
npx skills add kyleve/Stuff --skill running-tests
Clone the repo
git clone --depth 1 https://github.com/kyleve/Stuff

Made for: Claude Code, Codex.

Per session 44 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,376 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00044 $0.01376
Opus 5 $0.00022 $0.00688
Sonnet 5 $0.00009 $0.00275
Haiku 4.5 $0.00004 $0.00138

Measured yesterday against content hash 3c7d0900869e, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

running-tests scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.agents/skills/running-tests/SKILL.md · 136 lines

How it starts

The opening of the file, as written. The whole thing — 136 lines — stays where its author put it; the contents beside it link to each section on GitHub.

How to run tests in this repo. Read root AGENTS.md for always-on rules. Use ./test. Do not hand-roll tuist test or xcodebuild. Run checks in proportion to the change. Canonical flag list: ./test --help. Rationale for ./test over alternatives: header comment in test.

Documentation-only changes

Pure documentation or comment-only changes can skip ./test. Skip ./swiftformat --lint when the changed files are outside the formatter's scope. Record skipped checks and the reason in the commit or PR validation.

Do not classify a semantic change to configuration, scripts, generator inputs, executable examples, or app-rendered copy as documentation-only. Run the narrowest applicable checks below instead.

Pick a tier

Pick the narrowest tier that covers the change:

Tier Command When
Affected ./test Default — bundles touched by your diff against origin/main
One bundle ./test WhereCoreTests You know exactly what you touched
Unit suite ./test --all Change spans modules; before a wide commit
Image suite ./test --snapshots Triggers below
Everything ./test --everything Full revalidation; what CI runs

Examples:

  • Edited WhereCore only → ./test (or ./test WhereCoreTests if you want to be explicit)
  • Edited WhereCore + WhereUI./test or ./test --all before committing
  • Changed a stylesheet token that renders → ./test --snapshots (or ./test if the graph already pulls snapshots in)

Compare against a ref other than origin/main: ./test --base REF.

Architecture and host checks

Every normal invocation runs the complete Bumper Bowling sequence first. Use ./test --architecture-only to run only the configuration validation, rule tests, and architecture lint.

CI test jobs use --skip-architecture because the dedicated Bumper job owns that sequence. Do not use this flag for normal local validation.

Read the full file on GitHub · 136 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 136 lines · 44 tokens per session scan A 3c7d0900869e

Subscribe to this mod's changes

running-tests is a skill published in the GitHub repository kyleve/Stuff (2 stars, last pushed 2d ago), licensed Apache-2.0. It adds 44 tokens to every session and 1,376 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

nemoclaw-contributor-implement-issue

Implement an accepted NemoClaw GitHub issue in the current checkout. Use when a user asks to pick up an issue for implementation, implement or fix a named issue, or add the issue's tests. Confirm accepted scope, deliver the smallest independently valuable capability slice, and record validation and remaining gates…

NVIDIA/NemoClaw · 134 tokens

amazon-bestseller-listing

Amazon Best Sellers listing scraper: extract product cards from any Amazon Best Sellers (zgbs) or /gp/bestsellers/ category page — returns rank (position on chart), asin, title, url, image, imageAlt, price, stars, reviewCount, ratingRaw per item, plus category metadata (categoryName, categoryFullName, categoryUrl) and…

browser-act/skills · 300 tokens

amazon-reviews-api-skill

This skill helps users automatically extract Amazon product reviews via the Amazon Reviews API. Agent should proactively apply this skill when users express needs like getting reviews for Amazon product with ASIN B07TS6R1SF, analyzing customer feedback for a specific Amazon item, getting ratings and comments for a…

browser-act/skills · 124 tokens

amazon-competitor-analyzer

Scrapes Amazon product data from ASINs using browseract.com automation API and performs surgical competitive analysis. Compares specifications, pricing, review quality, and visual strategies to identify competitor moats and vulnerabilities.

browser-act/skills · 48 tokens

cabloy-worktree-environment

This skill must be used only when the user explicitly invokes /cabloy-worktree-environment or explicitly asks to perform the named Cabloy worktree-environment setup. It prepares a confirmation-gated, worktree-local Vona and Zova runtime environment for a linked Cabloy Basic or Cabloy Start Git worktree using Git…

cabloy/cabloy · 126 tokens

shopify-hydrogen

Hydrogen storefront implementation cookbooks. Some of the available recipes are: B2B Commerce, Bundles, Combined Listings, Custom Cart Method, Dynamic Content with Metaobjects, Express Server, Google Tag Manager Integration, Infinite Scroll, Legacy Customer Account Flow, Markets, Partytown + Google Tag Manager…

Shopify/Shopify-AI-Toolkit · 106 tokens