Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add qarium/goga --skill goga-define-successgit clone --depth 1 https://github.com/qarium/gogaWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/qarium/goga/goga-define-success)<a href="https://agentmods.dev/skills/qarium/goga/goga-define-success"><img src="https://agentmods.dev/badge/skills/qarium/goga/goga-define-success/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/qarium/goga/goga-define-success"><img src="https://agentmods.dev/badge/skills/qarium/goga/goga-define-success.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector pass
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00005 | $0.01389 |
| Opus 5 | $0.00003 | $0.00694 |
| Sonnet 5 | $0.00001 | $0.00278 |
| Haiku 4.5 | $0.00001 | $0.00139 |
Grade A, and why
goga-define-success scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 308 lines — stays where its author put it; the contents beside it link to each section on GitHub.
goga-define-success
Purpose
Define the criteria that determine whether the product change successfully solves the identified problem and delivers the intended user experience.
Transform the problem, goals, user experience, requirements, and scope into a concise set of observable success criteria that can be used to verify the completed product.
These criteria are an intermediate PRD artifact. They are not intended to become a long-term analytics or product-metrics framework.
Contract
consume:
- problem
- goals
- user_experience
- requirements
- scope
produce:
- success_criteria
Core Principle
Success criteria answer:
How will we know that the implemented product actually delivers what we decided to build?
They should verify the intended product outcome and behaviour.
Do not turn this stage into a metrics-design exercise.
Only define metrics when a quantitative measure is necessary to make the PRD sufficiently precise for engineering and subsequent validation.
Product Interview
Success criteria must reflect what the user considers a successful product outcome.
Do not invent success metrics or targets merely because they appear useful.
Interview the user when it is unclear:
- what outcome should define success;
- which outcome is most important;
- whether a quantitative measure is actually necessary;
- whether several possible success definitions imply different product priorities.
Prefer observable outcomes over arbitrary metrics.
If a quantitative target is not explicitly justified by the product context, do not invent one.
Process
1. Start from the goals
For each important goal, determine what evidence would demonstrate that the goal has been achieved.
Use:
Goal
↓
Expected outcome
↓
Success criterion
A criterion should make success observable rather than merely restating the goal.
2. Validate the user experience
Check whether the important user scenarios described in user_experience can be verified.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 8d ago First seen · 308 lines · 5 tokens per session scan A 70918fcf55c9
goga-define-success is a skill published in the GitHub repository qarium/goga (29 stars, last pushed 4d ago), licensed BSD-3-Clause. It adds 5 tokens to every session and 1,389 once invoked, about $0.0000 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.
Other skills, from other repositories
doubt-driven-review
In-flight adversarial check on a non-trivial decision BEFORE it stands — distinct from post-hoc review of a finished diff. Use on "stress-test this decision", "are we sure about this", "verify before commit", "poke holes in this", when working in unfamiliar code, or before an irreversible step (migration, prod deploy…
release-cut
Cut a new pi-agent-dashboard release: promote ## [Unreleased] in CHANGELOG.md, bump every workspace package.json per SemVer, commit, tag v , and push — triggering the Release workflow that publishes every non-private workspace, builds the Electron artifacts, and creates a GitHub Release. Use on "cut a release"…
spec-coherence-check
Sweep all active OpenSpec proposals for staleness, conflicts, and obsolescence against the current codebase and archived changes. Use when proposals may be outdated, when checking cross-proposal conflicts, or before starting a batch of implementations. Produces a gap-analysis report, updates a priority queue file, and…
ship-it
Worktree-side implementation orchestrator for an OpenSpec change. Idempotent: gates automated scenarios on filesystem reality, owns the red-test fix loop, runs the docker harness with always-teardown, then drives ship-change inline. Escape hatch writes SHIPITBLOCKED.md. Runnable headless. Triggers: "ship it", "build…
faq-mine
Mine docs/faq.md from README.md, docs/.md, and the pi-hermes memory stores. Dispatches @fast subagents per source, dedupes against the existing FAQ, and merges entries in caveman style. Use when asked to "build / regenerate / extend the FAQ", "mine docs into FAQ", "mine hermes memory into FAQ", "surface runtime…
session-to-guideline
Turn a pi session into a Markdown "how-we-did-it" collaboration guideline: reads the session's JSONL transcript and synthesizes a reusable playbook of which prompts worked, what had to be steered, and how to reproduce the result faster. Use when: "document this session", "write up how we did X with the AI", "make a…