decision-challenge

decision-challenge is a skill for Claude Code, Codex from ericrisco/rsc-harness. It costs 71 tokens per session (1,325 once invoked), scanned A, original, MIT.

A bounded challenge of a high-impact plan or architecture claim before the team commits to it.

In plain words
What is it for?
It helps assess irreversible migrations, production authentication changes, costly architecture choices, launch exceptions, and other consequential commitments.
Why use it?
It exposes weak assumptions and missing evidence while there is still time to change course. It ends with a clear decision to proceed, hold, or stop.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit It helps assess irreversible migrations, production authentication changes, costly architecture choices, launch exceptions, and other consequential commitments.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/ericrisco/rsc-harness/decision-challenge
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add ericrisco/rsc-harness --skill decision-challenge
Clone the repo
git clone --depth 1 https://github.com/ericrisco/rsc-harness

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for decision-challenge

README.md
[![agentmods](https://agentmods.dev/badge/skills/ericrisco/rsc-harness/decision-challenge.svg)](https://agentmods.dev/skills/ericrisco/rsc-harness/decision-challenge)
Your own site
<a href="https://agentmods.dev/skills/ericrisco/rsc-harness/decision-challenge"><img src="https://agentmods.dev/badge/skills/ericrisco/rsc-harness/decision-challenge.svg" alt="Measured on agentmods" height="20"></a>
Per session 71 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,325 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • NVIDIA SkillSpector pass 7 Sept 2026
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00071 $0.01325
Opus 5 $0.00036 $0.00662
Sonnet 5 $0.00014 $0.00265
Haiku 4.5 $0.00007 $0.00133

Measured 4d ago against content hash 1aacaa4224b0, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-07, from the pricing page.

Security

Grade A, and why

decision-challenge scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/decision-challenge/SKILL.md · 114 lines

How it starts

The opening of the file, as written. The whole thing — 114 lines — stays where its author put it; the contents beside it link to each section on GitHub.

decision-challenge — doubt with a stopping rule

Use this immediately before a consequential commitment: an irreversible migration, a production-auth change, a costly architecture bet, a launch exception, or a plan whose confidence is higher than its evidence. The purpose is to find the decision-changing doubt, not to perform scepticism forever.

This skill is read-only unless the user also asks to change the artifact. It emits a challenge record and a verdict. Finished code review stays with ../review/SKILL.md or ../code-review/SKILL.md; consistency across constitution/spec/plan/tasks stays with ../analyze/SKILL.md.

One bounded cycle

MATERIALIZE → ISOLATE → ATTACK → TEST → RECONCILE → VERDICT

1. MATERIALIZE the decision

Do not challenge a cloud of conversation. Write the concrete packet:

  • decision or claim being made;
  • artifact and exact revision it applies to;
  • constraints and invariants that must hold;
  • evidence the author relies on;
  • known unknowns;
  • cost of a false positive (unnecessary stop) and false negative (unsafe proceed).

If the decision cannot be stated in one sentence, split it into claims. “The migration is safe” is not atomic; lock duration, dependency completeness, rollback time and consumer compatibility are separate claims.

2. ISOLATE claim from persuasion

Create a compact challenge packet containing the artifact, contract, constraints and evidence — not the author’s confidence, status, prestige or preferred conclusion. This reduces anchoring.

A genuinely fresh pass is useful when the platform and user authorize one: another agent/model/context can receive only the packet. It is optional, never automatic, and never an excuse to leak data or broaden tool authority. When independent execution is unavailable, make the isolation explicit and challenge locally.

3. ATTACK each material claim

For each claim, ask:

  • What observation would make it false?
  • Which dependency or consumer is missing from the inventory?
  • Is the evidence direct, current and representative, or an analogy?
  • What timing, ordering, concurrency or partial-failure path is assumed away?
  • What privilege, data-quality or human handoff must work perfectly?
  • If rollback is promised, has restoration time and data reconciliation been proved?
  • Can the decision be made reversible or staged before accepting this risk?

Read the full file on GitHub · 114 lines

Files

What ships with it

2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 114 lines · 71 tokens per session scan A 1aacaa4224b0

Subscribe to this mod's changes

decision-challenge is a skill published in the GitHub repository ericrisco/rsc-harness (70 stars, last pushed today), licensed MIT. It adds 71 tokens to every session and 1,325 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.

Related

Other skills, from other repositories

merge-aliases

Folds two surface names for the same backend system into one canonical entity, keeping every original mention individually retrievable, and refuses to merge pairs that only share spelling.

ayeshakhalid192007-dev/graph-engineering-crash-course · 37 tokens

anchor-and-lock

Consults a check that sits outside the loop system before finalizing any decision the frozen facts bear on, and refuses every attempt by a loop to rewrite a node marked frozen, regardless of how convergent the loop's own reasoning looks.

ayeshakhalid192007-dev/graph-engineering-crash-course · 50 tokens

counter-metric-check

Compares a watched loop's headline reading against an independently owned counter-metric reading for each period in a governance graph, and produces a governance edge for every period where the counter-metric crosses its recorded ceiling.

ayeshakhalid192007-dev/graph-engineering-crash-course · 46 tokens

traverse-multi-hop

Expresses a multi-hop lineage question as a single variable-length path match against native graph storage, bounded by an explicit hop depth and an explicit relationship-type allowlist, instead of a recursive relational join that grows one level per hop.

ayeshakhalid192007-dev/graph-engineering-crash-course · 51 tokens

query-graph

Loads schema.sql into a local SQLite file, then answers availability and provenance questions against the nodes/edges tables with real SQL instead of re-reading source material.

ayeshakhalid192007-dev/graph-engineering-crash-course · 34 tokens

build-subgraph

Traverses a full graph and returns a bounded subgraph around one target node -- its depth-bounded dependencies plus any disputed claims attached to it -- while proving everything else was left out.

ayeshakhalid192007-dev/graph-engineering-crash-course · 40 tokens