independent-review-loop

independent-review-loop is a skill for Claude Code, Codex from DheerG/swarms. It costs 114 tokens per session (8,476 once invoked), scanned A, original, MIT.

A pre-delivery review process for code changes in which an independent reviewer reads the entire pull request, a proposed set of code changes, against the intended result. The author fixes relevant problems and the review repeats until none remain.

In plain words
What is it for?
Reviewing pull requests before delivery, finding functional problems, and confirming that in-scope fixes are complete.
Why use it?
It catches issues the original author may miss by using a separate reviewer and checking the full change repeatedly.

Skill for Claude CodeCodex

Written for Claude Code and Codex: user-invocable in frontmatter, but also runs codex exec. Also seen: mentions subagents; names the AskUserQuestion tool; mentions Codex.

Part of the swarms plugin — 13 skills, 10 commands, 2 agents, 2 hooks shipped together

Good fit Reviewing pull requests before delivery, finding functional problems, and confirming that in-scope fixes are complete.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/dheerg/swarms/independent-review-loop
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add DheerG/swarms --skill independent-review-loop
Clone the repo
git clone --depth 1 https://github.com/DheerG/swarms

Made for: Claude Code, Codex.

Or install swarms, the plugin that ships this one along with the rest of its 13 skills, 10 commands, 2 agents, 2 hooks.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for independent-review-loop

README.md
[![agentmods](https://agentmods.dev/badge/skills/dheerg/swarms/independent-review-loop/github.svg)](https://agentmods.dev/skills/dheerg/swarms/independent-review-loop)
Your own site
<a href="https://agentmods.dev/skills/dheerg/swarms/independent-review-loop"><img src="https://agentmods.dev/badge/skills/dheerg/swarms/independent-review-loop/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for independent-review-loop

Your own site · 80×15
<a href="https://agentmods.dev/skills/dheerg/swarms/independent-review-loop"><img src="https://agentmods.dev/badge/skills/dheerg/swarms/independent-review-loop.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 114 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 8,476 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • NVIDIA SkillSpector warn 7 Sept 2026
SkillSpector: 1 finding, up to high

These are SkillSpector’s own severities. On a checked sample its high-severity flags on skills were ~96% false positives — a documented command, a public API, a “never do X” rule — so we show them as a caution to read, not a verdict. Why →

  • high Prompt Injection · line 84
    Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
    Fix: Audit all comments and invisible characters. Remove any instructions that direct the agent to perform unauthorized actions. Use plain, reviewable content.
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00114 $0.08476
Opus 5 $0.00057 $0.04238
Sonnet 5 $0.00023 $0.01695
Haiku 4.5 $0.00011 $0.00848

Measured 10d ago against content hash 2582fcda2da9, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-09, from the pricing page.

Security

Grade A, and why

independent-review-loop scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/independent-review-loop/SKILL.md · 146 lines

How it starts

The opening of the file, as written. The whole thing — 146 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Operational spec for the team lead. This skill runs an independent review pass at delivery: a reviewer that fails differently from the author reads the whole PR against the approved outcome, the lead fixes what is in scope, and the loop repeats until the reviewer finds no more in-scope functional issues. It is the automated form of "ship, then loop the PR through Codex until it stops finding edge cases" — with the lead acting as the operator who keeps findings on-scope.

What distinguishes this from the team's own review and the recursive-refinement ladder is independence and exhaustiveness: the reviewer is a different model (Codex) or a fresh-context agent that fails differently from the authors, and it runs to exhaustion — until no in-scope functional finding remains — rather than the bounded rung ladder. (The ladder hunts bugs too, as it drives the work to the full scope of the outcome; the new axis here is the independent eye run to clean, not bug-hunting per se.) It is a distinct pass; it never replaces the ladder.

Only the lead runs this. Reviewers (Codex, or fresh subagents) are read-only; the lead is the sole writer, same as every other phase.

When this runs

Invoked from the unified pre-ship gate in code-mode's Refine/Deliver (and /swarm:refine) — the gate the lead presents once the team reaches 9/10+, offering (thoroughness-descending): recursive refinement + independent review / recursive refinement only / independent review loop only / ship as is. This skill runs for the two options that include the independent loop — recursive refinement + independent review (after the recursive ladder completes) and independent review loop only (on its own); recursive refinement only and ship as is do not invoke it. Deliver invokes this skill when the independent-review-loop pending run-state task recorded at the gate is open — it branches on that recorded task, not on recall of the pick.

It runs after the ship steps complete (PR-then-loop on the common path): when a PR was created, the loop reviews the PR's diff and pushes each round's fixes to it; on a commit-only / push-only ship with no PR, it reviews the pushed branch against its base. The diff base is resolvable from the PR when one exists (see the loop's base resolution below for the no-PR fallback).

Read the full file on GitHub · 146 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 10d ago First seen · 146 lines · 114 tokens per session scan A 2582fcda2da9

Subscribe to this mod's changes

independent-review-loop is a skill published in the GitHub repository DheerG/swarms (85 stars, last pushed 1mo ago), licensed MIT. It adds 114 tokens to every session and 8,476 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

pocket-development

Use when executing implementation plans of one or more tasks. Trigger on execute plan, delegate tasks, dispatch subagents. Combines delegate handoff discipline with prompt-engineering attention mechanics.

rfxlamia/pocketto · 40 tokens

pocket-planning

Converts a pocket-grinding spec into a TDD-structured execution plan of full Pocket Packets. Use when pocket-grinding handoff arrives (spec path + acceptance criteria). Trigger on "create plan", "build plan", "pocket-planning", or when pocket-grinding skill invokes this. Outputs tasks ready to dispatch via…

rfxlamia/pocketto · 76 tokens

pocket-grinding

BDD-driven feature/fix discovery before any implementation. Use when planning a feature, designing a fix, or exploring options before building. Trigger on "pocket-grinding", "brainstorm", "think through", "plan this", "before we build". Invokes pocket-planning at handoff.

rfxlamia/pocketto · 65 tokens

pocket-help

Onboarding and routing guide for the Pocket skill ecosystem. Use when someone asks what Pocket is, which Pocket skill to use, how the end-to-end flow works, or when Pocket is better than lighter Superpowers-style flows. Trigger on "what is pocket", "how do I use pocket", "which pocket skill", "explain pocket"…

rfxlamia/pocketto · 112 tokens

pocket-pitching

Pre-grinding problem exploration. Use BEFORE pocket-grinding when the problem is unformed or needs exploration. Guides diverge→converge with structured brainstorming methods and LLM-to-LLM curation (advisor tool), then produces a pitch exploration doc. Trigger on "pocket-pitching", "pitch this", "explore this idea"…

rfxlamia/pocketto · 108 tokens

structured-research

Standalone skill for validating an explicit assumption with structured research. Use when a user holds a belief that is still an assumption — a technical claim, a library's behavior, a "this is probably how X works" — and wants it investigated methodically before it leaks into planning or code. Recommends a research…

rfxlamia/pocketto · 147 tokens