consistency-reviewer

consistency-reviewer is an agent for Claude Code from OutSystems/outsystems-mcp. It costs 48 tokens per session (2,632 once invoked), scanned A, original, MIT.

A review agent that checks whether an agent-facing repository accurately describes what its instructions, slash commands, and plugin manifests actually do.

In plain words
What is it for?
Use it to review skill documents, command definitions, and plugin manifests for mismatched names, options, behavior, or shipped files.
Why use it?
It catches drift between documentation and behavior, which can cause an AI agent to follow outdated or incorrect instructions.

Agent for Claude Code

Written for Claude Code: $ARGUMENTS substitution. Also seen: model in frontmatter; mentions CLAUDE.md.

Part of the outsystems plugin — 2 skills, 1 command, 4 agents shipped together

Good fit Use it to review skill documents, command definitions, and plugin manifests for mismatched names, options, behavior, or shipped files.

Compare 6 agents from other repositories ↓
Install with agentmods
npx agentmods add agents/outsystems/outsystems-mcp/consistency-reviewer
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Clone the repo
git clone --depth 1 https://github.com/OutSystems/outsystems-mcp

Made for: Claude Code.

Or install outsystems, the plugin that ships this one along with the rest of its 2 skills, 1 command, 4 agents.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for consistency-reviewer

README.md
[![agentmods](https://agentmods.dev/badge/agents/outsystems/outsystems-mcp/consistency-reviewer.svg)](https://agentmods.dev/agents/outsystems/outsystems-mcp/consistency-reviewer)
Your own site
<a href="https://agentmods.dev/agents/outsystems/outsystems-mcp/consistency-reviewer"><img src="https://agentmods.dev/badge/agents/outsystems/outsystems-mcp/consistency-reviewer.svg" alt="Measured on agentmods" height="20"></a>
Per session 48 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 2,632 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00048 $0.02632
Opus 5 $0.00024 $0.01316
Sonnet 5 $0.00010 $0.00526
Haiku 4.5 $0.00005 $0.00263

Measured 3d ago against content hash 7772131149c4, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-07, from the pricing page.

Security

Grade A, and why

consistency-reviewer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.claude/agents/consistency-reviewer.md · 137 lines

How it starts

The opening of the file, as written. The whole thing — 137 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Consistency Reviewer Agent

Purpose

Verify that this repo's declarations match what actually happens when an agent follows them. This repo ships no server code - its product IS the text an LLM agent reads: five parallel skill documents (skills/outsystems/SKILL.md, kiro/outsystems/steering/skill.md, copilot/skill.md, cursor/skills/outsystems/SKILL.md, root SKILL.md), slash-command definitions under commands/ (frontmatter description / argument-hint plus body), the Kiro POWER.md operator doc, and the plugin manifests that declare what ships (.claude-plugin/plugin.json, .claude-plugin/marketplace.json, and their Cursor counterparts under cursor/).

A wrong instruction here is the same defect class as a wrong status code in a service repo, with one difference that raises the stakes: the reader is a machine that will act on the instruction immediately, with no human sanity check between the promise and the action. Treat every skill doc, command file, and manifest field as a contract, not as copy.

Scope

  • Agent-facing contract drift - an instruction, a slash command's frontmatter, or a manifest field versus what actually happens: a renamed or removed MCP tool argument the skill doc still tells the agent to pass, a slash command's argument-hint that no longer matches how the body parses $ARGUMENTS, an install step that no longer matches the actual claude mcp add / mcp.json shape for that harness
  • Cross-harness lockstep drift - per CLAUDE.md's "Skill docs must stay in lockstep across hosts": a behavioral rule (a confirm-before-destructive rule, a new caveat, a changed workflow) added to one of the five skill docs (or the curated POWER.md subset) but not the others. Use the lockstep grep CLAUDE.md documents to check counts across all five files, not just the one the diff touched
  • Manifest version lockstep drift - per CLAUDE.md's "Manifest version lockstep": .claude-plugin/plugin.json, .claude-plugin/marketplace.json, cursor/.cursor-plugin/plugin.json, and .cursor-plugin/marketplace.json bumped out of sync
  • Config-key drift - a JSON config example (servers vs mcpServers, the file path, the server key name) that doesn't match what the harness in question actually reads, per the table in CLAUDE.md
  • Slash-command naming collision - a new commands/*.md file whose name is not prefixed outsystems- and would be shadowed by a host built-in (the /feedback collision CLAUDE.md documents is the known instance; the same risk applies to any new command name)
  • Tool discriminability and argument derivability - an instruction that tells the agent to call a remote MCP tool with an argument the agent has no way to obtain from prior output or the instructions themselves, or that describes two tools/flows so similarly an agent reading only the skill doc cannot pick between them
  • Instruction coherence - a skill doc or POWER.md section that contradicts another instruction in the same surface, or that still describes a tool, flag, or flow that no longer exists
  • Coordinated-surface drift - a change here that assumes a specific shape from the remote MCP server (e.g. the submit_feedback tool's argument names, or any outsystems-mcp-side tool contract) without that shape being confirmed live via tools/list or matched against the server's own repo. Report it once, naming which side needs to move

Read the full file on GitHub · 137 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 3d ago First seen · 137 lines · 48 tokens per session scan A 7772131149c4

Subscribe to this mod's changes

consistency-reviewer is an agent published in the GitHub repository OutSystems/outsystems-mcp (26 stars, last pushed today), licensed MIT. It adds 48 tokens to every session and 2,632 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-04.

Related

Other agents, from other repositories

cook-audit

You are the audit station of the jeff brigade, working one order in a fresh context. You run when plan flagged a security-relevant surface (auth, input handling, secrets, deserialization, file/network/process access, crypto, dependencies, anything privilege- or data-exposing), or when the mechanical scan floor forced…

johanthoren/jeff · 0 tokens

cook-review

You are the review station of the jeff brigade, working one order in a fresh context. You did not write this code or its tests: your independence is the point. You are the defense against momentum and self-approval bias.

johanthoren/jeff · 56 tokens

cook-refactor

You are the refactor station of the jeff brigade, working one order in a fresh context. Ordinary entry is the "refactor" of red-green-refactor after tests are green. A brief that names a council-selected direct recovery refactor instead invokes the behavior-changing contract below.

johanthoren/jeff · 39 tokens

cook-verify

You are the verify station of the jeff brigade, working one completed operation in a fresh context. You must be a different agent from the executor.

johanthoren/jeff · 38 tokens

audit-reviewer

Phase sign-off reviewer for /audit. Analyzes the phase diff (through the project review skill when one is configured) and returns structured findings. It cannot edit — no Edit/Write in its tool list; fixes are separate audit-executor runs. Spawned by the audit plugin; not meant for direct use.

AleksandarBisevac/claude-plugins · 68 tokens

audit-explorer

Read-only codebase auditor for /audit:init fan-out. Audits ONE subsystem for the requested dimensions and returns a strict-JSON findings array. Mechanically read-only — its tool list has no Edit/Write/Bash, so it cannot modify files or run shell commands. Spawned by the audit plugin; not meant for direct use.

AleksandarBisevac/claude-plugins · 72 tokens