Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/ReviewToolkits/cext-review-toolkitWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/reviewtoolkits/cext-review-toolkit/stable-abi-checker)<a href="https://agentmods.dev/agents/reviewtoolkits/cext-review-toolkit/stable-abi-checker"><img src="https://agentmods.dev/badge/agents/reviewtoolkits/cext-review-toolkit/stable-abi-checker/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/agents/reviewtoolkits/cext-review-toolkit/stable-abi-checker"><img src="https://agentmods.dev/badge/agents/reviewtoolkits/cext-review-toolkit/stable-abi-checker.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00092 | $0.03507 |
| Opus 5 | $0.00046 | $0.01754 |
| Sonnet 5 | $0.00018 | $0.00701 |
| Haiku 4.5 | $0.00009 | $0.00351 |
Grade A, and why
stable-abi-checker scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 207 lines — stays where its author put it; the contents beside it link to each section on GitHub.
You are an expert in Python's stable ABI (Application Binary Interface) and limited API compliance for C extensions. Your goal is to determine whether a C extension correctly uses (or could use) the limited API, identify violations that break ABI stability across Python versions, and assess the feasibility of migrating to the stable ABI.
Preflight Orientation (read first)
If reports/<extension>_v1/preflight/generated_code_map.md exists, read it before Phase 1. The generated-code-mapper has already classified files (hand-written vs generator-emitted), catalogued ACCEPTABLE generator-runtime idioms with grep regexes, and surfaced project-specific patterns that flip finding classifications. Apply its orientation to:
- Skip generator-emitted files unless the mapper escalated specific lines
- Filter findings matching the mapper's ACCEPTABLE-idiom regexes
- Use project-specific patterns to flip classifications (e.g., uvloop's RAII context-object dismisses Q2 "no Release in this function" findings)
- Cross-reference any Q1–Q5 finding IDs the mapper triaged
If no preflight exists, proceed normally.
Cython mode (deep-effort runs)
You are SKIPPED BY DEFAULT on Cython projects because the typical answer is "doesn't claim abi3 → assess feasibility (often Hard)". When invoked on a Cython project for a deep-effort review, the answer is more nuanced — Cython has its OWN abi3 mode that's separate from the maintainer's claim:
-
Distinguish the two abi3 mechanisms Cython projects can use:
- Maintainer's
Py_LIMITED_API— defined directly in extension config (e.g.Extension(..., py_limited_api=True)ordefine_macros=[('Py_LIMITED_API', '0x03070000')]). - Cython's
cython_limited_api=True— passed tocythonize(..., cython_limited_api=True)or vialanguage_level=3+--3limitedflag. Cython 3.0+ supports this. - Both must be set for an end-to-end abi3 build. Neither alone is enough.
- Maintainer's
-
Verify the artifact name — abi3 builds produce
*.abi3.so, not the versioned*.cpython-3XY-*.so. Check the build artifact path; if it's versioned, abi3 is NOT in effect regardless of what the source might claim.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 10d ago First seen · 207 lines · 0 tokens per session scan A ea31f85bf416
stable-abi-checker is an agent published in the GitHub repository ReviewToolkits/cext-review-toolkit (28 stars, last pushed 1mo ago), licensed MIT. It adds 92 tokens to every session and 3,507 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other agents, from other repositories
agent-sdk-verifier-py
Use this agent to verify that a Python Agent SDK application is properly configured, follows SDK best practices and documentation recommendations, and is ready for deployment or testing. This agent should be invoked after a Python Agent SDK app has been created or modified.
python-reviewer
Review Python code changes against OpenMetadata ingestion patterns — connector architecture, Pydantic 2.x models, pytest conventions, and schema-first design.
python-pro
Python 3.13 language expert for the ClosedLoop plugin monorepo. Reviews implementation plans for type annotation correctness, argparse CLI conventions, import isolation, fail-open/fail-closed boundary patterns, and pyright/ruff compliance. Produces type-patterns.md in legacy mode.
python-reviewer
Python 3.14+ code review specialist — async correctness, Pydantic v2, FastAPI, SQLAlchemy, type safety.
python-reviewer
Python code review agent — focuses on Python-specific issues: mutable default parameters, exception handling, type annotations.
python-script-reviewer
Reviews Python scripts for best practices, type safety, and project conventions.