apim_throttle_expert

apim_throttle_expert is a skill for Claude Code, Codex from aiappsgbb/awesome-gbb. It costs 48 tokens per session (538 once invoked), scanned A, original, MIT.

A troubleshooting guide for HTTP 429 errors from the AI Citadel API gateway. A 429 means the service is temporarily refusing requests because a usage or rate limit was reached.

In plain words
What is it for?
Use it to inspect gateway logs and response headers, identify the limit that triggered, and choose the matching configuration or capacity fix.
Why use it?
It helps distinguish monthly product quotas, per-key request limits, and exhaustion of the backend's token capacity.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit Use it to inspect gateway logs and response headers, identify the limit that triggered, and choose the matching configuration or capacity fix.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/aiappsgbb/awesome-gbb/apim_throttle_expert
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add aiappsgbb/awesome-gbb --skill apim_throttle_expert
Clone the repo
git clone --depth 1 https://github.com/aiappsgbb/awesome-gbb

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for apim_throttle_expert

README.md
[![agentmods](https://agentmods.dev/badge/skills/aiappsgbb/awesome-gbb/apim_throttle_expert/github.svg)](https://agentmods.dev/skills/aiappsgbb/awesome-gbb/apim_throttle_expert)
Your own site
<a href="https://agentmods.dev/skills/aiappsgbb/awesome-gbb/apim_throttle_expert"><img src="https://agentmods.dev/badge/skills/aiappsgbb/awesome-gbb/apim_throttle_expert/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for apim_throttle_expert

Your own site · 80×15
<a href="https://agentmods.dev/skills/aiappsgbb/awesome-gbb/apim_throttle_expert"><img src="https://agentmods.dev/badge/skills/aiappsgbb/awesome-gbb/apim_throttle_expert.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 48 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 538 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00048 $0.00538
Opus 5 $0.00024 $0.00269
Sonnet 5 $0.00010 $0.00108
Haiku 4.5 $0.00005 $0.00054

Measured 9d ago against content hash af3eb08721d9, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-09, from the pricing page.

Security

Grade A, and why

apim_throttle_expert scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/azure-sre-agent/references/plugins/gbb-citadel/skills/apim_throttle_expert/SKILL.md · 57 lines

What it actually says

apim_throttle_expert

When to use

The SRE Agent should invoke this skill when:

  • A user reports a 429 from the Citadel gateway
  • A scheduled task flags a spike in 429 responses
  • Backend pool TPM utilization approaches its limit

Investigation flow

  1. Identify the failing operation — ask the user for the path (/llm/v1/chat/completions, /doc/analyze, etc.) and the time window.

  2. Pull APIM diagnostics from Log Analytics:

    ApiManagementGatewayLogs
    | where TimeGenerated > ago(2h)
    | where ResponseCode == 429
    | summarize count() by OperationId, ClientIpAddress, ProductId, bin(TimeGenerated, 5m)
    | order by TimeGenerated desc
    
  3. Classify the 429 by inspecting the response headers in the same log:

    Header present Cause Action
    Retry-After: <secs> + x-throttling-source: apim-product-quota Per-product monthly quota exhausted Check product config; consider raising or moving caller to a different product
    Retry-After: <secs> + x-throttling-source: rate-limit-by-key Per-key rate limit (typically per-spoke MI) Check rate-limit policy + named value
    Retry-After: <secs> + x-aoai-throttle: true Backend AOAI TPM exhaustion (not APIM) Scale up AOAI deployment or move to PTU
    No Retry-After Bug in policy — escalate Read the operation's policy XML
  4. For backend exhaustion, check the BackendPool config:

    az apim api operation policy list ...
    
  5. Output the classification, the specific policy / quota that triggered, the per-product utilization, and a recommended action.

Tools

This skill uses:

  • RunAzCliReadCommands
  • QueryLogAnalyticsByWorkspaceId

Safety

  • Never modify APIM policy, named values, or product subscriptions
  • Never read APIM subscription key contents
  • All actions must be reviewed by a human
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 9d ago First seen · 57 lines · 48 tokens per session scan A af3eb08721d9

Subscribe to this mod's changes

apim_throttle_expert is a skill published in the GitHub repository aiappsgbb/awesome-gbb (5 stars, last pushed 2d ago), licensed MIT. It adds 48 tokens to every session and 538 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

printing-press-amend

Amend a published CLI from one of two input sources: (1) dogfood mode mines the active Claude Code session transcript for friction (missing flags, hand- rolled API payloads, silent-null returns); (2) direct-input mode accepts user-supplied asks (rename a command, add commands or feeds, fix a named bug, optionally…

mvanhorn/cli-printing-press · 222 tokens

convex-performance-audit

Audits Convex performance for reads, subscriptions, write contention, and function limits. Use for slow features, insights findings, OCC conflicts, or read amplification.

openclaw/clawhub · 38 tokens

convex-insights

Query a running Convex app's logs + health in natural language (official MCP): failures, slow/expensive functions, deploy causality — scoped, evidence-backed, with a dashboard deep link.

openclaw/clawhub · 45 tokens

ssl-proxy-troubleshoot

Systematic workflow for troubleshooting SSL/proxy connectivity issues with government websites.

HKUDS/OpenSpace · 20 tokens

diagnose-backend-bug

Diagnose a bounded backend or multi-service failure from GitHub Issues, Jira, Aone, user-provided exports, logs, traces, responses, stack traces, or job records. Use when a service, API, RPC, worker, queue, CLI, or scheduled job bug needs correlation through the project's existing observability route before repair; do…

QoderAI/better-harness · 87 tokens

platform-apex-anonymous-run

Use this skill to run anonymous Apex against the connected Salesforce org — from a .apex file or a pasted snippet — capturing the debug log, surfacing compile and runtime errors, and summarizing results. Trigger on phrases like "run this anonymous apex", "execute this script against my org", "run this snippet of…

forcedotcom/sf-skills · 159 tokens