ai-models-gateway

ai-models-gateway is a skill for Claude Code from iofold/ainative-claude-plugins. It costs 131 tokens per session (1,754 once invoked), scanned A, original, MIT.

A guide for choosing AI models and connecting OpenAI, Google, and Anthropic models through Cloudflare AI Gateway. An AI gateway is a shared endpoint that can route requests to different model providers.

In plain words
What is it for?
Use it to compare current model options and prices, choose a model, configure a shared OpenAI-compatible endpoint, or set up multi-provider routing and fallbacks.
Why use it?
It helps match a model to needs such as price, speed, context length, coding, or reasoning without handling every provider separately. The guide also documents provider-specific settings and limits.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin. Also seen: positional $N argument.

Needs its repository: it runs a file that does not travel with it, so clone the repository first. The line is bun scripts/lib/openai-helper.ts "Your prompt".

Part of the ai-development plugin — 3 skills shipped together

Good fit Use it to compare current model options and prices, choose a model, configure a shared OpenAI-compatible endpoint, or set up multi-provider routing and fallbacks.

Compare 6 skills from other repositories ↓
Install

Getting it into your agent

It runs from inside its repository, so the clone comes first — what it calls does not travel with the file alone.

Clone the repo
git clone --depth 1 https://github.com/iofold/ainative-claude-plugins
agentmods
npx agentmods add skills/iofold/ainative-claude-plugins/ai-models-gateway

Made for: Claude Code.

Or install ai-development, the plugin that ships this one along with the rest of its 3 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for ai-models-gateway

README.md
[![agentmods](https://agentmods.dev/badge/skills/iofold/ainative-claude-plugins/ai-models-gateway.svg)](https://agentmods.dev/skills/iofold/ainative-claude-plugins/ai-models-gateway)
Your own site
<a href="https://agentmods.dev/skills/iofold/ainative-claude-plugins/ai-models-gateway"><img src="https://agentmods.dev/badge/skills/iofold/ainative-claude-plugins/ai-models-gateway.svg" alt="Measured on agentmods" height="20"></a>
Per session 131 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,754 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00131 $0.01754
Opus 5 $0.00066 $0.00877
Sonnet 5 $0.00026 $0.00351
Haiku 4.5 $0.00013 $0.00175

Measured 8d ago against content hash a3c6700a1710, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-08, from the pricing page.

Security

Grade A, and why

ai-models-gateway scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.

The scan reads SKILL.md. This mod also ships 1 executable file (references/models.ts), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

ai-development/skills/ai-models-gateway/SKILL.md · 151 lines

How it starts

The opening of the file, as written. The whole thing — 151 lines — stays where its author put it; the contents beside it link to each section on GitHub.

AI Models & Cloudflare AI Gateway

Overview

This skill provides current (2025-2026) AI model information, pricing comparisons, and guidance on using Cloudflare AI Gateway as a unified endpoint for OpenAI, Google, and Anthropic.

Helper script: scripts/lib/openai-helper.ts - Lightweight OpenAI-compatible client for hooks/scripts

Reference Files

For detailed information, consult:

  • references/models.ts - TypeScript definitions with all model IDs, gateway strings, pricing, and helper functions
  • references/api-parameters.md - Detailed API parameters, constraints, and provider-specific nuances

Quick Model Selection

Use Case Recommended Model Price (Input/MTok) Context
Cheapest gpt-5-nano $0.05 400K
Budget + 1M context gemini-2.5-flash-lite $0.10 1M
Balanced gpt-5-mini, gemini-2.5-flash $0.25-0.30 400K-1M
Flagship gpt-5.4, gemini-3.1-pro-preview $2.00-2.50 1M+
Code/agentic gpt-5, claude-sonnet-4-6 $1.25-3.00 200K-400K
Maximum intelligence gpt-5.4-pro, claude-opus-4-6 $5.00-30.00 200K-1M
Long context (1M+) gemini-2.5-flash, gpt-5.4 $0.30-2.50 1M+
Image generation gemini-3.1-flash-image-preview (Nano Banana 2) $0.50 1M
Image gen (max quality) gemini-3-pro-image-preview (Nano Banana Pro) $2.00 1M

Model Selection Decision Tree

Is cost the primary concern?
├── Yes → gpt-5-nano ($0.05) or gemini-2.5-flash-lite ($0.10, 1M context)
│
├── Need long context (>400K)?
│   ├── Budget → gemini-2.5-flash (1M, $0.30)
│   └── Quality → gpt-5.4 (1.05M, $2.50)
│
├── Need maximum intelligence?
│   ├── OpenAI → gpt-5.4-pro ($30, xhigh reasoning)
│   └── Anthropic → claude-opus-4-6 ($5)
│
├── Code/agentic tasks?
│   └── gpt-5 ($1.25) or claude-sonnet-4-6 ($3)
│
└── Balanced quality/cost?
    └── gpt-5-mini ($0.25) or gemini-3-flash-preview ($0.50)

Read the full file on GitHub · 151 lines

Files

What ships with it

2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 8d ago First seen · 151 lines · 131 tokens per session scan A a3c6700a1710

Subscribe to this mod's changes

ai-models-gateway is a skill published in the GitHub repository iofold/ainative-claude-plugins (3 stars, last pushed 20d ago), licensed MIT. It adds 131 tokens to every session and 1,754 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

gemini-api-agent-platform

Guides the usage of the Gemini API on Agent Platform with the Google Gen AI SDK for enterprise AI applications. Covers SDK usage (Python, JS/TS, Go, Java, C#), capabilities like Live API, tools, multimedia generation, caching, and batch prediction.

davila7/claude-code-templates · 61 tokens

open-source

Documentation reference for writing Python code using the browser-use open-source library. Use this skill whenever the user needs help with Agent, Browser, or Tools configuration, is writing code that imports from browseruse, asks about @sandbox deployment, supported LLM models, Actor API, custom tools, lifecycle…

browser-use/browser-use · 137 tokens

gemini-api-dev

Use this skill when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses, background research tasks, function calling, structured output, or migrating from the old generateContent API. Covers SDK usage and best…

google-gemini/gemini-skills · 73 tokens

deepstream-sop

Use this skill when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether operators perform assembly-line steps in order via event boundary detection (GEBD) plus VLM classification. Trigger even if the…

NVIDIA/skills · 219 tokens

azure-search-documents-dotnet

Azure AI Search SDK for .NET (Azure.Search.Documents). Use for building search applications with full-text, vector, semantic, and hybrid search. Covers SearchClient (queries, document CRUD), SearchIndexClient (index management), and SearchIndexerClient (indexers, skillsets). Triggers: "Azure Search .NET"…

microsoft/skills · 102 tokens

azure-search-documents-ts

Build search applications using Azure AI Search SDK for JavaScript (@azure/search-documents). Use when creating/managing indexes, implementing vector/hybrid search, semantic ranking, or building agentic retrieval with knowledge bases.

microsoft/skills · 48 tokens