Use this skill when building applications with Gemini models, Gemini API, working with multimodal content (text, images, audio, video), implementing function calling, using structured outputs, or needing current model specifications. Covers SDK usage (google-genai for Python, @google/genai for JavaScript/TypeScript…
Complete reference for generating and editing images with Gemini's Nano Banana models (Gemini 3 Pro Image, Gemini 3.1 Flash Image, Gemini 3.1 Flash Lite Image, legacy Gemini 2.5 Flash Image) via the Interactions API. ALWAYS trigger when the user wants to generate, edit, remix, or upscale images with Gemini, mentions…
Use this skill when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses, background research tasks, function calling, structured output, or migrating from the old generateContent API. This skill covers the…
Complete reference for building with the OpenAI API — Responses API (the recommended primitive), Chat Completions, text generation/prompting, vision/image understanding, GPT Image generation, audio/speech (realtime + request-based), Structured Outputs (JSON schema), tools (web search, file search, function calling…
Use when the user asks how to build with OpenAI products or APIs and needs up-to-date official documentation with citations (for example: Codex, Responses API, Chat Completions, Apps SDK, Agents SDK, Realtime, model capabilities or limits); prioritize OpenAI docs MCP tools and restrict any fallback browsing to…
Complete reference for generating and editing images with the OpenAI API using GPT Image models (gpt-image-2, gpt-image-1.5, gpt-image-1, gpt-image-1-mini). ALWAYS trigger this skill when the user wants to generate images with OpenAI, mentions "gpt-image", "GPT Image", DALL-E replacement, image generation via the…