fleet

fleet is a skill for Claude Code, Codex from nebius/nebius-physical-ai. It costs 78 tokens per session (8,870 once invoked), scanned C, original, Apache-2.0.

A deployment and operations workflow for fleets of Nebius Managed Kubernetes clusters, which are groups of machines managed through Kubernetes. It uses one declarative YAML specification to describe clusters across projects in a tenant.

In plain words
What is it for?
Use it to plan, deploy, destroy, and inspect multiple GPU training clusters, including clusters with strict reserved-capacity requirements and optional RTX PRO 6000 MIG partitioning.
Why use it?
It removes the need to set up and manage each training cluster separately. The same specification can describe identical or customized clusters and can create projects when needed.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit Use it to plan, deploy, destroy, and inspect multiple GPU training clusters, including clusters with strict reserved-capacity requirements and optional RTX PRO 6000 MIG partitioning.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/nebius/nebius-physical-ai/fleet
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add nebius/nebius-physical-ai --skill fleet
Clone the repo
git clone --depth 1 https://github.com/nebius/nebius-physical-ai

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for fleet

README.md
[![agentmods](https://agentmods.dev/badge/skills/nebius/nebius-physical-ai/fleet.svg)](https://agentmods.dev/skills/nebius/nebius-physical-ai/fleet)
Your own site
<a href="https://agentmods.dev/skills/nebius/nebius-physical-ai/fleet"><img src="https://agentmods.dev/badge/skills/nebius/nebius-physical-ai/fleet.svg" alt="Measured on agentmods" height="20"></a>
Per session 78 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 8,870 The whole file, excluding the scripts and references it only reads on demand.
Security scan C 1 finding. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00078 $0.08870
Opus 5 $0.00039 $0.04435
Sonnet 5 $0.00016 $0.01774
Haiku 4.5 $0.00008 $0.00887

Measured today against content hash 831c5a9daac1, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-07, from the pricing page.

Security

Grade C, and why

fleet scanned grade C with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured today.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Reaches for credential fileshighPrivilege escalation

SSH keys, cloud credentials, git-credentials, .npmrc, /etc/shadow: reading these is how a config file becomes a credential leak.

`ssh_public_key` or `~/.ssh/id_ed25519.pub` / `id_rsa.pub`.
skills/tools/fleet/SKILL.md · 636 lines

How it starts

The opening of the file, as written. The whole thing — 636 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Fleet (multi-cluster, multi-project Kubernetes)

When To Use

Use when a customer wants more than one Managed Kubernetes cluster stood up across one or many projects in a single Nebius tenant — e.g. per-team or per-tenant training clusters — and wants npa to drive it from one declarative file. npa fleet wraps the public nebius/nebius-solutions-library k8s-training recipe (the same recipe npa cluster up uses) once per cluster, and can create the target projects on demand.

For a single cluster, prefer npa cluster up. For Slurm-on-Kubernetes, use npa soperator.

Three-tier contract:

  • CLI: npa fleet plan|deploy|destroy|status|verify-storage|verify-graphics --spec <fleet.yaml>.
  • SDK: npa.sdk.fleet.deploy(spec) / destroy / plan / status with FleetSpec / ProjectSpec / ClusterSpec / NodePoolSpec.
  • YAML / agent: apiVersion: npa.fleet/v0.0.1 spec; workflow toolRef: infra.fleet.deploy (config key fleet_spec).

Verify an existing shared filesystem

Run npa fleet verify-storage --spec <private-fleet.yaml> --output json to qualify every CPU and GPU worker. The shared implementation is npa.fleet.storage_verification.verify_storage, also exported as npa.sdk.fleet.verify_storage. Use --only-projects, --only-clusters, --project-prefix, and --profile to preserve Fleet selection and identity semantics. Unknown selectors and missing or stale registered identity fail closed; explicitly disabled filesystems are skipped.

The verifier proves the exact read-write virtiofs source/path, reboot-safe nofail entry, capacity in binary GiB, unique host-file checksums, CSI health, and one RWX PVC shared across pods pinned to every exact worker. Every pod checks every worker's unique payload. It reads no pre-existing customer entries. Cleanup removes only owned probe paths and temporary resources with identity labels and UID preconditions, then proves absence on every node using server-synchronized Linux statx attributes after owned writers stop. Cached positive directory entries cannot substitute for fresh link-count evidence; unsupported synchronization and persistent linked entries fail closed. Partial evidence or cleanup failure cannot pass. Do not replace this with the vendored single-node shell smoke or infer storage health from node readiness.

Read the full file on GitHub · 636 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. today Changed · +244 lines 831c5a9daac1
  2. 4d ago Changed · +8 lines 202ea58712d1
  3. 8d ago First seen · 384 lines · 78 tokens per session scan C 3d2f6e3fc325

Subscribe to this mod's changes

fleet is a skill published in the GitHub repository nebius/nebius-physical-ai (27 stars, last pushed today), licensed Apache-2.0. It adds 78 tokens to every session and 8,870 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it C with 1 finding (reaches for credential files). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

azd-deployment

Deploy containerized frontend + backend applications to Azure Container Apps with remote builds, managed identity, and idempotent infrastructure.

sickn33/agentic-awesome-skills · 29 tokens

openshell-cli

Guide agents through using the OpenShell CLI (openshell) for sandbox management, gateway registration, provider configuration and refresh, policy iteration, settings, service exposure, BYOC workflows, and inference routing. Covers basic through advanced multi-step workflows. Trigger keywords - openshell, sandbox…

NVIDIA/OpenShell · 127 tokens

langbot-deploy

Deploy and configure a LangBot instance — Docker / Docker Compose, Kubernetes, the config.yaml model, the Box sandbox runtime, the plugin runtime, and the global API key. Use when installing, deploying, upgrading, or configuring LangBot in production or self-hosted environments. Triggers on "deploy langbot", "langbot…

langbot-app/LangBot · 104 tokens

compute-env-setup

Set up a compute environment on a remote provider so Claude Science jobs can run there. Covers direct SSH/conda hosts, Slurm clusters, container-via-bridge runners, and managed-API providers (Modal, GCP, RunPod). Use when standing up a new provider, porting an env to a different backend, adding a tool that needs its…

UnicomAI/wanwu · 134 tokens

azure-cloud-migrate

Assess and migrate cross-cloud workloads to Azure with reports and code conversion. Supports Lambda→Functions, Beanstalk/Heroku/App Engine→App Service, Fargate/Kubernetes/Cloud Run/Spring Boot→Container Apps. WHEN: migrate Lambda to Functions, AWS to Azure, migrate Beanstalk, migrate Heroku, migrate App Engine, Cloud…

microsoft/skills · 106 tokens

atmos-helmfile

Helmfile orchestration: sync/apply/destroy/diff, Kubernetes deployments, varfile generation, EKS integration, source management.

cloudposse/atmos · 33 tokens