Platform SRE for Kubernetes

Platform SRE for Kubernetes is an agent for Claude Code from github/awesome-copilot. It costs 32 tokens per session (958 once invoked), scanned A, original, MIT.

A Kubernetes-focused site reliability assistant. Kubernetes is a system for running and managing containerized applications, while site reliability engineering focuses on keeping those systems dependable.

In plain words
What is it for?
Use it to plan and verify Kubernetes deployments, prepare manifests, check changes before rollout, manage releases and rollbacks, and assess operational risks.
Why use it?
It helps reduce deployment risk by requiring planning, validation, monitoring, safe rollbacks, and security settings for production changes.

Agent for Claude Code

Written for Claude Code: a Claude Code subagent (agents/*.md).

About the project

Awesome GitHub Copilot is a community collection of custom agents, instructions, skills, hooks, workflows, plugins, and configuration for GitHub Copilot. It helps Copilot users customize coding and development tasks. Catalogue entries are individual Copilot add-ons from this collection.

github/awesome-copilot · 38,668 stars · on GitHub

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/github/awesome-copilot/platform-sre-kubernetes
Clone the repo
git clone --depth 1 https://github.com/github/awesome-copilot

Made for: Claude Code.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for Platform SRE for Kubernetes

README.md
[![agentmods](https://agentmods.dev/badge/agents/github/awesome-copilot/platform-sre-kubernetes.svg)](https://agentmods.dev/agents/github/awesome-copilot/platform-sre-kubernetes)
Your own site
<a href="https://agentmods.dev/agents/github/awesome-copilot/platform-sre-kubernetes"><img src="https://agentmods.dev/badge/agents/github/awesome-copilot/platform-sre-kubernetes.svg" alt="Measured on agentmods" height="20"></a>
Per session 32 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 958 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00032 $0.00958
Opus 5 $0.00016 $0.00479
Sonnet 5 $0.00006 $0.00192
Haiku 4.5 $0.00003 $0.00096

Measured 2d ago against content hash ce7da8d73aaf, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

Platform SRE for Kubernetes scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

Copies of this mod

2 near-identical copies found in the catalogue:

agents/platform-sre-kubernetes.agent.md · 117 lines

How it starts

The opening of the file, as written. The whole thing — 117 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Platform SRE for Kubernetes

You are a Site Reliability Engineer specializing in Kubernetes deployments with a focus on production reliability, safe rollout/rollback procedures, security defaults, and operational verification.

Your Mission

Build and maintain production-grade Kubernetes deployments that prioritize reliability, observability, and safe change management. Every change should be reversible, monitored, and verified.

Clarifying Questions Checklist

Before making any changes, gather critical context:

Environment & Context

  • Target environment (dev, staging, production) and SLOs/SLAs
  • Kubernetes distribution (EKS, GKE, AKS, on-prem) and version
  • Deployment strategy (GitOps vs imperative, CI/CD pipeline)
  • Resource organization (namespaces, quotas, network policies)
  • Dependencies (databases, APIs, service mesh, ingress controller)

Output Format Standards

Every change must include:

  1. Plan: Change summary, risk assessment, blast radius, prerequisites
  2. Changes: Well-documented manifests with security contexts, resource limits, probes
  3. Validation: Pre-deployment validation (kubectl dry-run, kubeconform, helm template)
  4. Rollout: Step-by-step deployment with monitoring
  5. Rollback: Immediate rollback procedure
  6. Observability: Post-deployment verification metrics

Security Defaults (Non-Negotiable)

Always enforce:

  • runAsNonRoot: true with specific user ID
  • readOnlyRootFilesystem: true with tmpfs mounts
  • allowPrivilegeEscalation: false
  • Drop all capabilities, add only what's needed
  • seccompProfile: RuntimeDefault

Resource Management

Define for all containers:

  • Requests: Guaranteed minimum (for scheduling)
  • Limits: Hard maximum (prevents resource exhaustion)
  • Aim for QoS class: Guaranteed (requests == limits) or Burstable

Health Probes

Implement all three:

  • Liveness: Restart unhealthy containers
  • Readiness: Remove from load balancer when not ready
  • Startup: Protect slow-starting apps (failureThreshold × periodSeconds = max startup time)

Read the full file on GitHub · 117 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 117 lines · 32 tokens per session scan A ce7da8d73aaf

Subscribe to this mod's changes

Platform SRE for Kubernetes is an agent published in the GitHub repository github/awesome-copilot (38,668 stars, last pushed today), licensed MIT. It adds 32 tokens to every session and 958 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.

Related

Other agents, from other repositories

infrastructure

Cloud infrastructure, Kubernetes, orchestration, and infrastructure as code. Use for cloud platforms, containerization, service mesh, and infrastructure design.

AgentWorkforce/relay · 31 tokens

helm-deployment

Author and maintain Helm charts, multi-env config, digest-based deploys, and rollback-safe delivery across localdev, staging, and production.

atretyak1985/swarmery · 32 tokens

deployment-verifier

Verifies local deployment health — checks ports, starts app, polls health endpoint, inspects Docker containers.

asysta-act/agent-flow · 24 tokens

Kubernetes Workload Optimizer

Tunes container resource requests/limits AND node-level autoscaling (Karpenter, Cluster Autoscaler) for the right balance of cost, scheduling latency, and pod stability. Covers VPA-driven rightsizing and consolidation policy in one discipline.

Cletrics/finops-agents · 54 tokens

Kubernetes FinOps Engineer

Specialist in Kubernetes cost allocation, namespace and label-based chargeback, and cluster-level optimization. Comfortable with OpenCost, Kubecost, Karpenter, cluster autoscaler, and vertical pod autoscaler.

Cletrics/finops-agents · 48 tokens

hpc-platform-architect

Expert in designing centralized High-Performance Computing (HPC) platforms for modern vehicles. Specializes in hypervisor selection, AUTOSAR Adaptive integration, resource allocation, safety partitioning, and migration strategies from distributed ECU architectures to centralized compute.

birol91/quorum-agents · 54 tokens