Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add vinayaklatthe/microsoft-security-skills --skill azure-security-benchmarkgit clone --depth 1 https://github.com/vinayaklatthe/microsoft-security-skillsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/vinayaklatthe/microsoft-security-skills/azure-security-benchmark)<a href="https://agentmods.dev/skills/vinayaklatthe/microsoft-security-skills/azure-security-benchmark"><img src="https://agentmods.dev/badge/skills/vinayaklatthe/microsoft-security-skills/azure-security-benchmark/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/vinayaklatthe/microsoft-security-skills/azure-security-benchmark"><img src="https://agentmods.dev/badge/skills/vinayaklatthe/microsoft-security-skills/azure-security-benchmark.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00092 | $0.01870 |
| Opus 5 | $0.00046 | $0.00935 |
| Sonnet 5 | $0.00018 | $0.00374 |
| Haiku 4.5 | $0.00009 | $0.00187 |
Grade A, and why
azure-security-benchmark scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 123 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Microsoft Cloud Security Benchmark (MCSB)
The Microsoft Cloud Security Benchmark (MCSB, successor to the Azure Security Benchmark) is Microsoft's canonical, prescriptive set of cloud security best practices, with controls mapped to CIS, NIST SP 800-53, PCI DSS, and ISO 27001, and monitored natively in Microsoft Defender for Cloud. MCSB is the default initiative in Defender for Cloud - you are already being scored against it.
When to use
Use this skill when the user wants a recognised, framework-mapped cloud security baseline, needs to translate compliance obligations into concrete Azure controls, or is trying to prioritise where to start with Defender for Cloud recommendations.
Do not use this skill for:
- Tactical Defender for Cloud remediation work (use
defender-for-cloud-hardening) - Azure Policy authoring details (use
azure-policy) - Workload-specific threat modelling (use
threat-modelling)
Pick the right MCSB control domain first
MCSB has 12 control domains. Do not try to fix them in parallel - pick by risk + dependency:
| Priority | Domain | Why this order | Typical first action |
|---|---|---|---|
| P0 | Identity Management + Privileged Access | Identity is the new perimeter; admin compromise = game over | Enforce MFA, remove standing Global Admin via PIM |
| P0 | Logging & Threat Detection | Cannot respond without visibility | Enable diagnostic settings on subscriptions + Sentinel/Defender |
| P1 | Network Security | Reduces blast radius of compromised identities | Default-deny NSGs, private endpoints for PaaS |
| P1 | Data Protection | Encryption, key management, classification | Customer-managed keys for crown-jewel data, sensitivity labels |
| P1 | Asset Management | Cannot protect what you cannot see | Tag policy + resource inventory |
| P2 | Posture & Vulnerability Management | Continuous improvement loop | Onboard Defender for Servers, weekly Secure Score review |
| P2 | Endpoint Security + Backup & Recovery | Resilience baseline | Defender for Endpoint + Azure Backup with immutable vault |
| P2 | Incident Response | Process maturity | Runbook + tabletop with Defender XDR / Sentinel |
| P3 | DevOps Security | Shift-left controls | GitHub Advanced Security or Azure DevOps with Defender for DevOps |
| P3 | Governance & Strategy | Sustains the others | Azure Policy + management group structure |
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 11d ago First seen · 123 lines · 92 tokens per session scan A 78722875ea27
azure-security-benchmark is a skill published in the GitHub repository vinayaklatthe/microsoft-security-skills (173 stars, last pushed 2mo ago), licensed MIT. It adds 92 tokens to every session and 1,870 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
power-automate
Use when operating Microsoft Power Automate cloud flows from code — create, enable, update, list or delete via the Dataverse Web API (workflow table, category 5) with Entra ID OAuth2, plus run-history debugging. NOT designing the flow definition (that is automation-flows), NOT picking a platform by billing model (that…
chronicle
Analyze Copilot session history for standup reports, usage tips, session search, and session reindexing. Use when the user asks for a standup, daily summary, usage tips, workflow recommendations, wants to search or find past sessions by keyword/file/PR, wants to reindex their session store, or asks about deleting…
generate-run-commands
Generate or modify run commands for the current session. Use when the user wants to set up or update run commands that appear in the session's Run button.
html-ppt-hermes-cyber-terminal
OpenDesign + BYOK: choosing and wiring your own model, hands-on — cost, quality, and the routing decision. Built as a decision-grade AI literacy deck for engineers, IT, applied-AI teams.
skill-writing-plans
Create zero-context implementation plans with bite-sized tasks — use for multi-step feature planning.
bf-to-agents-sdk-dotnet-migration
Use when migrating a Bot Framework .NET SDK bot to Microsoft 365 Agents SDK. Triggered by projects that depend on packages: Microsoft.Bot.Builder or Microsoft.Bot.Builder.Integration.AspNet.Core that want to migrate to Agents SDK.