network-troubleshooting

network-troubleshooting is a skill for Claude Code from arjunprabhulal/devops-skills. It costs 128 tokens per session (1,462 once invoked), scanned A, original, MIT.

A guide to diagnosing network connection failures one layer at a time, from name lookup through routing, encryption, HTTP, and the application.

In plain words
What is it for?
Use it to investigate timeouts and failed connections with tools such as `dig`, `curl`, `openssl`, `traceroute`, `mtr`, `tcpdump`, and `ss`.
Why use it?
It prevents wasted effort on application code when the actual problem is DNS, routing, a closed port, or a TLS certificate.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin.

Part of the devops-skills plugin — 56 skills shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/arjunprabhulal/devops-skills/network-troubleshooting
Any agent
npx skills add arjunprabhulal/devops-skills --skill network-troubleshooting
Clone the repo
git clone --depth 1 https://github.com/arjunprabhulal/devops-skills

Made for: Claude Code.

Or install devops-skills, the plugin that ships this one along with the rest of its 56 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for network-troubleshooting

README.md
[![agentmods](https://agentmods.dev/badge/skills/arjunprabhulal/devops-skills/network-troubleshooting.svg)](https://agentmods.dev/skills/arjunprabhulal/devops-skills/network-troubleshooting)
Your own site
<a href="https://agentmods.dev/skills/arjunprabhulal/devops-skills/network-troubleshooting"><img src="https://agentmods.dev/badge/skills/arjunprabhulal/devops-skills/network-troubleshooting.svg" alt="Measured on agentmods" height="20"></a>
Per session 128 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,462 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 1 finding. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00128 $0.01462
Opus 5 $0.00064 $0.00731
Sonnet 5 $0.00026 $0.00292
Haiku 4.5 $0.00013 $0.00146

Measured 6d ago against content hash 891a59b43299, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

network-troubleshooting scanned grade A with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Makes network callslowCapability

Not a fault in itself. Listed so you know the mod talks to something, and to what.

description: Covers diagnosing connectivity failures methodically, layer by layer, with the right tool per symptom — dig/nslookup for DNS, curl/openssl for TLS and HTTP, traceroute/mtr for routing, tcpdump for packet cap
skills/networking/network-troubleshooting/SKILL.md · 127 lines

How it starts

The opening of the file, as written. The whole thing — 127 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Network Troubleshooting

"It's not connecting" describes a symptom, not a cause, and the cause could be any of six or seven layers between the client and the application. The single biggest time-waster in network troubleshooting is guessing at the application layer first because it's the most familiar, when the fault is actually two layers down in DNS or routing. Working through the stack in order — resolution, then reachability, then the port, then the protocol, then the application — turns a guessing session into a bounded diagnosis.

Isolate the layer before you fix anything; a fix aimed at the wrong layer just adds noise to the next person's investigation.

For a symptom-driven command cookbook (dig, curl, traceroute, ss, tcpdump), read references/diagnostic-commands.md.

1. Confirm DNS resolves to what you expect before touching anything else

If the name doesn't resolve, or resolves to the wrong address, nothing past this point matters. dig gives more control and cleaner output than nslookup for this — use it against both the system resolver and the authoritative server directly to separate a propagation issue from a wrong record.

dig +short api.example.com                # what the system resolver returns
dig +short api.example.com @8.8.8.8       # bypass local cache, check a public resolver
dig api.example.com NS                    # who is authoritative
  • A mismatch between the two queries above means propagation lag or a stale local cache, not a wrong record — see dns-management for TTL behavior.
  • NXDOMAIN from the authoritative server itself means the record genuinely doesn't exist; stop looking downstream and fix the zone.

Done when: the resolved address is confirmed correct from both the client's resolver and the authoritative source, or the mismatch is understood and explained.

2. Confirm reachability and the path before assuming a firewall

Once the address is right, check whether packets can get there at all, and where they stop if not. traceroute (or mtr for a continuously updating view with loss percentages) shows each hop; a consistent stop at the same hop across repeated runs points at that hop specifically, not at the destination.

Read the full file on GitHub · 127 lines

Files

What ships with it

1 file beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 127 lines · 128 tokens per session scan A 891a59b43299

Subscribe to this mod's changes

network-troubleshooting is a skill published in the GitHub repository arjunprabhulal/devops-skills (3 stars, last pushed 11d ago), licensed MIT. It adds 128 tokens to every session and 1,462 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 1 finding (makes network calls). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

gke-ai-troubleshooting-jobset-interruption

Diagnoses GKE JobSet interruptions, restarts, and preemptions for AI/ML training workloads autonomously. Use when troubleshooting JobSet restart loops, spot VM preemptions, node readiness failures, host VM issues, or coordinator worker crashes. Don't use for general GKE cluster creation, basic workload deployment, or…

google/skills · 83 tokens

gke-workload-troubleshooting

Diagnoses GKE workload failures (CrashLoopBackOff, OOMKilled, ImagePullBackOff, Pending, etc.) via logs and events. Use when pods fail to start or crash repeatedly. Don't use for GKE cluster infrastructure provisioning, node pool creation, or non-Kubernetes Google Cloud services.

google/skills · 69 tokens

gke-node-notready

Diagnoses GKE nodes reporting NotReady or Unknown status by inspecting node conditions, events, kubelet/containerd logs, and node metrics, then proposing safe remediations. Use when nodes show NotReady, when the kubelet stops posting node status, or when workloads are evicted or stuck Pending due to node health. Don't…

google/skills · 112 tokens

gke-ai-troubleshooting-tpu-dynamic-slices-monitoring

Monitors, troubleshoots, and manages GKE TPU Dynamic Slices custom resources. Use when checking TPU slice lifecycle states, troubleshooting slice provisioning failures, validating single-slice or multi-slice (JobSet) workload manifests, or safely patching stuck finalizers and disabling the slice controller. Don't use…

google/skills · 107 tokens

datalineage-summary

Summarizes Google Cloud Data Lineage graphs to help users debug data quality issues and understand data provenance for BQ/GCS. Use when summarizing upstream and downstream data flows, and presenting complex lineage data as an intuitive Markdown report. Don't use for generic BigQuery queries, editing lineage…

google/skills · 92 tokens

gke-ai-troubleshooting-handle-disruption-gpu-tpu

Diagnoses, predicts, and mitigates node disruptions during Compute Engine host maintenance and hardware or software maintenance events for GPU and TPU workloads on GKE. Use when diagnosing node disruptions, predicting host maintenance events on GPU/TPU nodepools, inspecting node interruption PromQL metrics, auditing…

google/skills · 117 tokens