penetration-testing

A set of methods and checklists for authorized security testing, in which you look for ways software or systems could be attacked and document the results.

In plain words
What is it for?
Use it to plan black-box, gray-box, white-box, or red-team assessments and review issues such as access-control failures, injection, weak encryption, and outdated dependencies.
Why use it?
It provides an ordered process for reconnaissance, scanning, exploitation, follow-up checks, and reporting, reducing the chance of missing common weaknesses.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/saolalab/clawforce/penetration-testing
Any agent
npx skills add saolalab/clawforce --skill penetration-testing
Clone the repo
git clone --depth 1 https://github.com/saolalab/clawforce

Made for: Claude Code, Codex.

Per session 23 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 816 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00023 $0.00816
Opus 5 $0.00012 $0.00408
Sonnet 5 $0.00005 $0.00163
Haiku 4.5 $0.00002 $0.00082

Measured 3d ago against content hash 706bc3c9e95a, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

penetration-testing scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

marketplace/roles/security-engineer/workspace/skills/penetration-testing/SKILL.md · 129 lines

How it starts

The opening of the file, as written. The whole thing — 129 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Penetration Testing

Testing Methodology

Reconnaissance → Scanning → Exploitation → Post-Exploitation → Reporting

Assessment Types

Type Scope Knowledge Duration
Black Box External No info 1-2 weeks
Gray Box Internal Partial 1-2 weeks
White Box Full Complete 2-4 weeks
Red Team Full org Limited 2-4 weeks

OWASP Top 10 Testing Checklist

1. Broken Access Control

  • IDOR (Insecure Direct Object References)
  • Privilege escalation (vertical/horizontal)
  • Missing function-level access control
  • JWT/token manipulation

2. Cryptographic Failures

  • Weak encryption algorithms
  • Sensitive data in transit (HTTPS)
  • Sensitive data at rest
  • Hardcoded secrets

3. Injection

  • SQL injection
  • NoSQL injection
  • OS command injection
  • LDAP injection

4. Insecure Design

  • Business logic flaws
  • Missing rate limiting
  • Insufficient anti-automation

5. Security Misconfiguration

  • Default credentials
  • Unnecessary features enabled
  • Verbose error messages
  • Missing security headers

6. Vulnerable Components

  • Outdated dependencies
  • Known CVEs in stack
  • Unmaintained libraries

7. Authentication Failures

  • Brute force protection
  • Session management
  • Password policies
  • MFA implementation

8. Data Integrity Failures

  • Deserialization attacks
  • CI/CD pipeline security
  • Software supply chain

9. Logging & Monitoring

  • Sensitive data in logs
  • Insufficient logging
  • Log injection

10. SSRF

  • Internal resource access
  • Cloud metadata access
  • Protocol smuggling

Pentest Report Template

## Penetration Test Report

### Executive Summary
- **Engagement**: [Type and scope]
- **Date**: [Testing period]
- **Overall Risk**: [Critical/High/Medium/Low]
- **Key Findings**: [Count by severity]

### Scope
- **In Scope**: [Systems, applications, networks]
- **Out of Scope**: [Exclusions]
- **Testing Approach**: [Methodology used]

### Findings Summary

| # | Finding | Severity | Status |
|---|---------|----------|--------|
| 1 | [Title] | Critical | Open |

### Detailed Findings

#### Finding 1: [Title]
- **Severity**: [Rating]
- **CVSS**: [Score]
- **Location**: [Affected component]
- **Description**: [Technical details]
- **Proof of Concept**: [Steps/screenshots]
- **Impact**: [Business impact]
- **Remediation**: [Recommended fix]

### Positive Observations
- [Security controls that worked well]

### Recommendations
- [Strategic security improvements]

Read the full file on GitHub · 129 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 3d ago First seen · 129 lines · 23 tokens per session scan A 706bc3c9e95a

Subscribe to this mod's changes

penetration-testing is a skill published in the GitHub repository saolalab/clawforce (38 stars, last pushed 4mo ago), licensed Apache-2.0. It adds 23 tokens to every session and 816 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

agenticmail

🎀 AgenticMail — Full email, SMS, storage & multi-agent coordination for AI agents. 63 tools.

InternLM/WildClawBench · 28 tokens

ai-meeting-scheduling

Booking links fail for groups. SkipUp schedules meetings with 2-50 participants via email — one API call coordinates across timezones automatically. Also: check status, pause, resume, or cancel requests. Async only — does not instant-book, access calendars, or do free/busy lookups.

InternLM/WildClawBench · 66 tokens

agentic-paper-digest-skill

Fetches and summarizes recent arXiv and Hugging Face papers with Agentic Paper Digest. Use when the user wants a paper digest, a JSON feed of recent papers, or to run the arXiv/HF pipeline.

InternLM/WildClawBench · 54 tokens

eachlabs-voice-audio

Text-to-speech, speech-to-text, voice conversion, and audio processing using EachLabs AI models. Supports ElevenLabs TTS, Whisper transcription with diarization, and RVC voice conversion. Use when the user needs TTS, transcription, or voice conversion.

InternLM/WildClawBench · 60 tokens

arxiv-summarizer-orchestrator

End-to-end orchestration skill for periodic arXiv collection and reporting using three sub-skills: arxiv-search-collector, arxiv-paper-processor, and arxiv-batch-reporter. Supports manual language control across all markdown outputs and Stage-B processing strategy (subagentparallel default max 5, or serial).

InternLM/WildClawBench · 75 tokens

academic-literature-search

这是一个专注于学术文献检索的专业工具,集成了多个权威学术数据库,提供全面、快速、准确的文献检索服务。支持多数据库并发检索、高级过滤、智能排序和多种输出格式。.

InternLM/WildClawBench · 0 tokens