Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/internlm/wildclawbench/03_task4npx skills add InternLM/WildClawBench --skill 03_task4git clone --depth 1 https://github.com/InternLM/WildClawBenchWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00060 | $0.00703 |
| Opus 5 | $0.00030 | $0.00351 |
| Sonnet 5 | $0.00012 | $0.00141 |
| Haiku 4.5 | $0.00006 | $0.00070 |
Grade A, and why
03_task4 scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 91 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Slack Status Drafter Skill
Read recent Slack messages, reconcile the latest figures, and save a polished client-ready status report as a draft — never send directly.
Tools
All tools are defined in tmp_workspace/utils.py:
http_request— POST to any URL with an optional JSONbody; use for all Slack API callswrite_file— writecontenttopath; use to save the final report
Slack API
Base URL: http://localhost:9110
| Action | Endpoint | Required Body |
|---|---|---|
| List messages | POST /slack/messages |
{"days_back": 7, "max_results": 20} (all optional) |
| Get message | POST /slack/messages/get |
{"message_id": "<id>"} |
| Save draft | POST /slack/drafts/save |
{"to": "@recipient", "content": "..."} + optional "reply_to_message_id" |
⚠️ Draft-only task. Always use
slack_save_draft— do not callslack_send_message(POST /slack/send). The user must review before anything goes to the client.
Workflow
- List messages — fetch recent messages with
slack_list_messages - Identify relevant messages — filter for messages related to the project in question
- Read each in full — retrieve complete content via
slack_get_message - Reconcile numbers — where figures conflict or have been revised, use the most recent version; note any unresolved discrepancies
- Draft the report — write a clear, accurate, client-appropriate status update
- Save as draft — submit via
slack_save_draftfor user review before sending - Write report — save the full draft and reconciliation notes to
/tmp_workspace/results/results.md
Draft Format
Hi [Client Name],
Here's the latest status update on [Project Name]:
**Overall Status**: On track / At risk / Delayed
**Progress**
- [Area]: [current status, latest figures]
**Key Milestones**
- [Milestone]: [status, date]
**Next Steps**
- ...
Please let me know if you have any questions.
[Your Name]
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 3d ago First seen · 91 lines · 60 tokens per session scan A fb1ff1d03eac
03_task4 is a skill published in the GitHub repository InternLM/WildClawBench (516 stars, last pushed 16d ago), licensed MIT. It adds 60 tokens to every session and 703 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
agent-browser
Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a…
naga-config
Naga 自身配置管理技能。用于查看和修改 Naga 系统设置、添加 MCP 工具服务、导入自定义技能、搜索可用 MCP 工具。当用户要求修改设置、添加工具或技能时使用此技能。.
naga_control
通过 agentType: "nagacontrol" 调用,直接控制 Naga 自身的运行状态和配置。.
file-manager
文件管理技能。用于创建、移动、复制、删除文件和文件夹,整理目录结构。当用户需要管理文件、整理文件夹或批量处理文件时使用。.
live2d_controller
// 读取:system.characterbundle.loadcharacterskillsections -> system.config.buildtier1variables.characterbuiltinskillsprompt.
code-review
代码审查和质量分析技能。用于审查代码、发现潜在问题、提供改进建议。当用户请求代码审查、代码质量分析或最佳实践建议时使用。.