InternLM/WildClawBench

An in-the-wild benchmark for AI agents in the production harness.

517Stars on the repository
27Mods indexed here, across every type
22d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

self-improvement

25

InternLM/WildClawBench

Skill Claude Code

Captures learnings, errors, and corrections to enable continuous improvement. Use when: (1) A command or operation fails unexpectedly, (2) User corrects Claude ('No, that's wrong...', 'Actually...'), (3) User requests a capability that doesn't exist, (4) An external API or tool fails, (5) Claude realizes its knowledge…

not rated 517 22d ago A 102 tokens copy · 100% MIT

video-frames

26

InternLM/WildClawBench

Skill Claude CodeCodex

Extract frames or short clips from videos using ffmpeg.

not rated 517 22d ago A 15 tokens copy · 89% MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: