long-horizon-agents skills

35 tagged long-horizon-agents, measured the same way as everything else here.

Browse within: codex-cli 20cost-efficient-ai 20fable5 6long-horizon-tasks 6long-horizon-terminal-bench 6terminal-bench 6harness 5long-horizon 5longhorizon-harness 5

AMAP-ML/LongHorizon-Harness

Skill Claude CodeCodex

Reproduce CUA-Harness experiments on WeaveBench from a GitHub checkout. Use when the user wants an AI coding agent to set up dependencies, download WeaveBench assets, prepare the 120G VM, configure Qwen/Anthropic-compatible APIs, run smoke tests, launch full or subset evaluations, inspect logs, or summarize scores for…

not rated 1.5k +58 16d ago A 81 tokens original MIT

rewardkit

02

zli12321/LHTB

Skill Claude CodeCodex

Write Harbor task verifiers using Reward Kit. Use when creating or editing a task's tests/ directory, adding grading criteria, setting up LLM/agent judges, or designing verifiers that produce a reward score.

not rated 697 +5 8d ago A 46 tokens copy · 91% Apache-2.0

good-skill

03

sheawinkler/ContextLattice

Skill Claude CodeCodex

Use only for fixture validation of narrow skill routing. Avoid for normal repo work.

not rated 153 +1 9d ago A 20 tokens original Apache-2.0

skill-creator

04

GhabiX/SpineCodex

Skill Claude CodeCodex

Guide for creating effective skills. This skill should be used when users want to create a new skill (or update an existing skill) that extends Codex's capabilities with specialized knowledge, workflows, or tool integrations.

not rated 128 +4 10d ago A 46 tokens copy · 70% Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: