harness skills

969 tagged harness, measured the same way as everything else here.

Browse within: harness-engineering 148pi 75harness-ai 58deepseek 56Evaluation 55agent-orchestration 55effect 53explicit 53patterns 53rules 53agentic-coding 51loop-engineering 51agentic-workflow 46benchmark 39

alibaba/open-code-review

Skill Claude CodeCodex

Delegation mode for open-code-review (OCR). Instead of OCR calling an LLM endpoint, this skill instructs the host agent to perform the code review itself, using OCR only for deterministic engineering: file selection and rule resolution. Use when the host agent should drive the review with its own LLM capabilities.

22k 4d ago A 68 tokens original Apache-2.0

open-code-review

02

alibaba/open-code-review

Skill Claude CodeCodex

Performs AI-powered code review on Git changes using the ocr CLI from alibaba/open-code-review. Use when the user asks to review code, review a pull request, review staged/unstaged changes, review a commit, or compare branches for code quality issues. Produces line-level review comments and can automatically apply…

22k 4d ago A 98 tokens original Apache-2.0

add-new-model

03

pydantic/pydantic-ai

Skill Claude CodeCodex

Add support for a newly-released LLM model in pydantic-ai (e.g. openai:gpt-5.6, anthropic:claude-sonnet-5). Use when a provider ships a new model id and you need to wire literals, profile flags, and tests to recognize it. Handles SDK-lag, gateway list conventions, and capability probing.

20k 2d ago A 80 tokens original MIT

complete-partial-pr

04

pydantic/pydantic-ai

Skill Claude CodeCodex

Evaluate and complete an issue or PR where the submitted patch fixes only a narrow symptom of the reported pain point. Use when a contribution may miss adjacent integration surfaces, provider/spec semantics, roundtrip behavior, tests, docs, or historical maintainer decisions.

20k 2d ago A 55 tokens original MIT

pydantic/pydantic-ai

Skill Claude CodeCodex

Build AI agents with Pydantic AI — tools, capabilities (including on-demand loading), structured output, streaming, testing, and multi-agent patterns. Use when the user mentions Pydantic AI, imports pydanticai, or asks to build an AI agent, add tools/capabilities, defer capability loading, stream output, define agents…

20k 2d ago A 85 tokens original MIT

harness-creator

06

walkinglabs/learn-harness-engineering

Skill Claude CodeCodex

Build, audit, and improve harnesses that make AI coding agents reliable: AGENTS.md/CLAUDE.md instruction files, feature/state tracking, verification gates, scope boundaries, session handoff, memory persistence, context budgets, tool-permission safety, and multi-agent coordination. Use this whenever a coding agent is…

15k 6d ago A 142 tokens original MIT

browse

07

yc-software/qm

Skill Claude CodeCodex

Drive a real stealth browser from your shell — act on websites (order food, file an expense, pull data behind a login), with per-person persistent sign-ins via the provider's managed auth (Kernel, Anchor, or Browserbase — picked by which API key you have). Use for ACTING on a site; to just read a page, use curl/wget…

14k 3d ago A 101 tokens original MIT

google-drive-sheets

08

yc-software/qm

Skill Claude CodeCodex

Find, read, export, edit, and manage the user's Google Drive, Docs, Sheets, and Slides through per-user OAuth.

14k 3d ago A 31 tokens original MIT

popular-web-designs

09

yc-software/qm

Skill Claude CodeCodex

54 real-world design systems (Stripe, Linear, Vercel, Notion, Apple…) as ready-to-paste HTML/CSS — exact colors, typography, components, and spacing for styling a page after a known brand.

14k 3d ago A 51 tokens original MIT

mindfold-ai/Trellis

Skill Claude CodeCodex

Systematic first principles thinking for any problem domain. Use when the user says "analyze from first principles", "第一性原理", "从根本分析", "从零开始思考", "think from scratch", "question this design", "is this the right approach", "challenge assumptions", "挑战假设", "为什么要这样做", "有没有更好的方案", "why are we doing it this way", or needs…

14k 5d ago A 164 tokens AGPL-3.0

python-design

11

mindfold-ai/Trellis

Skill Claude CodeCodex

Python design patterns for CLI scripts and utilities — type-first development, deep modules, complexity management, and red flags. Use when reading, writing, reviewing, or refactoring Python files, especially in .trellis/scripts/ or any CLI/scripting context. Also activate when planning module structure, deciding…

14k 5d ago A 73 tokens AGPL-3.0

ts-sdk-author

12

mindfold-ai/Trellis

Skill Claude CodeCodex

Design, build, verify, and publish production-grade TypeScript SDKs as npm packages inside a pnpm monorepo. Covers workspace layout, public API and module boundaries, plugin extension points, branded types and library-tuned tsconfig, tsdown bundling (vs tsup/tsc-only/unbuild), package.json exports with dual ESM+CJS…

14k 5d ago A 219 tokens AGPL-3.0

harness

13

revfactory/harness

Skill Claude CodeCodex

A meta-skill for designing and maintaining a project harness: a coordinated set of specialist agents, skills, and project instructions. It defines their roles, connects them to the project, and updates the setup as the work changes.

8.9k 1mo ago A 170 tokens original Apache-2.0

chaitanyagiri/munder-difflin

Skill Claude CodeCodex

A skill for creating hand-drawn illustrations for Chinese articles, posts, blogs, notes, and documents, using a recurring black character and a mostly white style with a few colored annotations.

5.6k 2d ago A 122 tokens

capabilities

15

chaitanyagiri/munder-difflin

Skill Claude CodeCodex

Your capability catalog — read this at boot. Lists the temporal date-range skills and the external integrations (reached via the loopback broker) available to you as a spawned worker, and exactly how to call each. Read-only. Consult it whenever you're unsure what tools/integrations you have or how to invoke them.

5.6k 2d ago A 69 tokens

temporal

16

chaitanyagiri/munder-difflin

Skill Claude CodeCodex

Resolve ANY named time window — today, yesterday, thisWeek, lastWeek, last7Days, last30Days, last90Days, thisMonth, lastMonth, thisQuarter, lastQuarter, thisYear, lastYear, last12Months — or an arbitrary range (lastNdays / lastNweeks / lastNmonths) to a concrete ISO date range relative to your run time. Read-only: no…

5.6k 2d ago A 114 tokens

gh-pr-description

17

vercel/eve

Skill Claude CodeCodex

Drafts and reviews GitHub pull request descriptions for the eve repository. Use when opening, updating, or reviewing a PR, or when summarizing a branch for reviewers.

4.9k 2d ago A 38 tokens original Apache-2.0

technical-writing

18

vercel/eve

Skill Claude CodeCodex

Write, edit, review, or audit user-facing documentation for the eve repository. Use for changes under docs/, documentation tied to eve APIs or CLI behavior, docs work based on Slack or support feedback, and requests to make eve docs clearer, more natural, or less AI-patterned while verifying claims against current…

4.9k 2d ago A 78 tokens original Apache-2.0

toolkit-guide

19

vercel/eve

Skill Claude CodeCodex

How to triage an account and verify packaged skill resources with the toolkit CRM extension.

4.9k 2d ago A 17 tokens original Apache-2.0

dap-debugging

20

getkimchi/kimchi

Skill Claude CodeCodex

Diagnose runtime state with persistent DAP debugger sessions — breakpoints, expression eval, and stepping across Go, Python, TypeScript/JavaScript, and native binaries.

2.2k 2d ago A 37 tokens original Apache-2.0

improve

21

getkimchi/kimchi

Skill Claude CodeCodex

Run the curator — consolidate the agent-created skill library via umbrella-building.

2.2k 2d ago A 16 tokens original Apache-2.0

agent-mode

22

PhyAgentOS/PhyAgentOS-core

Skill Claude CodeCodex

Unified tool for managing agent LLM modes (add, remove, update, list, switch).

1.9k 2d ago A 22 tokens original MIT

image

24

PhyAgentOS/PhyAgentOS-core

Skill Claude CodeCodex

Unified tool for analyzing and displaying images (vision analysis and image display).

1.9k 2d ago A 16 tokens original MIT