software factory skills

99 tagged software factory, measured the same way as everything else here.

Browse within: agent-orchestration 57github-issues 57nextjs 57sqlite 57agentic-engineering 14pi 8prompt-engineering 7spec-driven-development 7Autonomous Agents 6autonomy 6multi-agent-systems 6agentic-coding 5agentic-workflow 5gemini 5

finn-build

01

finna/Finn-loop

Skill Claude CodeCodex

Claim the next safe agent-ready issue from Linear, implement it, and open a PR. Use when asked to run Finn-loop's builder, work the approved queue, or fix Finn-loop review feedback. Designed for /loop; one pass does one unit of work.

305 1mo ago A 57 tokens original MIT

finn-review

02

finna/Finn-loop

Skill Claude CodeCodex

Review open PRs against their linked Linear issues and required GitHub checks, then post a three-group verdict with Finn-loop labels. Use when asked to run Finn-loop's reviewer or review its PR queue. Designed for /loop; never merges or pushes code.

305 1mo ago A 56 tokens original MIT

finn-spec

03

finna/Finn-loop

Skill Claude CodeCodex

Interview the user about a raw idea until confident, then file a build-ready issue in Linear. Use when asked to run Finn-loop's spec interview, draft a queue-ready issue, or plan a feature. Interactive — requires the user present; never run unattended.

305 1mo ago A 56 tokens original MIT

tendril-debug-job

04

Ivy-Interactive/Ivy-Tendril

Skill Claude CodeCodex

Analyze a job's execution artifacts to identify issues and improvement opportunities in Tendril, the promptware instructions, memory, or tools.

174 3d ago A 0 tokens

tendril-debug-plan

05

Ivy-Interactive/Ivy-Tendril

Skill Claude CodeCodex

Debug a Tendril plan by analyzing its execution logs, session JSONL, verification results, and checking infrastructure. Produces actionable bugfix and improvement recommendations. Use when the user wants to investigate why a plan failed, behaved unexpectedly, or to audit plan execution quality.

174 3d ago A 59 tokens

tendrillable

06

Ivy-Interactive/Ivy-Tendril

Skill Claude CodeCodex

Find "Tendrillable" GitHub issues — open, recent, code-requiring issues that an agent can plan and one-shot WITHOUT asking clarifying questions, with high probability of success. Classifies a repo's open issues against the Tendrillable rubric and prints a ranked list of issue URLs. Use when asked to find tendrillable…

174 3d ago A 0 tokens

factory-implement

07

addyosmani/factory

Skill Claude CodeCodex

Claim and implement one ready GitHub issue, run fail-closed gates, obtain independent verification, and open a draft pull request.

162 10d ago A 30 tokens original MIT

factory-monitor

08

addyosmani/factory

Skill Claude CodeCodex

Inspect factory health, stale work, CI, security advisories, and recent run evidence without implementing fixes.

162 10d ago A 24 tokens original MIT

factory-spec

09

addyosmani/factory

Skill Claude CodeCodex

Turn an ambiguous factory issue into human-approved product, behavior, design, and implementation slices. Use interactively before implementation.

162 10d ago A 27 tokens original MIT

review-architecture

10

mrinalwadhwa/fluent

Skill Claude CodeCodex

Reviews the structural quality of code. Use when reviewing structural changes in a diff, checking a new or edited module boundary, or auditing the architecture of a codebase.

84 17d ago A 37 tokens original Apache-2.0

review-behaviors

11

mrinalwadhwa/fluent

Skill Claude CodeCodex

Reviews the quality and coherence of behavior statements. Use when reviewing behavior changes in a diff, checking a new or edited behavior statement, or auditing the behavior statements of a codebase.

84 17d ago A 41 tokens original Apache-2.0

mrinalwadhwa/fluent

Skill Claude CodeCodex

Reviews the accuracy and writing quality of documentation. Use when reviewing doc changes in a diff, checking a new or edited page, or auditing the documentation of a codebase.

84 17d ago A 38 tokens original Apache-2.0

factory-line

13

os-factory/har

Skill Claude CodeCodex

Factory line for executing one station of a declared multi-station program — read the installed line bundle (har line status), plan parallel work into isolated HAR slots, run the cumulative gate with har line gate, and hand off for human review. Use when asked to "run a factory line", "run the next station", "execute…

83 yesterday A 101 tokens original Apache-2.0

new-plugin

14

os-factory/har

Skill Claude CodeCodex

Factory line for adding a new HAR verification plugin (like playwright or rocketsim) for any framework — research the framework docs, build the template under src/templates/plugins/, register it everywhere, validate on a real repository, and open a PR. Use when asked to add/create a plugin, plugin template, or…

83 yesterday A 89 tokens original Apache-2.0

v1-milestone

15

os-factory/har

Skill Claude CodeCodex

Factory line for executing one milestone of the HAR v1.0.0 refactor (epic os-factory/har#225) — plan the wave of parallel subagents, implement each issue in its own HAR slot, ship stacked PRs, run the fixture-e2e milestone gate, and hand off for review. Use when asked to "run the next v1 milestone", "work on v1.0.0"…

83 yesterday A 110 tokens original Apache-2.0

speckit-archive-run

16

racecraft-lab/Paddock

Skill Claude CodeCodex

Archive merged feature specs into project memory with provenance, sweep discovery, and gated cleanup.

11 2mo ago A 23 tokens original MIT

speckit-cleanup-run

17

racecraft-lab/Paddock

Skill Claude CodeCodex

Post-implementation quality gate that reviews changes, fixes small issues (scout rule), creates tasks for medium issues, and generates analysis for large issues.

11 2mo ago A 37 tokens original MIT

racecraft-lab/Paddock

Skill Claude CodeCodex

Verify tasks marked [X] in tasks.md are implemented, not phantom completions (marked done but backed by missing or dead code).

11 2mo ago A 36 tokens original MIT

langgraph-proof

19

zrk222/code-factory

Skill Claude CodeCodex

Use when building, reviewing, or debugging a LangGraph flow that must prove a resumed run preserved its recorded semantic transitions and side-effect discipline.

6 2d ago A 32 tokens

bcp

23

mtthsnc/tempest

Skill Claude CodeCodex

Use when brand or marketing work needs to be on-brand and traceable — set up a Brand Context Protocol (BCP) for a business, capture brand truth, or produce/score a deliverable (landing page, deck, email, ad copy) against the brand. Triggers on "set up a brand", "make this on-brand", "scaffold a BCP", "brand context…

2 2mo ago A 118 tokens original MIT

bench

24

mtthsnc/tempest

Skill Claude CodeCodex

Use when you need to measure front-end performance or Core Web Vitals — LCP, CLS, INP/TBT, page load and hydration timing — for a page or a change, and judge it against a perf budget or baseline to catch regressions. Reach for it when a page feels slow, before/after a UI change, or when you must prove a perf budget…

2 2mo ago A 91 tokens original MIT