forger-labs-hq

24 mods across 1 repository, 6 stars between them.

forger-labs-hq/researchforge

Skill Claude CodeCodex

Set up the experiment contract and run the frozen baseline benchmark. Use when an improve-repository project needs its evaluation defined, the contract approved, or the baseline measured.

6 2d ago A 38 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Check that ResearchForge's dependencies (git, Python, optionally Docker) are available and explain any failures. Use when setup fails, before starting a project, or when the user asks whether their machine is ready.

6 2d ago A 47 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Generate testable, evidence-linked hypotheses from the research landscape and import them for validation. Use after the landscape exists, when the user wants hypotheses, experiment ideas, or "what should we try?".

6 2d ago A 46 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Synthesize stored papers into a research landscape — grouped directions with evidence claims — and import it for validation. Use after papers are stored, when the user wants directions, themes, or a map of the literature.

6 2d ago A 47 tokens original Apache-2.0

researchforge-paper

05

forger-labs-hq/researchforge

Skill Claude CodeCodex

Build the research package — BibTeX citations, related work, evidence matrix, paper outline, reproducibility bundle, and experiment data. Use when the user wants publication materials or a research write-up bundle.

6 2d ago A 45 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Search arXiv for papers relevant to the project objective and review what was stored. Use when the user wants literature, related work, or asks what papers ResearchForge found.

6 2d ago A 40 tokens original Apache-2.0

researchforge-plan

07

forger-labs-hq/researchforge

Skill Claude CodeCodex

Design experiment variants for a hypothesis — write patches, import the plan for validation, and get it approved. Use after a baseline exists, when the user wants to plan or implement experiments.

6 2d ago A 41 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Summarize a run's results — ranking, Pareto trade-offs, constraint violations, and rejected experiments — grounded strictly in recorded measurements. Use when the user asks how the experiments went or which variant won.

6 2d ago A 46 tokens original Apache-2.0

researchforge-run

09

forger-labs-hq/researchforge

Skill Claude CodeCodex

Execute an approved experiment plan through the screening → full benchmark funnel, or resume an interrupted run. Use when the user says run the experiments, or a run was interrupted.

6 2d ago A 38 tokens original Apache-2.0

researchforge-ship

10

forger-labs-hq/researchforge

Skill Claude CodeCodex

Ship a validated experiment — clean local branch reconstructed from the baseline, engineering report, and optional draft PR. Use when the user wants the winning change as a branch, a report, or a PR.

6 2d ago A 45 tokens original Apache-2.0

researchforge-start

11

forger-labs-hq/researchforge

Skill Claude CodeCodex

Start or resume a ResearchForge project — explore a research idea or improve a repository with benchmarked experiments. Use when the user wants to begin research, set up ResearchForge, or asks "where was I?" in an existing project.

6 2d ago A 51 tokens original Apache-2.0

forger-labs-hq/researchforge

Skill Claude CodeCodex

Run repeated validation benchmarks on a run's finalists so a result can honestly be called validated. Use after a run has a promising winner, or when the user asks to confirm/validate a result.

6 2d ago A 44 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Set up the experiment contract and run the frozen baseline benchmark. Use when an improve-repository project needs its evaluation defined, the contract approved, or the baseline measured.

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Check that ResearchForge's dependencies (git, Python, optionally Docker) are available and explain any failures. Use when setup fails, before starting a project, or when the user asks whether their machine is ready.

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Generate testable, evidence-linked hypotheses from the research landscape and import them for validation. Use after the landscape exists, when the user wants hypotheses, experiment ideas, or "what should we try?".

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Synthesize stored papers into a research landscape — grouped directions with evidence claims — and import it for validation. Use after papers are stored, when the user wants directions, themes, or a map of the literature.

6 2d ago A 0 tokens original Apache-2.0

researchforge-paper

17

forger-labs-hq/researchforge

Cursor rule

Build the research package — BibTeX citations, related work, evidence matrix, paper outline, reproducibility bundle, and experiment data. Use when the user wants publication materials or a research write-up bundle.

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Search arXiv for papers relevant to the project objective and review what was stored. Use when the user wants literature, related work, or asks what papers ResearchForge found.

6 2d ago A 0 tokens original Apache-2.0

researchforge-plan

19

forger-labs-hq/researchforge

Cursor rule

Design experiment variants for a hypothesis — write patches, import the plan for validation, and get it approved. Use after a baseline exists, when the user wants to plan or implement experiments.

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Summarize a run's results — ranking, Pareto trade-offs, constraint violations, and rejected experiments — grounded strictly in recorded measurements. Use when the user asks how the experiments went or which variant won.

6 2d ago A 0 tokens original Apache-2.0

researchforge-run

21

forger-labs-hq/researchforge

Cursor rule

Execute an approved experiment plan through the screening → full benchmark funnel, or resume an interrupted run. Use when the user says run the experiments, or a run was interrupted.

6 2d ago A 0 tokens original Apache-2.0

researchforge-ship

22

forger-labs-hq/researchforge

Cursor rule

Ship a validated experiment — clean local branch reconstructed from the baseline, engineering report, and optional draft PR. Use when the user wants the winning change as a branch, a report, or a PR.

6 2d ago A 0 tokens original Apache-2.0

researchforge-start

23

forger-labs-hq/researchforge

Cursor rule

Start or resume a ResearchForge project — explore a research idea or improve a repository with benchmarked experiments. Use when the user wants to begin research, set up ResearchForge, or asks "where was I?" in an existing project.

6 2d ago A 0 tokens original Apache-2.0

forger-labs-hq/researchforge

Cursor rule

Run repeated validation benchmarks on a run's finalists so a result can honestly be called validated. Use after a run has a promising winner, or when the user asks to confirm/validate a result.

6 2d ago A 0 tokens original Apache-2.0