forge

forge is an agent for coding agents from namastexlabs/automagik-tools. It costs 21 tokens per session (2,745 once invoked), scanned A, a copy of forge, Apache-2.0.

A planning orchestrator that breaks an approved software request into execution groups, task files, and validation steps. It works across different types of projects and follows a discovery, implementation, and verification structure.

In plain words
What is it for?
It is for preparing coordinated implementation plans, gathering context, recording blockers, defining validation hooks, and linking tasks to project tracking information.
Why use it?
It turns a broad approved request into smaller work items with clearer instructions and checks.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/namastexlabs/automagik-tools/forge
Clone the repo
git clone --depth 1 https://github.com/namastexlabs/automagik-tools

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for forge

README.md
[![agentmods](https://agentmods.dev/badge/agents/namastexlabs/automagik-tools/forge.svg)](https://agentmods.dev/agents/namastexlabs/automagik-tools/forge)
Your own site
<a href="https://agentmods.dev/agents/namastexlabs/automagik-tools/forge"><img src="https://agentmods.dev/badge/agents/namastexlabs/automagik-tools/forge.svg" alt="Measured on agentmods" height="20"></a>
Per session 21 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 2,745 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin 97% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00021 $0.02745
Opus 5 $0.00010 $0.01373
Sonnet 5 $0.00004 $0.00549
Haiku 4.5 $0.00002 $0.00275

Measured 5d ago against content hash d73f27ed2333, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-05, from the pricing page.

Security

Grade A, and why

forge scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

97% identical to forge — 3 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

.genie/agents/forge.md · 291 lines

How it starts

The opening of the file, as written. The whole thing — 291 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Framework Reference

This agent uses the universal prompting framework documented in AGENTS.md §Prompting Standards Framework:

  • Task Breakdown Structure (Discovery → Implementation → Verification)
  • Context Gathering Protocol (when to explore vs escalate)
  • Blocker Report Protocol (when to halt and document)
  • Done Report Template (standard evidence format)

Naming Convention (Code Domain): @.genie/code/spells/emoji-naming-convention.md - MANDATORY when creating Forge tasks for code

Customize phases below for execution breakdown and task planning.

Universal Forge Orchestrator

Identity & Mission

Forge translates an approved wish into coordinated execution groups with documented validation hooks, task files, and tracker linkage. Run it once the wish status is APPROVED; never alter the wish itself—produce a companion plan that makes execution unambiguous.

Works across all domains (code, create) by detecting context from the wish document.

Domain Detection

Detect domain from wish:

  • Code domain:
    • Wish contains <spec_contract>
    • Evidence in qa/ folder
    • Uses emoji naming for tasks
    • References GitHub issues
    • Branch strategy documented
  • Create domain:
    • Wish contains <quality_contract>
    • Evidence in validation/ folder
    • No emoji naming required
    • No GitHub issue reference
    • Optional branch strategy

Operating Context

  • Load the inline <spec_contract> or <quality_contract> from .genie/wishes/<slug>/<slug>-wish.md and treat it as the source of truth
  • Generate .genie/wishes/<slug>/task-<group>.md files so downstream agents can auto-load context via @ references
  • Capture dependencies, personas, and evidence expectations before implementation begins

Success Criteria

  • ✅ Plan saved to .genie/wishes/<slug>/reports/forge-plan-<slug>-<timestamp>.md
  • ✅ Each execution group lists scope, inputs (@ references), deliverables, evidence, suggested persona, dependencies
  • ✅ Groups map to wish evaluation matrix checkpoints (Discovery 30pts, Implementation 40pts, Verification 30pts)
  • ✅ Task files created as .genie/wishes/<slug>/task-<group>.md for easy @ reference
  • ✅ [Code] Branch strategy documented (default feat/<wish-slug>, existing branch, or micro-task)
  • ✅ Validation hooks specify which matrix checkpoints they validate and target score
  • ✅ Evidence paths align with review agent expectations
  • ✅ Approval log and follow-up checklist included
  • ✅ Chat response summarises groups, matrix coverage, risks, and next steps with link to the plan

Read the full file on GitHub · 291 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 5d ago First seen · 291 lines · 21 tokens per session scan A d73f27ed2333

Subscribe to this mod's changes

forge is an agent published in the GitHub repository namastexlabs/automagik-tools (15 stars, last pushed 9mo ago), licensed Apache-2.0. It adds 21 tokens to every session and 2,745 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. It is 97% identical to forge, differing in 3 lines, and is treated as a copy.

Related

Other agents, from other repositories

kingdee-qa-engineer

QA & Test Engineer for the kingdee-mcp project. Authors evals/ and tests/ cases, reproduces bugs against the live K3Cloud environment, and runs regression scans via bin/kmcp test.

WaHaiLong/KingdeeMCP · 49 tokens

security-reviewer

Use after changes to auth.py, credential handling, HTTP request code, sitemap fetching, or quota management. Reviews for SSRF vulnerabilities, credential leaks, API key exposure, and quota exhaustion risks. Read-only: reports issues without fixing them.

FlorianBruniaux/google-search-console-mcp · 0 tokens

gsc-sitemap-auditor

Audits all submitted sitemaps for a property. Use when asked about sitemap health, submitted vs. indexed counts, sitemap errors, or why certain pages are not getting crawled. Aussi déclenché en français par "mon sitemap est à jour", "problème de sitemap", "pourquoi les URLs de mon sitemap sont pas indexées", "combien…

FlorianBruniaux/google-search-console-mcp · 98 tokens

regen

Re-runs the Wiswa generator, applies post-processing, verifies the result is safe, and commits. Use when .wiswa.jsonnet or managed templates change.

Tatsh/wiswa-mcp · 36 tokens

badge-sync

Synchronises the badge list in docs/badges.rst with README.md. Use after editing the README badge block or when docs and README drift.

Tatsh/wiswa-mcp · 33 tokens

forecaster

Time series forecasting agent. Handles decomposition, stationarity testing, ARIMA/ETS model choice, and uncertainty quantification. Use when predicting future values from historical data.

ChrisGVE/localdata-mcp · 36 tokens