do-and-judge
25NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Execute a task with sub-agent implementation and LLM-as-a-judge verification with automatic retry loop.
Hand-crafted Claude Code Skills focused on improving agent results quality. Compatible with OpenCode, Cursor, Antigravity, Gemini CLI, and others. Includes CodeRabbit open-source alternative.
Context Engineering Kit is a collection of skills, agents, hooks, instructions, commands, and plugins that shape how coding agents use context and carry out development work. It is for developers using Claude Code and other supported coding agents who want more predictable results. The catalogue entries are components from this collection.
This repository also configures its own agents. See what context-engineering-kit tells them →
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Execute a task with sub-agent implementation and LLM-as-a-judge verification with automatic retry loop.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Execute tasks through competitive multi-agent generation, meta-judge evaluation specification, multi-judge evaluation, and evidence-based synthesis.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Run independent tasks concurrently across multiple files or targets using parallel sub-agents, with per-task model selection and LLM-as-a-judge verification. Use when tasks do not depend on each other and can run side by side.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Execute one complex task as ordered, dependent steps run sequentially, passing context from each step to the next, with per-step LLM-as-a-judge verification. Use when later steps depend on the results of earlier ones.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Systematically fix all failing tests after business logic changes or refactoring.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Use when adding metadata to commits without changing history, tracking review status, test results, code quality annotations, or supplementing commit messages post-hoc - provides git notes commands and patterns for attaching non-invasive metadata to Git objects.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Use when working on multiple branches simultaneously, context switching without stashing, reviewing PRs while developing, testing in isolation, or comparing implementations across branches - provides git worktree commands and workflow patterns for parallel development with multiple working directories.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Implement a task step by step with automated LLM-as-Judge verification at the end of each phase.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Evaluate solutions through multi-round debate between independent judges until consensus.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Launch a meta-judge then a judge sub-agent to evaluate results produced in the current conversation.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Use when Code implementation and refactoring, architecturing or designing systems, process and workflow improvements, error handling and validation. Provide tehniquest to avoid over-engineering and apply iterative improvements.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Launch an intelligent sub-agent with automatic model selection based on task complexity, specialized agent matching, Zero-shot CoT reasoning, and mandatory self-critique verification.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Load all open issues from GitHub and save them as markdown files.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Use to load open/unresolved PR review comments then aggregate them as tasks in .specs/comments/.md for parallel agents to fix.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Curates insights from reflections and critiques into CLAUDE.md using Agentic Context Engineering.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Design multi-agent architectures for complex tasks. Use when single-agent context limits are exceeded, when tasks decompose naturally into subtasks, or when specializing agents improves quality.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Iterative PDCA cycle for systematic experimentation and continuous improvement.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Refine a draft task specification into a fully planned, implementation-ready task with acceptance criteria, architecture, per-step sub-task files and verifiable phases.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Use this skill when you writing commands, hooks, skills for Agent, or prompts for sub agents or any other LLM interaction, including optimizing prompts, improving LLM outputs, or designing production prompt templates.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Execute complete FPF cycle from hypothesis generation to decision.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Search the FPF knowledge base and display hypothesis details with assurance information.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Reflect on previus response and output, based on Self-refinement framework for iterative improvement with complexity triage and verification.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Reset the FPF reasoning cycle to start fresh.
NeoLabHQ/context-engineering-kit
Skill Claude CodeCodex
Verify what PR review comments have been addressed (committed/pushed OR uncommitted local changes) and resolve the threads that are genuinely fixed or no longer relevant.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: