Orchestra-Research/AI-Research-SKILLs
Skill Claude CodeCodex
Expert guidance for GRPO/RL fine-tuning with TRL for reasoning and task-specific model training.
192 tagged Reasoning, measured the same way as everything else here.
Orchestra-Research/AI-Research-SKILLs
Skill Claude CodeCodex
Expert guidance for GRPO/RL fine-tuning with TRL for reasoning and task-specific model training.
Skill Claude CodeCodex
Write a monthly trend newsletter for this ai-agent-papers repo. Use when asked to create/update a monthly trend report, summarize a category's recent papers, or "make a newsletter". Encodes the house procedure: cover every coherent category with enough papers that month (no fixed number of issues; overlap between…
zhiliscope/baiyueguang-learning-skill
Skill Claude CodeCodex
Transform knowledge points into an investigable information network using the Baiyueguang Learning Method.
Plugin Claude Code
Grounds a coding agent's architecture decisions in real arXiv prior art: parallel isolated reads of papers found category-wise (divergence), converged into one recommended path with citations, a first step, and known failure modes to avoid.
Skill Claude CodeCodex
Grounds a coding agent's architecture decisions in real arXiv prior art before it builds something new. Reads arXiv category-wise via real HTTP fetch, spawns parallel isolated reads across the papers found, scores/clusters them, then converges on ONE recommended path with citations, a first step, and known prior-art…
Skill Claude CodeCodex
Determines whether a security incident involves personal data, triggers regulatory breach-notification obligations (e.g. GDPR Art. 33/34), and what notification timelines apply.
Skill Claude CodeCodex
Analyzes authentication and authorization events for failed-login clustering, privilege-escalation chains, credential-stuffing patterns, and MFA-bypass indicators.
Skill Claude CodeCodex
Detects anomalous behavior by authenticated users that may indicate insider threats through behavioral analysis of access patterns, privileged actions, and data movement.
Skill Claude CodeCodex
This skill is designed to rewrite user prompts to align with the expert-level aesthetic standards. It transforms simple descriptions into multi-dimensional, professional-grade artistic instructions that maximize scores across all fine-grained aesthetic attributes.
Skill Claude CodeCodex
This skill should be triggered when the user's request involves multiple objects, complex scene arrangements, or specific physical relationships between elements.
Skill Claude CodeCodex
This skill should be triggered when the user's request contains specific quotes, slogans, names, or any intended text that needs to appear accurately and aesthetically within the generated image.
Skill Claude CodeCodex
Use for composing DSPy modules with Ensemble, MultiChainComparison, ensemble voting, sequential pipelines, and multi-program workflows.
Skill Claude CodeCodex
Use for creating custom DSPy modules, extending dspy.Module, reusable components, stateful modules, serialization, and module testing.
Skill Claude CodeCodex
Guide for designing, implementing, and advising on MCP servers that use Code Mode — the pattern where an LLM writes and executes code to orchestrate API calls instead of calling individual tools one at a time. Use this skill whenever the user mentions Code Mode, MCP tool proliferation, context window bloat from MCP…
Skill Claude CodeCodex
Run an Agent-Native Architecture Audit using Thoughtbox Hub for multi-agent coordination. Spawns 3 auditor agents and 1 synthesizer that collaborate through structured channels, cross-reference findings, and build consensus on scores.
Skill Claude CodeCodex
Orchestrate a multi-agent collaboration demo on the Thoughtbox Hub. Spawns MANAGER, ARCHITECT, and DEBUGGER agents that coordinate through shared workspaces, problems, proposals, and channels.
Plugin Claude Code
Plugin marketplace listing 1 plugin: thoughtbox-claude-code.
Agent Claude Code
Proactively audit the assumption registry for stale, unverified, or newly-broken assumptions. Use on a schedule (weekly) or when starting work that depends on external systems. Unlike dependency-verifier (reactive, investigates on demand), this agent walks the full registry and flags what needs attention.
Agent Claude Code
Track aggregate token spend and API costs across all agent operations. Enforce budget constraints, detect cost anomalies, and recommend budget reallocation. Use for daily cost rollups, budget enforcement, and spend optimization.
Agent Claude Code
Adversarial reviewer that systematically attacks specs, implementations, and plans to find logical gaps, missed edge cases, implicit assumptions, and over/under-engineering. Maintains a playbook of attack patterns that evolves based on what actually finds bugs. Use on any artifact you want stress-tested before…
MCP server Claude CodeCodexCursor +2
Extended reasoning with structured processes, persistence, and workflow guidance. Runs locally from the @kastalien-research/thoughtbox npm package.
Skill Claude CodeCodex
Reasoning protocol distilled from Claude Fable 5. Makes any model reason like Fable — evidence-grounded claims, multi-hypothesis diagnosis, concrete simulation, adversarial self-review, calibrated outcome-first delivery. Its never-skipped Floor check catches simple-looking trick questions models answer confidently…
angrysky56/advanced-reasoning-mcp
Cursor rule Cursor
CodeGraph MCP usage guide — when to use which tool.
angrysky56/advanced-reasoning-mcp
Instructions file CodexOpenCode
Instructions for angrysky56/advanced-reasoning-mcp, covering agents.md — advanced reasoning mcp, what this server gives you, mental model, the core workflow and tools.
At most 3 mods per repository are shown here — the rest are on their repository pages: