wjgoarxiv

16 mods across 3 repositories, 30 stars between them.

autoresearch

01

wjgoarxiv/autoresearch-skill

Plugin Claude Code

Autonomous research loops for any domain. 10 commands: core loop, plan, debug, fix, predict, security, scenario, reason, ship.

30 2mo ago A tokens not measured original MIT

autoresearch

02

wjgoarxiv/autoresearch-skill

Plugin Claude Code

Autonomous research loops with 10 commands. Generalizes Karpathy's autoresearch loop to any domain with mechanical evaluation, overnight persistence, and zero dependencies.

30 2mo ago A tokens not measured original MIT

autoresearch-skill

03

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Autonomous research and experimentation toolkit with 10 commands. Core loop inspired by Karpathy's autoresearch — generalizes to any domain with mechanical evaluation, overnight persistence, and zero dependencies. TRIGGER when: user wants autonomous experiments; user mentions "autoresearch" or "auto-research"; user…

30 2mo ago A 153 tokens original MIT

pdf

04

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When the LLM (Claude, ChatGPT, Gemini, or others) needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

30 2mo ago A 63 tokens original MIT

pdf

05

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When the LLM (Claude, ChatGPT, Gemini, or others) needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

30 2mo ago A 63 tokens copy · 72% MIT

autoresearch

06

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Core autonomous research loop. Reads research.md, proposes hypotheses, runs experiments, evaluates results mechanically, keeps improvements, discards failures, and iterates until the target metric is achieved or the iteration budget is exhausted. TRIGGER when: user invokes "autoresearch" (no subcommand); research.md…

30 2mo ago A 79 tokens original MIT

autoresearch:debug

07

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Scientific bug hunting using falsifiable hypotheses. Forms hypotheses, designs falsifying tests, eliminates candidates systematically, and logs the full investigation trail in a structured debug/ folder. TRIGGER when: user has a bug to investigate scientifically; user wants systematic root-cause analysis; user says…

30 2mo ago A 143 tokens original MIT

autoresearch:fix

08

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Iterative error-crusher loop that auto-stops at 0 errors. Cascade-aware: fixes dependency errors before their dependents. Refuses anti-patterns that hide errors instead of fixing them. TRIGGER when: user has errors or failures to fix iteratively; user asks to "fix all errors"; user has a failing test suite; user has…

30 2mo ago A 133 tokens original MIT

autoresearch:learn

09

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Feedback-driven self-improvement protocol for autoresearch-skill. Converts failed runs, confusing transcripts, bad outputs, or user feedback into a bounded improvement plan, an eval scenario, and a patch checklist without executing the patch automatically. TRIGGER when: user says the skill failed, wants to improve…

30 2mo ago A 142 tokens original MIT

autoresearch:plan

10

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

7-step setup wizard that produces a complete, ready-to-run research.md without executing the research loop. Walks the user through goal, metric, search space, constraints, evaluator design, and baseline measurement, then writes the file. TRIGGER when: user wants to set up a research project; user wants to plan before…

30 2mo ago A 134 tokens original MIT

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Multi-perspective deliberation engine. Gathers independent positions from diverse personas, runs cross-examination and rebuttal rounds, detects herd behavior, and synthesizes a neutral judge verdict with confidence levels. TRIGGER when: user wants multi-perspective prediction, forecasting, scenario analysis, decision…

30 2mo ago A 99 tokens original MIT

autoresearch:reason

12

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Adversarial multi-round reasoning with blind-judge panel to reach rigorous conclusions. TRIGGER when: user wants rigorous reasoning or argument evaluation; user wants a decision analyzed from multiple angles; user wants devil's advocate critique; user asks "what are the strongest arguments for/against"; user wants a…

30 2mo ago A 123 tokens original MIT

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

12-dimension scenario exploration across user-specified domain modes. TRIGGER when: user wants to explore scenarios, edge cases, or what-if analysis; user asks "what could go wrong"; user wants failure mode analysis; user asks about best/worst case outcomes; user wants to stress-test a plan, design, or system; user…

30 2mo ago A 124 tokens original MIT

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Iterative security audit engine. Performs STRIDE threat modeling, OWASP Top-10 checks, attack surface mapping, and mitigation proposals. Loops until coverage target is reached or budget is exhausted. TRIGGER when: user wants a security audit, threat model, vulnerability assessment, penetration test review, "is this…

30 2mo ago A 94 tokens original MIT

autoresearch:ship

15

wjgoarxiv/autoresearch-skill

Skill Claude CodeCodex

Universal shipping workflow: 8-phase linear pipeline from verification to deploy. Reads type-checklists.md to select the right checklist for the artifact type. The ONLY pause is Phase 7 (user confirmation before irreversible deploy/publish). TRIGGER when: user wants to ship, release, publish, or deploy something; user…

30 2mo ago A 157 tokens original MIT

vmd-hydrate-mcp

16

wjgoarxiv/vmd-hydrate-mcp

MCP server Claude CodeCodexCursor

An MCP server that drives VMD for GROMACS/LAMMPS trajectory analysis, scripted rendering, and clathrate-hydrate cage science (H-bond networks, F3/F4 order parameters). Runs locally from the vmd-hydrate-mcp Python package.

0 2mo ago A tokens not measured original MIT