Skill Claude CodeCodex
Evaluate the probability of necessity (PN) for a causal factor — would the conclusion fail if this factor were absent?
Research Artifact Stress-Testing Engine — five-campaign adversarial validation producing weakness-annotated verification reports
Skill Claude CodeCodex
Evaluate the probability of necessity (PN) for a causal factor — would the conclusion fail if this factor were absent?
Skill Claude CodeCodex
Strategy: Probability of Necessity and Sufficiency (PNS/PS) — systematically evaluate whether each factor is necessary, sufficient, both, or neither for the conclusion.
Skill Claude CodeCodex
Rate failure mode occurrence probability 1-10. Estimates how likely each failure mode is to manifest during research execution.
Skill Claude CodeCodex
Identify all parameter dimensions along which a claim's validity might vary.
Skill Claude CodeCodex
Build a detailed adversarial persona with background, motivation, expertise, blind spots, and preferred attack patterns.
Skill Claude CodeCodex
Evaluates artifact from a specific assigned perspective. Produces assessment grounded in that viewpoint's values, priorities, and expertise.
Skill Claude CodeCodex
Execute Klein pre-mortem protocol — assume failure has occurred, generate plausible failure scenarios through prospective hindsight.
Skill Claude CodeCodex
Tactic: Pre-mortem rapid screening feeds high-risk items into full FMEA analysis. Bridges fast intuitive generation with systematic structured analysis.
Skill Claude CodeCodex
Execute a single attack probe against an artifact, record the result with evidence and severity classification.
Skill Claude CodeCodex
Strategy: Research execution process FMEA — analyzes how the research process itself can fail during execution, distinct from design-level failures.
Skill Claude CodeCodex
Strategy: Klein pre-mortem — assume the artifact has failed, then retrospect plausible causes. Generates rapid failure scenario catalog.
Skill Claude CodeCodex
Re-evaluate S/O/D scores after mitigation measures are in place. Validates that mitigations actually reduce risk as expected.
Skill Claude CodeCodex
Strategy: Systematic adversarial probing retuned for truth-seeking. Threat surface = the set of load-bearing claims. Output is NOT a resilience score and NOT a hardening list — it is, per claim, the specific observation/computation that would refute it, plus which attacks succeeded. Methods: UFMCS Key Assumptions…
Skill Claude CodeCodex
Campaign: Systematic adversarial attack from military/intelligence/AI-safety traditions. Core question: Can systematic adversarial attacks find fatal flaws? Methods: UFMCS Red Team Handbook v9.0, CIA SAT, Anthropic Red Teaming, NIST AI RMF, Inie et al. 12-strategy taxonomy.
Skill Claude CodeCodex
Strategy: Action Priority matrix — classifies failure modes into H/M/L priority using severity-weighted scoring per AIAG-VDA 2019 Action Priority tables.
Skill Claude CodeCodex
Rate failure mode severity 1-10 based on end-effect impact. Follows AIAG-VDA severity scale calibrated for research artifacts.
Skill Claude CodeCodex
Remove one specified factor from the artifact's support structure and reason about how the conclusion changes.
Skill Claude CodeCodex
Strategy: Multi-agent collaborative debate based on Du et al. Society of Mind. Agents share perspectives iteratively until convergence or divergence is detected.
Skill Claude CodeCodex
Strategy: Military-grade assumption testing — Key Assumptions Check, Devil's Advocacy, Team A/B analysis to expose hidden dependencies and unexamined beliefs.
Skill Claude CodeCodex
Tactic: Progressive debate escalation based on confidence thresholds. Each round increases attack sophistication until defender collapses or proves resilient.
Skill Claude CodeCodex
Paper landscape scan returning abstracts and metadata. Import of literature-engine/paper-overview skill. Abstracts only — no conclusions from abstracts.
Skill Claude CodeCodex
Paper full text access via alphaxiv answerpdfqueries or getpapercontent(fullText=true). Import of literature-engine/paper-research skill. Raw extracted text for precise claims.
Skill Claude CodeCodex
Paper AI summary report via alphaxiv getpapercontent. Import of literature-engine/paper-search skill. Structured AI-generated intermediate report.
Skill Claude CodeCodex
Tactic: Sequential perspective evaluation with divergence aggregation. Each agent evaluates from a distinct viewpoint, then disagreements are surfaced and resolved.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: