zjunlp

154 mods across 10 repositories, 3.1k stars between them.

zjunlp/DataMind

Skill Claude CodeCodex

Answers questions about the structure, metadata, field definitions, and business rules of the dabstep payment processing dataset. Use this skill when questions ask about: column names, field meanings, fee rule structure, which factors affect fees, fee formula, boolean factor effects on cost, volume/fraud/capture-delay…

135 2d ago A 101 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve dabstep FeeDeltaandImpactSimulation questions: computing fee deltas when a fee's rate changes, and identifying which merchants are affected by fee rule changes. Use when asked about fee impact, delta payments, rate changes, or which merchants would be affected by modifying a fee rule.

135 2d ago A 0 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve fraud analysis and general macro transaction analysis questions on the dabstep payment dataset. Use this skill for questions about fraud rates, transaction distributions, merchant rankings, card scheme analysis, country breakdowns, correlation studies, and yes/no comparative questions on payment data.

135 2d ago A 61 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve dabstep dataset questions that ask to identify the most expensive Merchant Category Code (MCC) or the most expensive Authorization Characteristics Indicator (ACI) for a given transaction. Use this skill when the question asks "what is the most expensive MCC for a transaction of X euros" or "what is the most…

135 2d ago A 85 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solves payment routing and cost optimization problems in the dabstep dataset. Use when the question asks which card scheme a merchant should steer traffic to (for minimum or maximum fees), or which ACI (Authorization Characteristics Indicator) to incentivize for fraudulent transactions to achieve the lowest possible…

135 2d ago A 108 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Skill for computing total payment processing fees for a merchant over a specific day, date range, or month in the dabstep dataset. Use this skill whenever the question asks for "total fees", "fees paid", or "fees charged" for a merchant over some time period. The computation requires matching each transaction to a fee…

135 2d ago A 101 tokens

analyzer

80

zjunlp/DataMind

Agent Claude Code

Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.

135 2d ago A 0 tokens

comparator

81

zjunlp/DataMind

Agent Claude Code

Compare two outputs WITHOUT knowing which skill produced them.

135 2d ago A 0 tokens

grader

82

zjunlp/DataMind

Agent Claude Code

Evaluate expectations against an execution transcript and outputs.

135 2d ago A 0 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Analyze a company in an SEC 10-K SQLite database and produce high-quality evidence-grounded financial QA pairs. Use this whenever the user asks to analyze a company by CIK/ticker, inspect 10-K financial trends, generate finance QA datasets, or work with filings/financialfacts tables.

135 2d ago A 67 tokens

ddr-globem-analysis

84

zjunlp/DataMind

Skill Claude CodeCodex

Analyze a specific participant's longitudinal passive-sensing and psychological data in the GLOBEM digital depression research dataset. Use this skill whenever the task involves: analyzing a user's mental health or behavioral data from wearables/smartphones, generating QA pairs about behavioral/psychological changes…

135 2d ago A 98 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive strategy for analyzing individual patient records in MIMIC-IV EHR database and generating high-quality, diverse QA pairs. Use this skill whenever the task involves analyzing a specific patient's clinical data from MIMIC-IV (or similar EHR databases), querying across hospital and ICU tables, and…

135 2d ago A 118 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive individual-user analysis on the GLOBEM dataset — a longitudinal passive-sensing + mental-health study of college students. Use this skill whenever a task involves analyzing a specific participant (e.g. "Analyze user INS-W002") from the GLOBEM dataset, exploring behavioral patterns from smartphone…

135 2d ago A 88 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive patient analysis using the MIMIC-IV clinical database. Use this skill whenever asked to analyze, summarize, or investigate a patient's medical history, hospital admissions, diagnoses, medications, procedures, or clinical course from a MIMIC-IV SQLite database. Triggers on prompts like "Analyze patient…

135 2d ago A 100 tokens

longds-bench

88

zjunlp/DataMind

Skill Claude CodeCodex

Self-evaluate the current agent on LongDS-Bench (zjunlp/DataMind): the long-horizon, multi-turn agentic data-analysis benchmark. Use this when the user asks to run, score, or benchmark an agent on LongDS / LongDS-Bench / DataMind longds, or to measure multi-turn data-analysis ability. This does NOT use DSGym's Docker…

135 2d ago A 176 tokens

mechanist

89

zjunlp/Mechanist

Plugin Claude Code

Plugin marketplace listing 1 plugin: mechanist.

49 6d ago A tokens not measured original MIT

mechanist

90

zjunlp/Mechanist

Plugin Claude Code

Mechanist — AI Systems as Scientific Instruments for Understanding the Mechanisms of Intelligence.

49 6d ago A tokens not measured original MIT

claim

91

zjunlp/Mechanist

Agent

The claim agent of /auto. Runs the /auto-claim skill under two orthogonal axes — BEHAVIORSOURCE (given / given-validation / discovery) sets where the behavior comes from and whether it is validated; MECHANISM (given / discovery) sets who picks the mechanism method. discovery generates ranked, novelty-checked ideas…

49 6d ago A 129 tokens original MIT

experiment

92

zjunlp/Mechanist

Agent

The experiment agent of /auto. Wraps the /auto-experiment skill, which folds mechanism-family routing inline before implementing, code-reviewing, and deploying the experiment suite. Supports two-step invocation — first call returns candidate families for the orchestrator's mini-prompt, second call (with chosenfamily)…

49 6d ago A 72 tokens original MIT

iteration

93

zjunlp/Mechanist

Agent

The iteration agent of /auto. Runs the /auto-iteration-loop skill — an autonomous review loop that consumes /auto-verify's four-state output (PASS / FAIL / INCONCLUSIVE / ZEROELIGIBLEVARIANTS) plus the orthogonal deferred bucket and routes each claim to the right back-edge (① variant-only fix / ② baseline-script fix /…

49 6d ago A 143 tokens original MIT

verify

94

zjunlp/Mechanist

Agent

The verify agent of /auto. Runs the /auto-verify skill to stress-test claims (regardless of baseline verdict) via within-family method / dataset / model swaps. Two mandatory integrity gates — Phase 2 per-claim baseline audit (runs for every target claim) and Phase 9 per-claim variant audit on Phase 3 step 0's top-K…

49 6d ago B 151 tokens original MIT

ablation-planner

95

zjunlp/Mechanist

Skill Claude CodeCodex

Use when main results pass result-to-claim (claimsupported=yes or partial) and ablation studies are needed for paper submission. The external LLM reviewer (via llm-chat MCP) designs ablations from a reviewer's perspective, CC reviews feasibility and implements.

49 6d ago C 58 tokens original MIT

analyze-results

96

zjunlp/Mechanist

Skill Claude CodeCodex

Analyze ML experiment results, compute statistics, generate comparison tables and insights. Use when user says "analyze results", "compare", or needs to interpret experimental data.

49 6d ago A 37 tokens copy · 97% MIT