zjunlp/DataMind

[ICLR/AAAI/KDD/EMNLP2026] Open-Source LLM-Based Data Analysis Agents

135Stars on the repository
21Mods indexed here, across every type
2d agoLast push, which is what freshness is scored on
noneNo LICENSE: all rights reserved, so bodies are not copied

skill-creator

01

zjunlp/DataMind

Skill Claude CodeCodex

Guide for creating effective skills. This skill should be used when users want to create a new skill (or update an existing skill) that extends Claude's capabilities with specialized knowledge, workflows, or tool integrations.

135 2d ago A 45 tokens

applicable-fee-ids

02

zjunlp/DataMind

Skill Claude CodeCodex

Solve questions about which fee IDs apply to a payment merchant, transaction characteristics, or time period in the dabstep dataset. Use this skill for any question asking "which fee IDs apply to X", "what are the applicable fee IDs for merchant Y", "which merchants are affected by fee Z", or any query involving…

135 2d ago A 81 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve dabstep AverageFeeEstimation problems. Use this skill for questions asking about average payment processing fees, which card scheme is cheapest/most expensive in an average scenario, and fee calculations based on card scheme, account type, MCC description, credit/debit type, or any combination of these filters.

135 2d ago A 69 tokens

csv-analysis

04

zjunlp/DataMind

Skill Claude CodeCodex

Use this skill for CSV data analysis tasks that require reading a local CSV file, checking row counts and columns, grouping records, computing rates or aggregates, creating a chart, and writing a short Markdown report.

135 2d ago A 44 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Skill for computing average transaction value statistics from payment transaction data in the dabstep dataset. Use this skill when a question asks about average transaction amount/value grouped by a categorical field (e.g., shopperinteraction, issuingcountry, acquirercountry, aci), possibly filtered by merchant, card…

135 2d ago A 72 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Answers questions about the structure, metadata, field definitions, and business rules of the dabstep payment processing dataset. Use this skill when questions ask about: column names, field meanings, fee rule structure, which factors affect fees, fee formula, boolean factor effects on cost, volume/fraud/capture-delay…

135 2d ago A 101 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve dabstep FeeDeltaandImpactSimulation questions: computing fee deltas when a fee's rate changes, and identifying which merchants are affected by fee rule changes. Use when asked about fee impact, delta payments, rate changes, or which merchants would be affected by modifying a fee rule.

135 2d ago A 0 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve fraud analysis and general macro transaction analysis questions on the dabstep payment dataset. Use this skill for questions about fraud rates, transaction distributions, merchant rankings, card scheme analysis, country breakdowns, correlation studies, and yes/no comparative questions on payment data.

135 2d ago A 61 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solve dabstep dataset questions that ask to identify the most expensive Merchant Category Code (MCC) or the most expensive Authorization Characteristics Indicator (ACI) for a given transaction. Use this skill when the question asks "what is the most expensive MCC for a transaction of X euros" or "what is the most…

135 2d ago A 85 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Solves payment routing and cost optimization problems in the dabstep dataset. Use when the question asks which card scheme a merchant should steer traffic to (for minimum or maximum fees), or which ACI (Authorization Characteristics Indicator) to incentivize for fraudulent transactions to achieve the lowest possible…

135 2d ago A 108 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Skill for computing total payment processing fees for a merchant over a specific day, date range, or month in the dabstep dataset. Use this skill whenever the question asks for "total fees", "fees paid", or "fees charged" for a merchant over some time period. The computation requires matching each transaction to a fee…

135 2d ago A 101 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Analyze a company in an SEC 10-K SQLite database and produce high-quality evidence-grounded financial QA pairs. Use this whenever the user asks to analyze a company by CIK/ticker, inspect 10-K financial trends, generate finance QA datasets, or work with filings/financialfacts tables.

135 2d ago A 67 tokens

ddr-globem-analysis

13

zjunlp/DataMind

Skill Claude CodeCodex

Analyze a specific participant's longitudinal passive-sensing and psychological data in the GLOBEM digital depression research dataset. Use this skill whenever the task involves: analyzing a user's mental health or behavioral data from wearables/smartphones, generating QA pairs about behavioral/psychological changes…

135 2d ago A 98 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive strategy for analyzing individual patient records in MIMIC-IV EHR database and generating high-quality, diverse QA pairs. Use this skill whenever the task involves analyzing a specific patient's clinical data from MIMIC-IV (or similar EHR databases), querying across hospital and ICU tables, and…

135 2d ago A 118 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive individual-user analysis on the GLOBEM dataset — a longitudinal passive-sensing + mental-health study of college students. Use this skill whenever a task involves analyzing a specific participant (e.g. "Analyze user INS-W002") from the GLOBEM dataset, exploring behavioral patterns from smartphone…

135 2d ago A 88 tokens

zjunlp/DataMind

Skill Claude CodeCodex

Comprehensive patient analysis using the MIMIC-IV clinical database. Use this skill whenever asked to analyze, summarize, or investigate a patient's medical history, hospital admissions, diagnoses, medications, procedures, or clinical course from a MIMIC-IV SQLite database. Triggers on prompts like "Analyze patient…

135 2d ago A 100 tokens

longds-bench

17

zjunlp/DataMind

Skill Claude CodeCodex

Self-evaluate the current agent on LongDS-Bench (zjunlp/DataMind): the long-horizon, multi-turn agentic data-analysis benchmark. Use this when the user asks to run, score, or benchmark an agent on LongDS / LongDS-Bench / DataMind longds, or to measure multi-turn data-analysis ability. This does NOT use DSGym's Docker…

135 2d ago A 176 tokens