Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/flonat/flonat-research/data-analysisnpx skills add flonat/flonat-research --skill data-analysisgit clone --depth 1 https://github.com/flonat/flonat-researchWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/flonat/flonat-research/data-analysis)<a href="https://agentmods.dev/skills/flonat/flonat-research/data-analysis"><img src="https://agentmods.dev/badge/skills/flonat/flonat-research/data-analysis.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00041 | $0.01964 |
| Opus 5 | $0.00020 | $0.00982 |
| Sonnet 5 | $0.00008 | $0.00393 |
| Haiku 4.5 | $0.00004 | $0.00196 |
Grade A, and why
data-analysis scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 146 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Data Analysis Pipeline
Generate, execute, and verify analysis scripts across R, Python, Stata, and Julia.
Modes
| Mode | What it does | Phases |
|---|---|---|
| EDA | Exploratory data analysis only | 1–2 |
| Estimation | Estimation + publication output (requires locked design) | 1, 3–4 |
| Full | Complete pipeline | 1–5 |
Default: Full. Detect mode from user request or ask if ambiguous.
When to Use
- "Analyse this data" / "Run EDA on this CSV" / "Estimate the model"
- "Generate results tables" / "Create publication figures"
- Any task requiring data → script → output pipeline
When NOT to Use
- Experimental design or power analysis →
experiment-design - Generating synthetic data for testing →
synthetic-data - Auditing identification strategy →
causal-design - Proofreading or compiling the paper →
proofread,latex
Shared References
- Method probing questions:
shared/method-probing-questions.md— ask before running any analysis - Validation tiers:
shared/validation-tiers.md— declare tier before examining results - Escalation protocol:
shared/escalation-protocol.md— escalate when methodology answers are vague - Distribution diagnostics:
shared/distribution-diagnostics.md— mandatory DV checks before model selection - Engagement-stratified sampling:
shared/engagement-stratified-sampling.md— stratify by engagement tiers for social media data - Inter-coder reliability:
shared/intercoder-reliability.md— per-category reliability for content analysis and LLM annotation
Workflow
Phase 1: Setup
- Detect project structure: Read
CLAUDE.md, check fordata/,code/,paper/directories. - Detect language: Check existing scripts, user preference, or ask. Read
shared/multi-language-conventions.mdfor the chosen language's conventions. - Locate data: Find datasets in
data/raw/ordata/processed/. Never modifydata/raw/(perdata-sensitivityrule). For social-media datasets, followshared/engagement-stratified-sampling.mdwhen constructing analysis samples. - Confirm validation tier per
shared/validation-tiers.md. Tier dictates claim-strength language allowed in Phase 4 outputs and how strict the locked-design gate (step 5) is enforced. - Check for locked design: Look for analysis plan in
log/plans/,.context/project-recap.md, orMEMORY.mdestimand registry. If running Estimation or Full mode and no design exists, stop and warn: "No locked research design found. Runexperiment-designorcausal-designfirst, or confirm the specification before proceeding." Useshared/method-probing-questions.mdto probe gaps if the user pushes back on the gate.
What ships with it
9 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- references/econ-visualisation.md 3.0 KB
- references/estimation-recipes.md 3.0 KB
- references/language-conventions.md 2.0 KB
- references/matplotlib.md 11 KB
- references/plotly.md 6.7 KB
- references/seaborn.md 19 KB
- references/statsmodels.md 19 KB
- references/sympy.md 13 KB
- references/table-formatting.md 4.1 KB
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- yesterday First seen · 146 lines · 41 tokens per session scan A 0f05e14a659a
data-analysis is a skill published in the GitHub repository flonat/flonat-research (131 stars, last pushed 10d ago), licensed MIT. It adds 41 tokens to every session and 1,964 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.
Other skills, from other repositories
latex-compile
Compile a LaTeX document and fix every error plus aesthetic issue (overfull/underfull boxes, widows, alignment, fonts) for a clean PDF and log. Use this instead of running pdflatex/latexmk manually — it avoids the latexmk stale-log trap and silent grep failures on binary log output, and it reformats rather than…
nb-to-wolfbook
Convert Mathematica .nb or .m files to Wolfbook .wb format so they open and run in VS Code. Use when bringing existing .nb/.m files into Wolfbook, or to make an existing .wb bridge-safe.
sync-wb-nb
Propagate a change made in a Wolfbook .wb notebook into the paired .nb notebook so the two stay identical. Use immediately after every .wb edit.
wolfram-headless
Run heavy Wolfram Language (wolframscript) computations from Claude Code reliably, and diagnose the misleading "The product exited because of a license error". Use whenever invoking wolframscript on a non-trivial computation, when a wolframscript job dies with a "license error" despite a valid license, or when Wolfram…
cross-validate
Format a result, derivation, or numerical value for independent verification by a second model. Use when you want a cross-check on an important or contested result.
verify-citation
Confirm a paper actually exists (arXiv / Semantic Scholar / OpenAlex) before citing it. Use before writing any new citation.