nature-data

nature-data is a skill for Claude Code, Codex from YuanyuanMa03/academic-research-skills. It costs 90 tokens per session (1,292 once invoked), scanned A, a copy of nature-data, MIT.

A writing aid for preparing Nature journal data-availability statements and related data-sharing details. Nature is a group of scientific journals with specific requirements for explaining where research data can be found and how it can be accessed.

In plain words
What is it for?
Use it when preparing a manuscript’s Data Availability section, repository plan, dataset references, or FAIR metadata checklist. FAIR means data should be findable, accessible, interoperable, and reusable.
Why use it?
It turns vague or incomplete data-sharing information into clearer publication-ready wording and identifies missing details such as repositories, access restrictions, accession numbers, or dataset citations.

Skill for Claude CodeCodex

Written for Claude Code and Codex: argument-hint in frontmatter, but also agents/openai.yaml present.

Part of the nature-data plugin — 1 skill shipped together

Good fit Use it when preparing a manuscript’s Data Availability section, repository plan, dataset references, or FAIR metadata checklist. FAIR means data should be findable, accessible, interoperable, and reusable.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/yuanyuanma03/academic-research-skills/nature-data
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add YuanyuanMa03/academic-research-skills --skill nature-data
Clone the repo
git clone --depth 1 https://github.com/YuanyuanMa03/academic-research-skills

Made for: Claude Code, Codex.

Or install nature-data, the plugin that ships this one along with the rest of its 1 skill.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for nature-data

README.md
[![agentmods](https://agentmods.dev/badge/skills/yuanyuanma03/academic-research-skills/nature-data/github.svg)](https://agentmods.dev/skills/yuanyuanma03/academic-research-skills/nature-data)
Your own site
<a href="https://agentmods.dev/skills/yuanyuanma03/academic-research-skills/nature-data"><img src="https://agentmods.dev/badge/skills/yuanyuanma03/academic-research-skills/nature-data/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for nature-data

Your own site · 80×15
<a href="https://agentmods.dev/skills/yuanyuanma03/academic-research-skills/nature-data"><img src="https://agentmods.dev/badge/skills/yuanyuanma03/academic-research-skills/nature-data.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 90 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,292 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin 98% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00090 $0.01292
Opus 5 $0.00045 $0.00646
Sonnet 5 $0.00018 $0.00258
Haiku 4.5 $0.00009 $0.00129

Measured 12d ago against content hash f14fff4abe44, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

nature-data scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

98% identical to nature-data — 2 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

plugins/nature-data/skills/nature-data/SKILL.md · 120 lines

How it starts

The opening of the file, as written. The whole thing — 120 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Nature Data Availability Skill

Use this skill to turn a manuscript's supporting data into a transparent, Nature-ready data availability package: statement text, repository plan, dataset citations, and missing-information flags.

The governing policy layer is Springer Nature / Nature Portfolio data policy. The implementation layer is FAIR data practice and DataCite-style citation metadata.

Chinese-user operating mode

When the user writes in Chinese, provides a Chinese manuscript note, or asks for "中文对应", "中英对照", "数据可用性声明", "数据获取声明", "原始数据", "数据存储库", or "受限数据":

  • Accept Chinese input naturally, but draft the final submission-ready statement in English unless the user explicitly asks for Chinese only.
  • Preserve a short Chinese explanation of unresolved decisions when it helps the author act.
  • Translate intent, not wording. Chinese phrases such as "可向通讯作者索取" are usually too vague for Nature-style English unless the restriction and access process are specified.
  • Convert Chinese repository/status descriptions into precise publication terms: 数据可用性声明 -> Data Availability; 原始数据 -> raw data; 处理后数据 -> processed data; 源数据 -> source data; 补充材料 -> Supplementary Information; 受限数据 -> restricted data; 合理请求 -> reasonable request, only with reason and review route.
  • Use references/chinese-author-alignment.md for Chinese terminology, common CN-to-EN failure modes, and bilingual intake questions.

Default stance

  • Treat the Data Availability statement as a link between the paper's claims and the evidence needed to inspect, reproduce, or reuse them.
  • Do not invent DOIs, accession numbers, repository names, licences, embargo dates, ethics approvals, access committees, or data-use conditions.
  • Prefer public, discipline-specific repositories. Use generalist or institutional repositories only when no suitable community repository exists.
  • Describe both newly generated data and reused third-party data.
  • If data cannot be openly shared, state why, who controls access, how requests are evaluated, and what metadata or representative data can still be public.
  • Separate data, code, materials, and protocols unless the journal asks for a combined availability section.
  • Keep this skill focused on availability and metadata. Do not rewrite methods, analyze statistics, or polish the manuscript unless the user asks for those tasks separately.
  • Flag "available upon request" as weak unless there is a specific legal, ethical, commercial, or third-party restriction.

Read the full file on GitHub · 120 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 120 lines · 90 tokens per session scan A f14fff4abe44

Subscribe to this mod's changes

nature-data is a skill published in the GitHub repository YuanyuanMa03/academic-research-skills (63 stars, last pushed 22d ago), licensed MIT. It adds 90 tokens to every session and 1,292 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. It is 98% identical to nature-data, differing in 2 lines, and is treated as a copy.

Related

Other skills, from other repositories

scientific-writing

Core skill for the deep research and writing tool. Write scientific manuscripts in full paragraphs (never bullet points). Use two-stage process with (1) section outlines with key points using research-lookup then (2) convert to flowing prose. IMRAD structure, citations (APA/AMA/Vancouver), figures/tables, reporting…

LeonChaoX/qinyan-academic-skills · 87 tokens

venue-templates

Access comprehensive LaTeX templates, formatting requirements, and submission guidelines for major scientific publication venues (Nature, Science, PLOS, IEEE, ACM), academic conferences (NeurIPS, ICML, CVPR, CHI), research posters, and grant proposals (NSF, NIH, DOE, DARPA). This skill should be used when preparing…

LeonChaoX/qinyan-academic-skills · 97 tokens

rowan

Cloud-based quantum chemistry platform with Python API. Preferred for computational chemistry workflows including pKa prediction, geometry optimization, conformer searching, molecular property calculations, protein-ligand docking (AutoDock Vina), and AI protein cofolding (Chai-1, Boltz-1/2). Use when tasks involve…

LeonChaoX/qinyan-academic-skills · 115 tokens

cbioportal-database

Query cBioPortal for cancer genomics data including somatic mutations, copy number alterations, gene expression, and survival data across hundreds of cancer studies. Essential for cancer target validation, oncogene/tumor suppressor analysis, and patient-level genomic profiling.

LeonChaoX/qinyan-academic-skills · 57 tokens

monarch-database

Query the Monarch Initiative knowledge graph for disease-gene-phenotype associations across species. Integrates OMIM, ORPHANET, HPO, ClinVar, and model organism databases. Use for rare disease gene discovery, phenotype-to-gene mapping, cross-species disease modeling, and HPO term lookup.

LeonChaoX/qinyan-academic-skills · 68 tokens

interpro-database

Query InterPro for protein family, domain, and functional site annotations. Integrates Pfam, PANTHER, PRINTS, SMART, SUPERFAMILY, and 11 other member databases. Use for protein function prediction, domain architecture analysis, evolutionary classification, and GO term mapping.

LeonChaoX/qinyan-academic-skills · 62 tokens