data-engineering skills

386 tagged data-engineering, measured the same way as everything else here.

Browse within: data-governance 76business-intelligence 72analytics-engineering 56data-ingestion 48data-integration 48data-science 48apache-spark 39airflow 38databricks 32dbt 27bigquery 23ai-automation 22apache-airflow 22ckan 20

kafka-shadowtraffic

25

lensesio/agentic-engineering-for-apache-kafka

Skill Claude CodeCodex

Part of kafka-skills

Generate a ShadowTraffic configuration to populate a Kafka topic with realistic synthetic data. Discovers the target topic, its key and value schemas, and the correct serializers from the live cluster via any attached Kafka MCP server, then writes a ready-to-run shadowtraffic-config.json and Docker command. Use when…

56 12d ago A 134 tokens original MIT

aidp-fusion-seed

26

oracle-samples/oracle-aidp-samples

Skill Claude CodeCodex

Part of oracle-ai-data-platform-fusion-autopilot

Turn a natural-language seed request into a correct, guarded aidp-fusion-autopilot run --mode seed. Parses intent into scope flags (--datasets / --layers / --strict-scope / --resume), auto-satisfies preconditions (validate + /aidp-fusion-bootstrap + cluster), guards the destructive replace-on-silver/gold behaviour…

46 3d ago A 176 tokens UPL-1.0

mart-author

27

oracle-samples/oracle-aidp-samples

Skill Claude CodeCodex

Part of oracle-ai-data-platform-fusion-autopilot

Author a new medallion node (gold mart, silver dim, or additive column) when the live gold layer AND the content pack both cannot serve a business need. Takes the user's business logic, inspects the Fusion PVO SOURCE schema (not bronze) for available raw fields, and authors the lowest-cost, additive, non-destructive…

46 3d ago A 176 tokens UPL-1.0

northgraindata/dbt-doctor

Skill Claude CodeCodex

CSS and UI animation patterns for responsive, polished interfaces. Use when implementing hover effects, tooltips, button feedback, transitions, or fixing animation issues like flicker and shakiness.

46 +20 1mo ago A 42 tokens original MIT

dbt-doctor

30

northgraindata/dbt-doctor

Skill Claude CodeCodex

Static analysis and health checks for dbt projects. Use before committing SQL/YAML or when enforcing CI quality gates.

46 +20 1mo ago A 28 tokens original MIT

vaquarkhan/data-engineering-agent-skills

Skill Claude CodeCodex

Guides agents through schema-registry-backed event contracts. Use when managing Avro, Protobuf, or JSON Schema for event streams, compatibility policies, producer and consumer evolution, or contract enforcement in messaging systems.

39 +1 2mo ago A 51 tokens original MIT

vaquarkhan/data-engineering-agent-skills

Skill Claude CodeCodex

Guides agents through DuckDB-based local analytics and development workflows. Use when prototyping models locally, validating transformations, reproducing data issues quickly, or building lightweight analytical tooling without a full warehouse.

39 +1 2mo ago A 47 tokens original MIT

MiguelElGallo/iparq

Skill Claude CodeCodex

Inspect Parquet file metadata with the iParq CLI, including compression, encodings, physical and logical types, row groups, sort order, statistics, dictionary pages, page indexes, page locations, Bloom filters, and storage sizes. Use when an agent needs to explain how one or more .parquet files were written, compare…

25 18d ago A 99 tokens original MIT

azure

36

clawdata/clawdata

Skill Claude CodeCodex

Manage Azure cloud resources -- resource groups, storage, databases, functions, and data services using the az CLI.

24 5mo ago A 24 tokens

data-analysis

37

clawdata/clawdata

Skill Claude CodeCodex

Skill "data-analysis" from clawdata/clawdata, covering data analysis, rules, chart code block format, chart spec reference and chart type selection guide.

24 5mo ago A 0 tokens

duckdb

38

clawdata/clawdata

Skill Claude CodeCodex

Query and explore a local DuckDB warehouse -- list tables, inspect schemas, run SQL, ingest CSV/JSON/Parquet files.

24 5mo ago A 30 tokens

datacoolie-build

39

datacoolie/datacoolie

Skill Claude CodeCodex

Build, modify, materialize, run, and verify DataCoolie projects. Use for workspace bootstrap, metadata authoring, environment overlays, capability checks, runners, notebooks, custom functions, narrow unsupported adapters, local tests, immutable builds, and project-owned build/CI automation. This is the sole…

10 10d ago A 92 tokens AGPL-3.0

datacoolie-discover

40

datacoolie/datacoolie

Skill Claude CodeCodex

Inspect data sources and produce verified source evidence for DataCoolie design or build work. Use for every new DataCoolie project before design, for every declared source type, and when an existing source is new, changed, missing evidence, or contradictory. Discovery is read-only evidence and never creates runtime…

10 10d ago A 76 tokens AGPL-3.0

datacoolie-release

41

datacoolie/datacoolie

Skill Claude CodeCodex

Plan, preflight, deploy, promote, roll back, or author consume-only release automation for an exact verified DataCoolie build. Use for explicit deployment-lifecycle work; read-only planning may precede authorization, while target mutation requires exact authorization. This skill consumes immutable build artifacts and…

10 10d ago A 85 tokens AGPL-3.0

debabsah/analytics-office

Skill Claude CodeCodex

Part of analytics-office

Use when a finished thing — a source, a result, code, or the record — is about to be trusted or consumed; the gate fires before the work leans on it. Fire FIRST - before a number is BUILT on inherited sources or PRESENTED from them - when the silent assumptions baked into procs, queries, exports, or workbooks have not…

9 2mo ago A 249 tokens original MIT

brief-my-findings

43

debabsah/analytics-office

Skill Claude CodeCodex

Part of analytics-office

Use when work is leaving the desk — findings, a status, or a number that must hold up in the room. The analysis is finished and the findings need communicating to a stakeholder or decision-maker. Composes the brief from the evidence on hand and makes every claim carry its provenance and status, so open questions stay…

9 2mo ago A 180 tokens original MIT

review-my-query

44

debabsah/analytics-office

Skill Claude CodeCodex

Part of analytics-office

Use when a finished thing — a source, a result, code, or the record — is about to be trusted or consumed; the gate fires before the work leans on it. The code behind a number - a SQL query, view, or stored procedure, a dbt or semantic model, a measure, calc group, or RLS rule - needs reviewing for correctness before…

9 2mo ago A 236 tokens original MIT

create-lookalike

45

narrative-io/narrative-skills-marketplace

Skill Claude CodeCodex

Part of narrative-audience

Create a look-alike audience from a seed audience and a candidate population dataset. Classifies Rosetta Stone attributes, generates the same materialized-view scoring pipeline Lookalike Studio emits (Naive-Bayes categorical weights + Gaussian continuous similarity), gates on approval, submits via…

8 6d ago A 138 tokens original MIT

write-nql

46

narrative-io/narrative-skills-marketplace

Skill Claude CodeCodex

Part of narrative-common

Write, validate, and (optionally) execute an NQL query against a Narrative dataset. Drafts the query from the user's question, runs narrativenqlvalidate until it compiles, explains the query in plain English, and only runs it on explicit approval (or when invoked with --run). Use when: "write an NQL query for X"…

8 6d ago A 133 tokens original MIT

narrative-io/narrative-skills-marketplace

Skill Claude CodeCodex

Part of narrative-identity

Compare your data to a partner's data in the marketplace. Given a dataset you already own with person/edge data, this skill walks you through picking a partner data source to match against, choosing which identifier types to match on, optionally selecting which enrichment attributes to attach, and then submits the…

8 6d ago A 154 tokens original MIT

luccapinto/agentic-data-kit

Skill Claude CodeCodex

Use when the user wants to create, add, extend, or modify an agent, skill, or workflow in this kit (e.g. "create a skill for our dbt naming standard", "add an agent for X", "new workflow"). Decides whether to build an agent, a skill, or nothing, enforces lean quality standards, and keeps every installed AI-tool folder…

7 1mo ago A 110 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: