Data and AI skills

10,049 tagged Data and AI, measured the same way as everything else here.

Browse within: LLM 249agents 148agent 128agentic-ai 113cli 76skills 69anthropic-claude 59machine-learning 58chatgpt 56ai-scientist 55coding-agents 53bioinformatics 44azure 39PyTorch 36

NVIDIA-NeMo/Guardrails

Skill Claude CodeCodex

Helps developers create a NeMo Guardrails configuration for an LLM application. Use when users want to build, scaffold, configure, test, or iterate on input, output, retrieval, dialog, execution, Colang, or catalog-based guardrails. Trigger keywords - create guardrails, build guardrails, scaffold config, write rails…

not rated 7.1k +39 yesterday A 101 tokens

clawrouter

26

BlockRunAI/ClawRouter

Skill Claude CodeCodex

Hosted-gateway LLM router — save 84% on inference costs. A local proxy that forwards each request to the blockrun.ai gateway, which routes to the cheapest capable model across 76 models from OpenAI, Anthropic, Google, DeepSeek, xAI, Z.AI, and more. 7 free open-weight models included. Also exposes realtime market data…

not rated 6.6k +7 changed 4d ago A 222 tokens original MIT

cutedsl_megamoe

27

flashinfer-ai/flashinfer

Skill Claude CodeCodex

Skill "cutedsl_megamoe" from flashinfer-ai/flashinfer, covering updating the cutedsl megamoe kernel src, layout, when the kernel team drops a new version of src/ and what not to update here.

not rated 6.3k +28 today C 0 tokens original Apache-2.0

Tencent/puerts

Skill Claude CodeCodex ✓ vendor

Guide for developing LLM agents based on PuerTsAgent framework — covers resource directory structure, system-prompt, skills, builtin modules, and best practices.

not rated 6.2k +6 today A 39 tokens

autorag-setup

29

Marker-Inc-Korea/AutoRAG

Skill Claude CodeCodex

Install and configure AutoRAG, or repair its single search model, approved document roots, retrieval indexes, datasource skills, and health checks without exposing credentials. Use when autorag is missing, init/refresh/health fails, indexes are stale, or the user wants to add folders or datasources.

not rated 5.1k +3 changed yesterday A 65 tokens

withcoral/coral

Skill Codex

Create or update a Coral source spec YAML for a custom HTTP API or local dataset. Use when authoring a standalone source for coral source add --file, or when adapting that spec into a Coral repo source under sources/core or sources/community.

not rated 5.0k +9 yesterday A 59 tokens original Apache-2.0

varlock

31

dmno-dev/varlock

Skill Claude CodeCodex needs its repo

Secure environment variable management with Varlock. Use when handling secrets, API keys, credentials, or any sensitive configuration. Ensures secrets are never exposed in terminal, logs, or LLM context. Provides guidance around integrating varlock into a project, reading/editing .env.schema and other .env files…

not rated 4.3k +47 yesterday A 109 tokens original MIT

docetl

32

ucbepic/docetl

Skill Claude CodeCodex

Build and run LLM-powered data processing pipelines with DocETL. Use when users say "docetl", want to analyze unstructured data, process documents, extract information, or run ETL tasks on text. Helps with data collection, pipeline creation, execution, and optimization.

not rated 4.1k +4 today A 59 tokens original MIT

graph-designer

33

yifanfeng97/Hyper-Extract

Skill Claude CodeCodex

Design YAML extraction templates for graph types (graph, hypergraph, temporalgraph, spatialgraph). Use when user says: "design graph", "create knowledge graph", "extract relationships", "temporal data", "spatial data". Trigger: User wants to extract entity relationships, multi-party events, or time/location-based…

not rated 3.9k +9 today A 84 tokens

pysr

34

astroautomata/PySR

Skill Claude CodeCodex

Use when fitting equations to data with PySR or SymbolicRegression.jl, when a user wants an interpretable formula, symbolic model, scaling law, or empirical relation discovered from numeric data, or when debugging a PySR search that is slow, stuck, or giving poor equations.

not rated 3.7k +3 yesterday A 61 tokens original Apache-2.0

truera/trulens

Skill Claude CodeCodex

Configure feedback functions and selectors for TruLens evaluations.

not rated 3.5k +5 2d ago A 17 tokens original MIT

seedance-prompt-en

36

dexhunter/seedance2-skill

Skill Claude CodeCodex

Write effective prompts for Jimeng Seedance 2.0 multimodal AI video generation. Use when users want to create video prompts using text, images, videos, and audio inputs with the @ reference system. Covers camera movements, effects replication, video extension, editing, music beat-matching, e-commerce ads, short…

not rated 3.5k +32 6mo ago A 76 tokens original MIT

evalscope

37

modelscope/evalscope

Skill Claude CodeCodex

LLM evaluation & inference performance testing via the evalscope CLI. Translates natural language requests into evalscope commands for: (1) Model accuracy evaluation — runs 160+ benchmarks against local checkpoints or API endpoints (OpenAI-compatible, Anthropic, LiteLLM); (2) Performance stress testing — TTFT, TPOT…

not rated 3.4k +26 5d ago A 197 tokens original Apache-2.0

magnitude

38

magnitudedev/magnitude

Skill Claude CodeCodex

Set up and operate Magnitude local inference through its CLI, recommend hardware-fit local models, monitor acquisition and loading, and connect an agent harness. Use for Magnitude service, model, catalog, setup, or harness requests.

not rated 3.3k +1.7k today A 48 tokens original Apache-2.0

NVIDIA/physicsnemo

Skill Claude CodeCodex ✓ vendor

Official NVIDIA-authored guidance for PhysicsNeMo ShardTensor domain parallelism — integrate domain parallelism into training/inference scripts (new or existing) with DDP or FSDP2, write and register shard patches to enable new layers/ops, and bootstrap multi-GPU correctness tests. Use when working with ShardTensor…

not rated 3.2k +12 yesterday A 152 tokens original Apache-2.0

memmachine-memory

40

MemMachine/MemMachine

Skill Codex

Use when an agent or model needs durable project, user, or session context from MemMachine, needs to save information to MemMachine memory, has requests involving mem-cli, memmachine, or memmachineclient, has insufficient conversation context, or is tempted to search local files for prior context that should come from…

not rated 3.2k +5 yesterday A 96 tokens original Apache-2.0

harbor

41

av/harbor

Skill Claude Code

CLI toolkit for managing containerized LLM services. Use when the user wants to start, stop, configure, or manage AI/LLM services like Ollama, Open WebUI, llama.cpp, vLLM, LiteLLM, ComfyUI, and 250+ others. Triggers on requests to "run a model", "start ollama", "set up an LLM", "configure harbor", "manage services"…

not rated 3.2k +5 7d ago A 114 tokens original Apache-2.0

ai-agent-dev

42

Snailclimb/interview-guide

Skill Claude CodeCodex

A Chinese-language interview-question guide for roles focused on AI-agent development. It covers agent design, language-model calls, tool integration, MCP, retrieval-augmented generation, context management, and multiple agents working together.

not rated 3.2k +30 23d ago A 59 tokens AGPL-3.0

TanStack/ai

Skill Claude CodeCodex needs its repo

Image, audio, video, speech (TTS), and transcription generation using activity-specific adapters: generateImage() with openaiImage/geminiImage/byteplusImage, generateAudio() with geminiAudio/falAudio, generateVideo() with async polling (openaiVideo/geminiVideo/grokVideo/falVideo/byteplusVideo/openRouterVideo…

not rated 3.1k +17 changed 4d ago A 147 tokens original MIT

ax-agent-rlm

44

ax-llm/ax

Skill Claude CodeCodex

This skill helps an LLM generate correct AxAgent RLM/runtime code using @ax-llm/ax. Use when the user asks about RLM code execution, AxJSRuntime, contextFields, contextPolicy, liveRuntimeState, promptLevel, stage prompt controls, executorModelPolicy, maxRuntimeChars, agent.test(...), llmQuery(...), recursionOptions…

not rated 2.9k +2 today A 88 tokens original Apache-2.0

llama-cpp

45

moltis-org/moltis

Skill Claude CodeCodex

Run LLM inference with llama.cpp on CPU, Apple Silicon, AMD/Intel GPUs, or NVIDIA — plus GGUF model conversion and quantization (2–8 bit with K-quants and imatrix). Covers CLI, Python bindings, OpenAI-compatible server, and Ollama/LM Studio integration. Use for edge deployment, M1/M2/M3/M4 Macs, CUDA-less…

not rated 2.8k +6 3d ago A 91 tokens original MIT

ade

46

landing-ai/ade-cli

Skill Claude CodeCodex

Parse documents and extract schema-shaped data with the ADE (Agentic Document Extraction) v2 APIs through the ade CLI. A local job-item store makes every run idempotent, resumable, and citable — repeat runs are free, interrupted runs resume, and every answer can cite element ids with visual evidence.

not rated 2.4k +1 18d ago A 65 tokens original Apache-2.0

prompt-review

47

spacedriveapp/spacebot

Skill Claude CodeCodex

This skill should be used when the user asks to "review the prompt", "audit the system prompt", "check prompt quality", "inspect what the LLM sees", "debug prompt issues", or "find prompt engineering problems". Pulls the live rendered prompt via the API, explains how it's composed, and reviews it for issues.

not rated 2.4k +9 19d ago A 70 tokens

dstack-prototyping

48

dstackai/dstack

Skill Claude CodeCodex

Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven. Guides task-first prototyping on real hardware, choosing fleets/backends that can reuse idle instances and caches, checking vLLM/SGLang sources, and verifying the final…

not rated 2.2k +2 changed 2d ago A 80 tokens original MPL-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: