NVIDIA-NeMo/Nemotron

Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models

2.0kStars on the repository
12Mods indexed here, across every type
yesterdayLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Prepare, validate, build, and use Nemotron Customizer airgap image bundles for offline clusters. Use when planning airgapped deployments, editing deploy/nemotron-customizer/airgap/airgap.yaml, selecting workflow targets, grouping step execution images, baking repo overlays or wheel additions, resuming airgap runner…

not rated 2.0k +13 yesterday A 89 tokens original Apache-2.0

nemotron-add-model

02

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Onboard a new model family (Nemotron or third-party) into skills/ — paper chunks, recipe summaries, context packs, and model card. Use when a contributor wants downstream skills like /nemotron-customize to be able to route to a new model.

not rated 2.0k +13 yesterday A 58 tokens original Apache-2.0

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Add a cross-cutting decision pattern under src/nemotron/steps/patterns/. Use when a recurring ML decision (tokenizer lock, eval bookends, LoRA-on-small-data, etc.) must be encoded so other skills can fire it during planning.

not rated 2.0k +13 yesterday A 58 tokens original Apache-2.0

nemotron-add-step

04

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Add a new step under src/nemotron/steps/ / / — manifest (step.toml), runner glue, configs, and per-step README.md. Use when extending the catalog so /nemotron-customize can route to it.

not rated 2.0k +13 yesterday A 56 tokens original Apache-2.0

nemotron-customize

05

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Part of nemotron-customize

Plan, configure, and chain repo-native Nemotron customization steps into single-step or multi-step pipelines: curation, translation, SFT/PEFT (AutoModel or Megatron-Bridge), pretraining/CPT, RL alignment (DPO/RLVR/GRPO/RLHF), BYOB/MCQ benchmarks, checkpoint conversion, ModelOpt optimization, env profiles, and…

not rated 2.0k +13 yesterday A 159 tokens original Apache-2.0

nemotron-nano3

06

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Reference desk for Nemotron 3 Nano / Llama-Nemotron Nano 3 — architecture, training data, recipes, evaluation, quantization, deployment. Use when the user asks facts about the model rather than building a pipeline.

not rated 2.0k +13 yesterday A 53 tokens original Apache-2.0

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Generates BYO custom safety policies for NVIDIA Nemotron content-safety guardrails — Nemotron-Content-Safety-Reasoning-4B (text) and multimodal Nemotron-3-Content-Safety. Produces a Markdown policy, JSON taxonomy, and drop-in inference prompts. Maps rough words or an existing policy to V2 categories, adding custom…

not rated 2.0k +13 yesterday A 85 tokens original Apache-2.0

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Use when planning, debugging, tuning, evaluating, exporting, or deploying public Nemotron embed/rerank retrieval recipes.

not rated 2.0k +13 yesterday A 36 tokens original Apache-2.0

nemotron-super3

09

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Reference desk for NVIDIA Nemotron 3 Super — architecture, training data, recipes (pretrain/SFT/RL/eval/quantization), and deployment notes. Use when the user asks facts about Super3 rather than building a pipeline.

not rated 2.0k +13 yesterday A 53 tokens original Apache-2.0

nemotron-ultra

10

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference. Use when the user asks facts about Ultra rather than building a pipeline.

not rated 2.0k +13 yesterday A 68 tokens original Apache-2.0

NVIDIA-NeMo/Nemotron

Skill Claude CodeCodex

Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA fine-tuning of the 550B hybrid Mamba-Transformer MoE, ending at a saved adapter. Use when the user wants to run this cookbook…

not rated 2.0k +13 yesterday A 113 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: