Megatron-LM skills

11 tagged Megatron-LM, measured the same way as everything else here.

Browse within: SGLang 11reinforcement-learning 11FP8 6INT4 6enterprise 6moe 6GLM 5GRPO 5Post-Training 5

miles-rl-training

01

davila7/claude-code-templates

Skill Claude CodeCodex

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime. Use when training large MoE models with FP8/INT4, needing train-inference alignment, or requiring speculative RL for maximum throughput.

not rated 31k today A 51 tokens MIT

slime-rl-training

02

davila7/claude-code-templates

Skill Claude CodeCodex

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

not rated 31k today A 52 tokens MIT

slime-rl-training

03

synthetic-sciences/openscience

Skill Claude CodeCodex

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

not rated 3.4k today A 52 tokens Apache-2.0

miles-rl-training

04

synthetic-sciences/openscience

Skill Claude CodeCodex

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime. Use when training large MoE models with FP8/INT4, needing train-inference alignment, or requiring speculative RL for maximum throughput.

not rated 3.4k today A 51 tokens Apache-2.0

miles-rl-training

05

OpenLAIR/dr-claw

Skill Claude CodeCodex

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime. Use when training large MoE models with FP8/INT4, needing train-inference alignment, or requiring speculative RL for maximum throughput.

not rated 1.1k 3d ago A 51 tokens

slime-rl-training

06

OpenLAIR/dr-claw

Skill Claude CodeCodex

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

not rated 1.1k 3d ago A 52 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: