flashinfer-ai/flashinfer

FlashInfer: Kernel Library for LLM Serving

6.3kStars on the repository
7Mods indexed here, across every type
yesterdayLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

add-cuda-kernel

01

flashinfer-ai/flashinfer

Skill Claude CodeCodex

Step-by-step tutorial for adding new CUDA kernels to FlashInfer.

6.3k +21 yesterday A 18 tokens original Apache-2.0

cutedsl_megamoe

04

flashinfer-ai/flashinfer

Skill Claude CodeCodex

Skill "cutedsl_megamoe" from flashinfer-ai/flashinfer, covering updating the cutedsl megamoe kernel src, layout, when the kernel team drops a new version of src/ and what not to update here.

6.3k +21 yesterday C 0 tokens original Apache-2.0

flashinfer-ai/flashinfer

Skill Claude CodeCodex

This tree vendors the kernel team's SM90 FP8 MegaMoE drop — a fork of the same kernel repo that kernelsrc/cutedslmegamoe vendors (Bangyu's SM100 tree). The SM90 work (Vincent's hoppermegamoe branch) moved the shared runtime forward, so this tree duplicates common/, src/, and moenvfp4swapab/ at its own revision instead…

6.3k +21 yesterday C 0 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: