Mesh-LLM

32 mods across 1 repository, 3.3k stars between them.

release-validation

01

Mesh-LLM/mesh-llm

Agent Codex

Validates the next MeshLLM release against the last GitHub release on approved real hosts and produces a formal evidence-backed readiness report.

3.3k yesterday A 30 tokens original Apache-2.0

benchmark-tune

02

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when running, debugging, interpreting, or documenting mesh-llm benchmark tune model-serving throughput trials, including choosing ctx/batch/ubatch/mmap/mlock/speculative-decoding sweeps, running benchmark tune on local or SSH hosts, collecting JSON evidence, and applying tolerance-aware recommendations.…

3.3k yesterday A 106 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when adding, renaming, removing, validating, or exposing mesh-llm config settings, including built-in settings, plugin config schemas, owner-control apply behavior, CLI validation, and UI configuration surfaces.

3.3k yesterday A 48 tokens original Apache-2.0

connect-agents

04

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when connecting agent tools or OpenAI clients to mesh-llm — launching or configuring Goose, Claude Code, OpenCode, Pi, curl, or any OpenAI-compatible client against a local or remote mesh, picking a model, or validating tool-call reliability.

3.3k yesterday A 59 tokens original Apache-2.0

deploy-linux-gpu

05

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when deploying, installing, launching, or serving mesh-llm on a remote Linux GPU node (rented GPUs like Vast.ai or RunPod, or a self-managed CUDA server), including installing the CUDA build, choosing a model, keeping it alive under a supervisor, and verifying it serves.

3.3k yesterday D 67 tokens original Apache-2.0

deploy-macos

06

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when deploying, installing, launching, or serving mesh-llm on a macOS machine (local or remote over SSH), including installing a release, shipping a dev build bundle, codesign/quarantine fixes, choosing a model, and verifying it serves.

3.3k yesterday E 58 tokens original Apache-2.0

deploy-windows

07

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when installing, deploying, launching, serving, or troubleshooting mesh-llm on a Windows machine — PowerShell install via install.ps1, flavor selection (CUDA/ROCm/Vulkan/CPU), source builds, the contrib helper scripts, and verifying it serves.

3.3k yesterday A 60 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the artifact to Hugging Face.

3.3k yesterday A 55 tokens original Apache-2.0

hf-gguf-quant-jobs

09

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use when creating, monitoring, validating, or documenting low-memory Hugging Face Jobs or local runs that quantize split BF16/FP16 GGUF model repos into custom quant GGUF repos with skippy-quantize.

3.3k yesterday A 54 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use when changing mesh-llm automation or CLI flows that discover Hugging Face GGUF models, plan CPU Hugging Face Jobs for layer-package splitting, estimate max cost, or publish skippy layer packages/catalog entries.

3.3k yesterday A 50 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use when running quantization of a BF16/FP16 GGUF repo and Skippy layer-package creation as one local or Hugging Face Jobs workflow, publishing both artifacts to Hugging Face.

3.3k yesterday A 48 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction failures.

3.3k yesterday A 49 tokens original Apache-2.0

llama-patch-changes

13

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use when changing mesh-llm's llama.cpp patch queue, upstream pin, prepare/build scripts, or carried RPC, MoE, and mesh-hook llama.cpp patches.

3.3k yesterday C 41 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or patch queue.

3.3k yesterday C 51 tokens original Apache-2.0

manage-ci

15

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill as the mandatory starting point whenever inspecting, running, debugging, defining, editing, reviewing, or documenting MeshLLM CI/CD. It governs GitHub Actions workflows and local actions, triggers and routing, runners, caches, artifacts, permissions, releases, deployments, and CI infrastructure.

3.3k yesterday C 62 tokens original Apache-2.0

mesh-join

16

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when creating, joining, publishing, or connecting mesh-llm nodes into a mesh — private meshes with invite tokens, the public mesh via --auto, named/published meshes, client-only nodes, NAT/firewall/bind issues, or verifying multi-node setups.

3.3k yesterday C 60 tokens original Apache-2.0

metrics-server

17

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when working on benchmark telemetry ingest, metrics-server run lifecycle, OTLP collection, SQLite storage, benchmark report export, or separating telemetry/reporting ownership from staged runtime servers.

3.3k yesterday A 40 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when maintaining the plugin web UI projection contract, docs, exemplar coverage, or recovery flow for mesh-llm plugin web UI work.

3.3k yesterday A 35 tokens original Apache-2.0

release-validation

19

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when validating a MeshLLM release candidate or current HEAD against the last GitHub release, assembling the canonical feature/fix/modification inventory, testing locally built release bundles on user-approved real hosts and private meshes, deciding release readiness, or producing a formal…

3.3k yesterday A 62 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when starting, supervising, debugging, holding open, or stopping any remote process over SSH that needs an operator-like interactive environment, a TTY, login-shell startup files, long-running observation, logs, readiness checks, or later inspection.

3.3k yesterday A 55 tokens original Apache-2.0

skippy-bench

21

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when running benchmark orchestration, local single-stage or split benchmarks, benchmark report flow, or performance-oriented skippy runtime checks.

3.3k yesterday A 33 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when benchmarking Skippy exact-prefix cache across model families, comparing Skippy against llama-server, producing README benchmark tables, updating crates/skippy-cache/README.md evidence, or diagnosing cache benchmark gaps by family or Hugging Face use case.

3.3k yesterday A 57 tokens original Apache-2.0

skippy-correctness

23

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when validating skippy staged execution against full-model execution, adding model families, changing split boundaries, testing activation wire dtypes, or diagnosing mismatch behavior.

3.3k yesterday A 39 tokens original Apache-2.0

Mesh-LLM/mesh-llm

Skill Claude CodeCodex

Use this skill when certifying a GGUF model family for skippy stage-split serving, reviewing capability data, promoting family evidence into topology policy, or updating staged split certification docs.

3.3k yesterday A 43 tokens original Apache-2.0