little-planet/ascend-tune-lab

A repository for deploying agents/skills for parameter tuning, profiling and analysis.

2Stars on the repository
34Mods indexed here, across every type
22d agoLast push, which is what freshness is scored on
noneNo LICENSE: all rights reserved, so bodies are not copied

little-planet/ascend-tune-lab

Skill Claude CodeCodex

Estimates max concurrency under TTFT/TPOT SLO for arbitrary vLLM-Ascend models (Qwen any size, GLM, DeepSeek, MiniMax, …) by analyzing vllm-ascend attention dispatch (usemla/usesparse) and msmodeling profilingdatabase. Requires local clones of both repos. Use as sub-skill of serving-parallel-strategy-tuning.

not rated 2 22d ago A 90 tokens

little-planet/ascend-tune-lab

Skill Claude CodeCodex

Extract and compare configuration switches between vLLM and vLLM-Ascend repos. Invoke when user needs to audit, compare, or document config options across vLLM and vLLM-Ascend.

not rated 2 22d ago A 53 tokens

vllm-ascend-tuning

27

little-planet/ascend-tune-lab

Skill Claude CodeCodex

A guide to tuning vLLM-Ascend, a system for running language models on Ascend hardware. It covers standalone use, pipeline use, and quantization tuning, which reduces model precision to change performance or resource use.

not rated 2 22d ago B 147 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: