moe agents

15 tagged moe, measured the same way as everything else here.

Browse within: PyTorch 15blackwell 15cuda 15llm-serving 15

ad-onboard-reviewer

01

NVIDIA/TensorRT-LLM

Agent Claude Code

Independent reviewer for AutoDeploy model onboarding. Validates created model and test files against all onboarding requirements. Use after completing model onboarding work.

15k 2d ago A 33 tokens

NVIDIA/TensorRT-LLM

Agent Claude Code

Expert in GPU performance profiling for TRT-LLM workloads with nvidia-smi, Nsight Systems (nsys), Nsight Compute (ncu), and PyTorch profiler. This agent can execute shell commands directly. Delegate to this agent for: (1) Running workloads (Python scripts, CUDA binaries, shell commands), (2) Measuring performance…

15k 2d ago A 152 tokens

perf-test-sync

03

NVIDIA/TensorRT-LLM

Agent Claude Code

Use this agent when the user needs to synchronize performance test cases between development (dev) and QA directories, compare test configurations, update test lists, or analyze gaps between dev and QA perf test coverage. This includes syncing aggregated and disaggregated performance test cases, updating QA test lists…

15k 2d ago A 334 tokens