NVIDIA/TensorRT-Model-Connect

From PyTorch model to end-to-end TensorRT inference experience in two commands—AI-native, cross-platform, and built for the best possible user experience.

188Stars on the repository
12Mods indexed here, across every type
yesterdayLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

debug-trt-mismatch

01

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when TensorRT output diverges from a model reference, model-first validation fails, generated text or media is wrong, or a family change introduces a numerical mismatch. Routes the investigation by model modality and escalates from the first divergent boundary to the smallest responsible family-owned operation.

188 yesterday A 61 tokens original Apache-2.0

doc-sync

02

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use for documentation maintenance scans that keep the canonical website journey, repo-local skills, commands, API reference, architecture and design, extension guides, feature context, ADRs, and traceability status aligned with the current GitHub main branch. Covers the Source/Internal CI boundary and model-owned…

188 yesterday A 65 tokens original Apache-2.0

fp16-trt-network

03

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when adding, reviewing, or debugging FP16/BF16 precision in a family-owned, strongly typed TensorRT network. Covers dtype threading, explicit FP32 boundaries, typed constants, compact GQA/MQA state, bundle evidence, and low-precision validation.

188 yesterday A 59 tokens original Apache-2.0

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when evaluating FP16, BF16, or supported quantization formats for a TensorRT-Model-Connect model. Establishes a model-owned correctness baseline, changes one effective build option at a time, detects ineffective precision, and retains comparable parity, memory, bundle, and performance evidence.

188 yesterday A 64 tokens original Apache-2.0

pr-babysitter

05

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when monitoring GitHub pull request CI, diagnosing failed checks, rebasing branches onto github/main, applying narrowly scoped fixes, and updating PRs until their latest checks are green or a human blocker is identified.

188 yesterday A 48 tokens original Apache-2.0

profile-model

06

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when diagnosing one model's runtime cost or producing comparable TensorRT-Model-Connect performance evidence. Routes quick investigation to the unified profiler and release or qualification claims to the checked-in performance matrix and model-owned performance contract.

188 yesterday A 47 tokens original Apache-2.0

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Prepare a TensorRT-Model-Connect development or deployment-validation environment from a fresh checkout on an unfamiliar host. Use before builds, tests, packaging, or runtime work when no working repo environment is known.

188 yesterday A 48 tokens original Apache-2.0

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when converting QA findings, black-box failures, red-team reports, regression evidence, or local bug notes into GitHub Issues for NVIDIA/TensorRT-Model-Connect. Standardizes checking issue templates, checking labels, de-duplicating existing issues, drafting a bug report, creating the issue on GitHub, applying the…

188 yesterday A 82 tokens original Apache-2.0

submit-github-pr

09

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when publishing an existing TensorRT-Model-Connect change as a GitHub pull request. Verifies authenticated repository access, branch and diff scope, validation evidence, commit identity, reviewer-facing text, exact pushed head, and the created draft PR without merging it.

188 yesterday A 58 tokens original Apache-2.0

transform-model

10

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Use when onboarding a Hugging Face model into TensorRT-Model-Connect or extending an existing family to produce a .bundle bundle. Drives ownership-first implementation across Python builder, native runtime, and model-owned E2E descriptors, then requires reference-consistency and runtime evidence before support is…

188 yesterday A 63 tokens original Apache-2.0

write-git-messages

11

NVIDIA/TensorRT-Model-Connect

Skill Claude CodeCodex

Draft, revise, or review Git commit messages, PR titles, PR descriptions, and squash or rebase merge messages. Use when Codex needs to summarize a diff for reviewers, convert rough notes into a commit or PR message, check a message against Git and Conventional Commits style, or prepare repository contribution text…

188 yesterday A 75 tokens original Apache-2.0