NVIDIA/TensorRT-Model-Connect

From PyTorch model to end-to-end TensorRT inference experience in two commands—AI-native, cross-platform, and built for the best possible user experience.

About the project

NVIDIA/TensorRT-Model-Connect is a collection of C++ reference implementations for deploying supported Hugging Face models with NVIDIA TensorRT, a system that runs trained models to produce inference results. It is for developers who want to build and run supported models or evaluate model integrations through TensorRT. The catalogue entries provide skills and instructions for using this model deployment workflow.

Latest release trtmc-nightly-20260814T132928Z-75e25af75499 — TRTMC Nightly 75e25af75499 · 14 Aug 2026

These files are NVIDIA/TensorRT-Model-Connect's own configuration. They tell Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.

229Stars on the repository
1Files it configures its agents with
834Tokens loaded in every session
2Agents configured

Instructions