NVIDIA-TAO

60 mods across 1 repository, 86 stars between them.

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Deformable DETR for 2D object detection. Uses deformable attention for efficient multi-scale feature processing, lighter than DINO with competitive accuracy. Use when training, evaluating, exporting, quantizing, or running inference for a TAO Deformable-DETR model. Trigger phrases include "train deformable-detr"…

86 3d ago A 92 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Monocular depth estimation using Metric Depth Anything v2 or Relative Depth Anything architectures. Predicts per-pixel depth from single RGB images. Use when training, evaluating, exporting, or running inference for a TAO monocular depth model. Trigger phrases include "train monocular depth", "DepthAnything v2"…

86 3d ago A 85 tokens original Apache-2.0

tao-train-dino

51

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

DINO (DETR with Improved DeNoising Anchor Boxes) for 2D object detection. Transformer-based detector with denoising training, multi-scale features, and optional distillation support. Use when training, evaluating, exporting, distilling, quantizing, or running inference for a TAO DINO detector. Trigger phrases include…

86 3d ago A 100 tokens original Apache-2.0

tao-train-dinov3

52

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

DINOv3 continual self-supervised pre-training. Domain-adapts public DINOv3 ViT backbones on unlabeled images via teacher-student self-distillation (DINO + iBOT + KoLeo, optional Gram anchoring) and converts the EMA teacher into a timm-format backbone for downstream tasks. Trigger phrases include "train DINOv3"…

86 3d ago A 117 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Real-time stereo depth estimation using FastFoundationStereo (FFS), the distilled bp2 commercial variant of FoundationStereo. Predicts disparity maps from stereo image pairs with 10× lower latency than full FoundationStereo. Use when training, evaluating, exporting, or running inference for a TAO FastFoundationStereo…

86 3d ago A 101 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Stereo depth estimation using FoundationStereo. Predicts disparity maps from stereo image pairs for 3D reconstruction. Use when training, evaluating, exporting, or running inference for a TAO FoundationStereo model. Trigger phrases include "train stereo depth", "FoundationStereo", "stereo disparity estimation", "3D…

86 3d ago A 74 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Grounding DINO for open-set object detection. Combines DINO-style detection with a BERT text encoder for language-guided detection — detects objects described by text prompts without a fixed class vocabulary. Use when training, evaluating, exporting, quantizing, or running inference for a TAO Grounding DINO model.…

86 3d ago A 103 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

PyTorch-based TAO image classification. Supports a wide range of backbones (FAN, EfficientNet, ResNet, etc.) with distillation and quantization for deployment. Use when training, evaluating, distilling, quantizing, exporting, or running inference for a TAO image-classification (PyT) model. Trigger phrases include…

86 3d ago A 103 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Masked Auto-Encoder (MAE) for self-supervised pretraining and fine-tuning. Masks random patches and reconstructs them to learn visual representations; supports pretrain and finetune stages. Use when training, evaluating, exporting, or running inference for a TAO MAE backbone. Trigger phrases include "pretrain MAE"…

86 3d ago A 103 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

MAL (Mask Auto-Label) for weakly-supervised segmentation. Produces segmentation masks from minimal annotations (point or box annotations) using a ViT-MAE backbone. Use when training, evaluating, or running inference for a TAO MAL model. Trigger phrases include "train MAL", "Mask Auto-Label", "weakly-supervised…

86 3d ago A 94 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Mask Grounding DINO for grounded instance segmentation. Extends Grounding DINO with a mask-prediction head for open-set segmentation guided by text prompts. Use when training, evaluating, exporting, quantizing, or running inference for a TAO Mask-Grounding-DINO model. Trigger phrases include "train Mask Grounding…

86 3d ago A 100 tokens original Apache-2.0

NVIDIA-TAO/tao-skill-bank

Skill Claude CodeCodex

Mask2Former for universal image segmentation (panoptic, instance, and semantic). Transformer-based with masked attention for high-quality segmentation results. Use when training, evaluating, exporting, quantizing, or running inference for a TAO Mask2Former model. Trigger phrases include "train Mask2Former", "universal…

86 3d ago A 89 tokens original Apache-2.0