Evaluate and improve models and agents using environments
NVIDIA-NeMo/Gym is a library and infrastructure for evaluating and improving models and agents inside environments, where each environment defines tasks, agent interaction, verification, and execution state. It is for teams running reproducible evaluations or training at scale across settings such as code execution, tool calling, and sandboxes, and the catalogue entries provide skills and instructions for working with it.
Latest release v0.5.1 — NVIDIA NeMo-Gym 0.5.1 · 3 Sept 2026
These files are NVIDIA-NeMo/Gym's own configuration. They tell Claude Code, Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.
AGENTS.md A 2,289 tok CLAUDE.md A 3 tok .agents/skills/add-benchmark/SKILL.md A 120 tok .agents/skills/gh-stack/SKILL.md A 65 tok .agents/skills/nemo-gym-blade-analysis/SKILL.md A 106 tok .agents/skills/nemo-gym-debugging/SKILL.md A 68 tok .agents/skills/nemo-gym-docs/SKILL.md A 80 tok .agents/skills/nemo-gym-pivot-datasets/SKILL.md A 97 tok .agents/skills/nemo-gym-reward-profiling/SKILL.md A 78 tok