reinforcement learning instructions

16 tagged reinforcement learning, measured the same way as everything else here.

Gym AGENTS.md

01

NVIDIA-NeMo/Gym

Instructions file CodexOpenCode

Instructions for NVIDIA-NeMo/Gym, covering agents.md, quality bar, what this is, architecture and creating environments.

1.1k 2d ago A 2,289 tokens original Apache-2.0

Gym CLAUDE.md

02

NVIDIA-NeMo/Gym

Instructions file

Instructions for NVIDIA-NeMo/Gym, a project described as: Evaluate and improve models and agents using environments.

1.1k 2d ago A 3 tokens copy · 100% Apache-2.0

UniRL AGENTS.md

03

Tencent-Hunyuan/UniRL

Instructions file CodexOpenCode

AGENTS.md instructions for Tencent-Hunyuan/UniRL, a project described as: UniRL is a Framework for Unified Multimodal Model Reinforcement Learning.

921 4d ago A 11 tokens

UniRL CLAUDE.md

04

Tencent-Hunyuan/UniRL

Instructions file

Claude Code instructions for Tencent-Hunyuan/UniRL, covering claude.md, 1. think before coding, 2. simplicity first, 3. surgical changes and 4. goal-driven execution.

921 4d ago A 1,414 tokens

mobilegym AGENTS.md

05

Purewhiter/mobilegym

Instructions file CodexOpenCode

Instructions for Purewhiter/mobilegym, covering agents.md, project overview, type-checking strategy, eslint and with data graph generation.

777 4d ago A 6,877 tokens original Apache-2.0

utilForever/RosettaStone

Instructions file CodexOpenCode

AGENTS.md instructions for utilForever/RosettaStone, covering agents.md, what this repository is, golden rules (read before any change), common task flow and hearthstone model.

679 14d ago A 3,108 tokens AGPL-3.0

mjswan AGENTS.md

08

ttktjmt/mjswan

Instructions file CodexOpenCode

AGENTS.md instructions for ttktjmt/mjswan, covering agents.md, project, philosophy, layout and python workflow.

325 2d ago A 464 tokens Apache-2.0

MontrealAI/AGI-Alpha-Agent-v0

Instructions file CodexOpenCode

META‑AGENTIC α‑AGI 👁️✨ — Mission 🎯 End‑to‑end: Identify 🔍 → Out‑Learn 📚 → Out‑Think 🧠 → Out‑Design 🎨 → Out‑Strategise ♟️ → Out‑Execute ⚡.

293 4mo ago A 6,566 tokens original Apache-2.0

utilForever/baba-is-auto

Instructions file CodexOpenCode

Instructions for utilForever/baba-is-auto, covering agents.md, what this repository is, golden rules (read before any change), the most common task: changing simulator behavior and repository map.

190 9d ago A 1,650 tokens original MIT

omnisim AGENTS.md

13

omnilink-tech/omnisim

Instructions file CodexOpenCode

Instructions for omnilink-tech/omnisim, covering agents.md — running omnisim as an ai coding agent, 0. for agents: read this first (60 seconds), run this first turn, first moves by task type and hard-won rules (don't relearn these).

82 2d ago B 49,325 tokens original Apache-2.0

omnisim CLAUDE.md

14

omnilink-tech/omnisim

Instructions file

Instructions for omnilink-tech/omnisim, a project described as: Open-source robotics simulator for coding agents: HTTP/JSON + MCP control, Newton physics, wgpu rendering, ROS 2, and reproducible benchmarks.

82 2d ago C 136 tokens original Apache-2.0

Oshawott324/datalox-gated-runtime

Instructions file CodexOpenCode

AGENTS.md instructions for Oshawott324/datalox-gated-runtime, covering agent instructions, product boundary, provider behavior, runtime safety and public data boundary.

0 yesterday A 1,159 tokens original Apache-2.0