redai-infra/Relax

An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

580Stars on the repository
22Mods indexed here, across every type
5d agoLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

algorithm-expert

01

redai-infra/Relax

Agent

RL algorithm expert. Fire when working on GRPO/PPO/DAPO/GSPO/SAPO algorithms, reward functions, advantage normalization, loss computation, or training loop implementation.

580 5d ago A 37 tokens original Apache-2.0

fsdp-expert

02

redai-infra/Relax

Agent

FSDP backend expert. Fire when working on FSDP-based training, parameter sharding, FSDP weight update, CPU offloading, or troubleshooting FSDP-related issues.

580 5d ago A 38 tokens original Apache-2.0

launcher-expert

03

redai-infra/Relax

Agent

Ray orchestration & service deployment expert. Fire when working on Ray Serve deployment, placement groups, service lifecycle, rollout engine management, health monitoring, or troubleshooting job launch and GPU allocation issues.

580 5d ago A 38 tokens original Apache-2.0

megatron-expert

04

redai-infra/Relax

Agent

An expert assistant for integrating and configuring Megatron, a system for training very large machine-learning models across multiple processors or machines.

580 5d ago A 29 tokens original Apache-2.0

ray-expert

05

redai-infra/Relax

Agent

Ray framework expert. Fire when working on Ray cluster management, ray.init/ray.remote/ray.get patterns, placement groups, scheduling strategies, Ray Serve deployments, ray job submit, runtime environments, or troubleshooting Ray-specific errors (serialization, object store, GCS, scheduling failures).

580 5d ago A 56 tokens original Apache-2.0