Uni-Agent: Train Long-Horizon Agents at Scale
Uni-Agent is an open framework for scalable long-horizon agent RL. We share our insights on harness adaptation, RL training, and practical recipes for agentic RL.
Tagged
4 posts
Uni-Agent is an open framework for scalable long-horizon agent RL. We share our insights on harness adaptation, RL training, and practical recipes for agentic RL.
A complete GRPO workflow using verl, Qwen3-30B-A3B, SGLang, and PyTorch FSDP ran for 100 consecutive steps on four AMD Instinct MI455X GPUs, establishing an end-to-end RL baseline for CDNA 5.
A release focused on higher-throughput diffusion rollout, reusable omni adapters, and broader recipe coverage.
Keep the Tinker Cookbook loop you know, and run SFT, RL, and distillation on verl-managed GPU workers you control.