Uni-Agent: Train Long-Horizon Agents at Scale
Uni-Agent is an open framework for scalable long-horizon agent RL. We share our insights on harness adaptation, RL training, and practical recipes for agentic RL.
Posts by
3 posts
Uni-Agent is an open framework for scalable long-horizon agent RL. We share our insights on harness adaptation, RL training, and practical recipes for agentic RL.
A complete GRPO workflow using verl, Qwen3-30B-A3B, SGLang, and PyTorch FSDP ran for 100 consecutive steps on four AMD Instinct MI455X GPUs, establishing an end-to-end RL baseline for CDNA 5.
Keep the Tinker Cookbook loop you know, and run SFT, RL, and distillation on verl-managed GPU workers you control.