Posts by Xiaohong Kou

veRL on AMD: Production-Ready RL Post-Training on ROCm

Reinforcement learning post-training on AMD Instinct GPUs is here — with a turnkey container, AITER-accelerated vLLM and SGLang rollout, and accuracy validated on both MI300 and MI355.

Read more ...