Posts by Xiaofei Zheng
Hyperloom: A Multi-Agent Harness for Autonomous Inference Optimization on AMD GPUs
- 21 September 2026
ROCm™ Hyperloom delivered a median 1.73× inference speedup in extensive unattended evaluation on AMD Instinct™ GPUs, with gains ranging from 1.35× to 7.31×. It profiles each workload, searches framework and kernel optimizations, validates changes end to end, and carries proven results forward, without per-model human tuning.