Posts by Tanya Roosta
Hyperloom: A Multi-Agent Harness for Autonomous Inference Optimization on AMD GPUs
- 21 September 2026
ROCm™ Hyperloom delivered a median 1.73× inference speedup in extensive unattended evaluation on AMD Instinct™ GPUs, with gains ranging from 1.35× to 7.31×. It profiles each workload, searches framework and kernel optimizations, validates changes end to end, and carries proven results forward, without per-model human tuning.
Hyperloom - Autonomous Agentic Inference Optimization for AMD GPUs
- 23 July 2026
AMD is excited to introduce ROCm™ Hyperloom, a new open-source, agentic system aimed at automating the time-consuming task of optimizing end-to-end inference workloads. Using Hyperloom reduces the time to optimize an end-to-end workload from weeks to hours, saving time and enabling valuable resources to be allocated to other critical tasks. By combining various tools into an autonomous optimization loop, Hyperloom allows you to get the best performance out of your model and custom configuration on AMD Instinct GPUs.