Posts by Juergen Frick
Spur: Modern GPU Job Scheduling for HPC and AI Workloads
- 22 July 2026
The explosion of AI and large-scale model training has fundamentally changed how organizations think about GPU clusters. Traditional HPC schedulers were built in an era when CPUs dominated, and GPU support was retrofitted as an afterthought. Meanwhile, Kubernetes emerged from the cloud-native world with powerful orchestration primitives—but batch GPU scheduling and Slurm-style job workflows remain friction points unless you assemble a full ecosystem of operators, custom schedulers, and queueing layers on top of the default control plane.
Introducing ROCm™ AMD Infinity Context: A Purpose-Built KV Cache Tier for Distributed Inference
- 22 July 2026
In this blog, you will learn how ROCm™ AMD Infinity Context (ROCm AIC) addresses one of the fastest-growing bottlenecks in production AI systems: key-value (KV) cache management. You will explore the problem, the solution, the components that make up the stack, and how ROCm AIC works with existing AMD Instinct™ GPU deployments.