Posts by HongWei Meng

Local Quantization and Multi-Backend Deployment with AMD Quark on Strix Halo

For local AI, running a model on the target device is only part of the deployment workflow. Model preparation, including quantization and export, is another important step. AMD Quark provides memory-efficient quantization workflows that make it possible to optimize large models directly on platforms such as AMD Strix Halo.

Read more ...