Overview
EXTENSOR is an inference software package purpose-built for the AMD Ryzen™ AI Max platform that enables deployment of large language models and mixture-of-experts (MoE) models on local AI systems, bringing frontier-class AI capabilities to a broader range of developer, workstation, and edge form factors.
The innovation is captured in the name: EXTENSOR keeps each model's mixture-of-experts weights on SSD and uses adaptive, predictive prefetching to stream the right experts into memory just before they're needed. This approach significantly reduces memory requirements while maintaining high-quality model performance, enabling larger and more capable AI models to run on a single Ryzen AI Max system.
Downloads
EXTENSOR is launching with support for two frontier models out of the box:
| Model Name | Description | Support OS | Release Date |
| DSV4 Flash | Stronger Agent capabilities and top-tier reasoning | Linux | 7/22/2026 |
| GLM 5.2 | Advanced, multi-effort coding capabilities | Linux | 7/22/2026 |
Both models exceed available system memory on Strix Halo; EXTENSOR's expert-offload and prefetch engine is unlocking an entirely new class of experiences without sacrificing usability.