Overview
EXTENSOR is an inference software package purpose-built for the AMD Ryzen™ AI Max platform that enables deployment of large language models and mixture-of-experts (MoE) models on local AI systems, bringing frontier-class AI capabilities to a broader range of developer, workstation, and edge form factors.
The innovation is captured in the name: EXTENSOR keeps each model's mixture-of-experts weights on SSD and uses adaptive, predictive prefetching to stream the right experts into memory just before they're needed. This approach significantly reduces memory requirements while maintaining high-quality model performance, enabling larger and more capable AI models to run on a single Ryzen AI Max system.
Downloads
EXTENSOR is launching with support for two frontier models out of the box:
| Model Name | Size | Checksum |
| gfx1151 | 7.44 MB | dc2935700681fd55d3a4444457db0868 |
| gfx1201 | 7.35 MB | 91a70136bed06aefbc71e229abf0ed38 |
| gfx1152 | 7.41 MB | 9e214afebe0f3f10cb210156d4f242c |
EXTENSOR's expert-offload and prefetch engine is unlocking an entirely new class of experiences without sacrificing usability.