AMD Ryzen AI MAX Series

Overview

EXTENSOR is an inference software package purpose-built for the AMD Ryzen™ AI Max platform that enables deployment of large language models and mixture-of-experts (MoE) models on local AI systems, bringing frontier-class AI capabilities to a broader range of developer, workstation, and edge form factors.

The innovation is captured in the name: EXTENSOR keeps each model's mixture-of-experts weights on SSD and uses adaptive, predictive prefetching to stream the right experts into memory just before they're needed. This approach significantly reduces memory requirements while maintaining high-quality model performance, enabling larger and more capable AI models to run on a single Ryzen AI Max system.

Downloads

EXTENSOR is launching with support for two frontier models out of the box:

Model Name Size Checksum
gfx1151 7.44 MB dc2935700681fd55d3a4444457db0868
gfx1201 7.35 MB 91a70136bed06aefbc71e229abf0ed38
gfx1152 7.41 MB 9e214afebe0f3f10cb210156d4f242c

EXTENSOR's expert-offload and prefetch engine is unlocking an entirely new class of experiences without sacrificing usability.