Skip to main content

Enterprise AI Inference at Scale with AMD

abstract background

Abstract

Enterprise AI is scaling quickly in production, requiring infrastructure that optimizes performance, governance, simplicity, and cost. In this session we show how the AMD Enterprise AI reference stack enables heterogenous inference across on-premises and hosted environments using a vLLM-based Semantic Router to intelligently direct workloads across AMD Instinct™ MI350P, MI350X, and hosted models. Learn how enterprises can streamline deployment, strengthen AI governance, and scale AI responsibly.

July 22, 2026 2:00 PM - 2:45 PM PDT

Speakers


Presented By