The Token Tsunami: Infrastructure Bets That Win in 2026
Abstract
The agents have arrived and they're hungry. Every token has a price, and agentic AI is sending the bill soaring. This session decodes the token tsunami: where agentic demand is headed, how costs shift across model families and tasks, and how AMD delivers leadership performance and TCO at every tier, closing with strategic guidance from Chirag Dekate, VP Analyst for AI infrastructure from Gartner®. Make the infrastructure bets that win in 2026. Don't just ride the wave - provision for it.
GARTNER is a trademark of Gartner, Inc. and/or its affiliates.
July 22, 2026 3:00 PM - 3:45 PM PDT
Speakers
Presented By
VP Analyst | Gartner
Director of Marketing Intelligence Datacenter GPU | AMD
Sr. Director, Product Marketing for Datacenter GPU | AMD
Related Sessions
-
Supercomputing for All: Bringing Exascale Innovation to Enterprise AI
Supercomputing for All: Bringing Exascale Innovation to Enterprise AI
HPE and AMD have built some of the world's most advanced supercomputers, including Frontier and El Capitan. Join HPE leaders who helped deploy these systems and learn how the technologies, expertise, and operational practices developed for exascale computing are enabling enterprise AI. Discover how organizations can accelerate AI adoption, scale workloads, and reduce deployment risk.;HPE and AMD have built some of the world's most advanced supercomputers, including Frontier and El Capitan. Join HPE leaders who helped deploy these systems and learn how the technologies, expertise, and operational practices developed for exascale computing are enabling enterprise AI. Discover how organizations can accelerate AI adoption, scale workloads, and reduce deployment risk.
July 23, 2026
-
Unlocking Secure Enterprise Intelligence at Scale with Cisco
Unlocking Secure Enterprise Intelligence at Scale with Cisco
As organizations transition from AI experimentation to production-scale infrastructure, demand for high-performance compute must be matched by security and reliability. This session explores Cisco's vision for secure, high-performance AI environments and a framework for accelerating AI deployment while mitigating risks associated with large-scale data processing. Learn how the Cisco UCS C845A M8 and AMD are enabling the next generation of enterprise AI.;As organizations transition from AI experimentation to production-scale infrastructure, demand for high-performance compute must be matched by security and reliability. This session explores Cisco's vision for secure, high-performance AI environments and a framework for accelerating AI deployment while mitigating risks associated with large-scale data processing. Learn how the Cisco UCS C845A M8 and AMD are enabling the next generation of enterprise AI.
July 23, 2026
-
Panel Discussion: From Models to Production—A Blueprint for AI at Scale
Panel Discussion: From Models to Production—A Blueprint for AI at Scale
Moving AI from training to production takes more than GPUs. Hear how Microsoft and Chai AI built scalable AI infrastructure on Vultr using AMD Instinct GPUs and ROCm. Learn best practices for data locality, secure networking, Kubernetes orchestration, benchmarking, cost optimization, and scale-out operations. Leave with a practical blueprint for deploying fast, portable, production-ready AI workloads.;Moving AI from training to production takes more than GPUs. Hear how Microsoft and Chai AI built scalable AI infrastructure on Vultr using AMD Instinct GPUs and ROCm. Learn best practices for data locality, secure networking, Kubernetes orchestration, benchmarking, cost optimization, and scale-out operations. Leave with a practical blueprint for deploying fast, portable, production-ready AI workloads.
July 23, 2026
-
Right Size Your Memory Footprint to Move IT Refresh Forward
Right Size Your Memory Footprint to Move IT Refresh Forward
Memory has rarely been in such short supply and is impeding customer data center refresh plans. In this interactive conversation, we’ll discuss tips and tools for right-sizing memory configurations to help move your data center efficiency initiatives forward and preserve ROI. Bring your questions and our experts will provide answers!;Memory has rarely been in such short supply and is impeding customer data center refresh plans. In this interactive conversation, we’ll discuss tips and tools for right-sizing memory configurations to help move your data center efficiency initiatives forward and preserve ROI. Bring your questions and our experts will provide answers!
July 23, 2026
-
From Tokens to Outcomes: Driving AI ROI with Lenovo Hybrid Infrastructure
From Tokens to Outcomes: Driving AI ROI with Lenovo Hybrid Infrastructure
AI success is increasingly measured by business outcomes, not model size. As agentic AI accelerates inference demand, organizations must improve token efficiency, infrastructure utilization, and energy consumption to maximize ROI. Learn how Lenovo Hybrid AI Factories, powered by AMD, help enterprises deploy AI from personal systems to rack-scale environments while reducing token costs, increasing control and utilization, and supporting more sustainable AI growth.;AI success is increasingly measured by business outcomes, not model size. As agentic AI accelerates inference demand, organizations must improve token efficiency, infrastructure utilization, and energy consumption to maximize ROI. Learn how Lenovo Hybrid AI Factories, powered by AMD, help enterprises deploy AI from personal systems to rack-scale environments while reducing token costs, increasing control and utilization, and supporting more sustainable AI growth.
July 23, 2026
-
Redefining Scalable AI Performance: OCI Supercomputing in the Cloud
Redefining Scalable AI Performance: OCI Supercomputing in the Cloud
Organizations building frontier AI models need infrastructure designed for performance at scale. This session shows how OCI combines AMD Instinct, AMD EPYC, and Pensando in Oracle Acceleron to enable ultra-low-latency networking for high-throughput distributed workloads, with practical guidance for designing infrastructure for large language, multimodal, and scientific AI models.;Organizations building frontier AI models need infrastructure designed for performance at scale. This session shows how OCI combines AMD Instinct, AMD EPYC, and Pensando in Oracle Acceleron to enable ultra-low-latency networking for high-throughput distributed workloads, with practical guidance for designing infrastructure for large language, multimodal, and scientific AI models.
July 23, 2026
-
Accelerating Inference at Scale: Crusoe's Experience with AMD
Accelerating Inference at Scale: Crusoe's Experience with AMD
As a customer and operator of AMD technology, Crusoe’s Managed Inference team has built a production inference stack designed for speed, efficiency, and scale. This session will show how AMD Instinct, including MI355X, helped shape its serverless inference offering and what teams can apply when building production AI services that balance performance, memory bandwidth, and cost.;As a customer and operator of AMD technology, Crusoe’s Managed Inference team has built a production inference stack designed for speed, efficiency, and scale. This session will show how AMD Instinct, including MI355X, helped shape its serverless inference offering and what teams can apply when building production AI services that balance performance, memory bandwidth, and cost.
July 23, 2026
-
Samsung + AMD: Architecting for AI Breakthroughs with Co-Designed Solutions
Samsung + AMD: Architecting for AI Breakthroughs with Co-Designed Solutions
Samsung and AMD are collaborating on advanced memory technologies for AI and data center workloads. Learn how advances in memory bandwidth and power efficiency enable next-generation AI infrastructure, including rack-scale architectures powered by AMD Instinct GPUs, AMD EPYC CPUs, and AMD Helios. Discover how co-designed memory and compute technologies can help optimize AI system performance.;Samsung and AMD are collaborating on advanced memory technologies for AI and data center workloads. Learn how advances in memory bandwidth and power efficiency enable next-generation AI infrastructure, including rack-scale architectures powered by AMD Instinct GPUs, AMD EPYC CPUs, and AMD Helios. Discover how co-designed memory and compute technologies can help optimize AI system performance.
July 23, 2026