Creating CPU Inference and Agentic Performance Transparency
Abstract
Your finance team doesn't care about tokens per second. They care about predictable costs, compliance risk, and vendor lock-in. With agentic AI, the metrics for tracking success are even more complex. But benchmarks don't answer the question that actually matters: Should you undertake this effort and is it viable for your business? In this interactive technical discussion, we’ll break down the tradeoffs, work through the math, and pressure-test the strategy together.
July 22, 2026 3:30 PM - 4:00 PM PDT
Speakers
Presented By
Sr Director, Product Management | AMD
Director, Product for Private Cloud | Rackspace
Session Type
Meet the Experts
Related Sessions
-
Unlock the Power of AI with AI Producer Studio
Unlock the Power of AI with AI Producer Studio
AI Producer Studio is optimized for AMD Ryzen AI processors to automate multi-camera meetings, livestreams, and recordings. See how AI detects presenters, follows conversations, and automatically manages camera switching, framing, and layouts in real time. Learn how organizations can simplify professional video production and enhance hybrid collaboration with AMD.;AI Producer Studio is optimized for AMD Ryzen AI processors to automate multi-camera meetings, livestreams, and recordings. See how AI detects presenters, follows conversations, and automatically manages camera switching, framing, and layouts in real time. Learn how organizations can simplify professional video production and enhance hybrid collaboration with AMD.
July 23, 2026
-
Supercomputing for All: Bringing Exascale Innovation to Enterprise AI
Supercomputing for All: Bringing Exascale Innovation to Enterprise AI
HPE and AMD have built some of the world's most advanced supercomputers, including Frontier and El Capitan. Join HPE leaders who helped deploy these systems and learn how the technologies, expertise, and operational practices developed for exascale computing are enabling enterprise AI. Discover how organizations can accelerate AI adoption, scale workloads, and reduce deployment risk.;HPE and AMD have built some of the world's most advanced supercomputers, including Frontier and El Capitan. Join HPE leaders who helped deploy these systems and learn how the technologies, expertise, and operational practices developed for exascale computing are enabling enterprise AI. Discover how organizations can accelerate AI adoption, scale workloads, and reduce deployment risk.
July 23, 2026
-
Domain-Specific AI at Scale: Open Models, Post-Training, and AI Infrastructure
Domain-Specific AI at Scale: Open Models, Post-Training, and AI Infrastructure
Learn how domain-specific AI moves beyond generic models using post-training, domain evals, and scalable open infrastructure. Using Open Telco Models as a case study, this session covers curated data, reward loops, unified training and serving, and AMD Instinct/ROCm-based stacks for building specialized AI systems at enterprise scale.;Learn how domain-specific AI moves beyond generic models using post-training, domain evals, and scalable open infrastructure. Using Open Telco Models as a case study, this session covers curated data, reward loops, unified training and serving, and AMD Instinct/ROCm-based stacks for building specialized AI systems at enterprise scale.
July 23, 2026
-
Unlocking Secure Enterprise Intelligence at Scale with Cisco
Unlocking Secure Enterprise Intelligence at Scale with Cisco
As organizations transition from AI experimentation to production-scale infrastructure, demand for high-performance compute must be matched by security and reliability. This session explores Cisco's vision for secure, high-performance AI environments and a framework for accelerating AI deployment while mitigating risks associated with large-scale data processing. Learn how the Cisco UCS C845A M8 and AMD are enabling the next generation of enterprise AI.;As organizations transition from AI experimentation to production-scale infrastructure, demand for high-performance compute must be matched by security and reliability. This session explores Cisco's vision for secure, high-performance AI environments and a framework for accelerating AI deployment while mitigating risks associated with large-scale data processing. Learn how the Cisco UCS C845A M8 and AMD are enabling the next generation of enterprise AI.
July 23, 2026
-
Zyphra: Large-Model Training Lessons on AMD
Zyphra: Large-Model Training Lessons on AMD
Learn what it took to train ZAYA1-74B, a 74B-parameter mixture-of-experts model, end-to-end on AMD Instinct MI300X. This session shares key engineering lessons from designing an efficient training stack, optimizing long-context performance, and building a reinforcement learning pipeline for math, code, and agentic AI workloads. Discover practical insights for training and deploying large AI models on AMD infrastructure.;Learn what it took to train ZAYA1-74B, a 74B-parameter mixture-of-experts model, end-to-end on AMD Instinct MI300X. This session shares key engineering lessons from designing an efficient training stack, optimizing long-context performance, and building a reinforcement learning pipeline for math, code, and agentic AI workloads. Discover practical insights for training and deploying large AI models on AMD infrastructure.
July 23, 2026
-
Right Size Your Memory Footprint to Move IT Refresh Forward
Right Size Your Memory Footprint to Move IT Refresh Forward
Memory has rarely been in such short supply and is impeding customer data center refresh plans. In this interactive conversation, we’ll discuss tips and tools for right-sizing memory configurations to help move your data center efficiency initiatives forward and preserve ROI. Bring your questions and our experts will provide answers!;Memory has rarely been in such short supply and is impeding customer data center refresh plans. In this interactive conversation, we’ll discuss tips and tools for right-sizing memory configurations to help move your data center efficiency initiatives forward and preserve ROI. Bring your questions and our experts will provide answers!
July 23, 2026
-
Build an MRI Analysis Agent with AMD Blueprints
Build an MRI Analysis Agent with AMD Blueprints
Build and deploy an AI-powered MRI analysis agent in minutes using the AMD mri-doc Solution Blueprint. Run a Gradio-based pipeline that accepts DICOM, NIfTI, and standard image formats, applies tissue segmentation and anomaly detection, and generates LLM-drafted clinical reports on AMD Instinct GPUs. Then customize: swap the LLM AIM, reuse an existing model endpoint, or extend the pipeline for your specific clinical workflow.;Build and deploy an AI-powered MRI analysis agent in minutes using the AMD mri-doc Solution Blueprint. Run a Gradio-based pipeline that accepts DICOM, NIfTI, and standard image formats, applies tissue segmentation and anomaly detection, and generates LLM-drafted clinical reports on AMD Instinct GPUs. Then customize: swap the LLM AIM, reuse an existing model endpoint, or extend the pipeline for your specific clinical workflow.
July 23, 2026
-
From Tokens to Outcomes: Driving AI ROI with Lenovo Hybrid Infrastructure
From Tokens to Outcomes: Driving AI ROI with Lenovo Hybrid Infrastructure
AI success is increasingly measured by business outcomes, not model size. As agentic AI accelerates inference demand, organizations must improve token efficiency, infrastructure utilization, and energy consumption to maximize ROI. Learn how Lenovo Hybrid AI Factories, powered by AMD, help enterprises deploy AI from personal systems to rack-scale environments while reducing token costs, increasing control and utilization, and supporting more sustainable AI growth.;AI success is increasingly measured by business outcomes, not model size. As agentic AI accelerates inference demand, organizations must improve token efficiency, infrastructure utilization, and energy consumption to maximize ROI. Learn how Lenovo Hybrid AI Factories, powered by AMD, help enterprises deploy AI from personal systems to rack-scale environments while reducing token costs, increasing control and utilization, and supporting more sustainable AI growth.
July 23, 2026