惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Vercel News
Vercel News
博客园 - 【当耐特】
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
小众软件
小众软件
Hugging Face - Blog
Hugging Face - Blog
aimingoo的专栏
aimingoo的专栏
WordPress大学
WordPress大学
G
Google Developers Blog
博客园 - 叶小钗
大猫的无限游戏
大猫的无限游戏
P
Proofpoint News Feed
J
Java Code Geeks
U
Unit 42
云风的 BLOG
云风的 BLOG
阮一峰的网络日志
阮一峰的网络日志
N
Netflix TechBlog - Medium
宝玉的分享
宝玉的分享
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
D
Docker
V
Visual Studio Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
H
Help Net Security
V
V2EX
T
Tailwind CSS Blog

SiliconANGLE

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws - SiliconANGLE AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle - SiliconANGLE Team Cymru launches Total Insights Feed to replace legacy threat intelligence lists - SiliconANGLE AI Mode in Chrome adds split-screen view to enhance the web search experience - SiliconANGLE Resolve AI raises $40M at $1.5B valuation to optimize production environments - SiliconANGLE How Zscaler and OpenAI turn zero-trust security into an AI accelerator - SiliconANGLE OpenAI ratchets up Codex's agentic capabilities to rival Claude Code - SiliconANGLE Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements - SiliconANGLE Slash raises $100M at a $1.4B valuation to expand AI-powered banking platform for online businesses - SiliconANGLE Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work - SiliconANGLE Data center, consumer device chips boost TSMC’s revenue - SiliconANGLE Mission-critical security cannot be bolted on, says Oracle - SiliconANGLE Agentic infrastructure reshapes enterprise AI - SiliconANGLE Data quality, and data freedom, foundational for AI success - SiliconANGLE Data trust is a bedrock in successful, scalable AI outcomes - SiliconANGLE Google introduces new agentic AI-ready tools and resources for Android developers  - SiliconANGLE Agentic AI orchestration separates winners from laggards - SiliconANGLE Data-driven tools turning the tide against human trafficking - SiliconANGLE Achieving trusted AI development goes beyond 'vibes' - SiliconANGLE Impinj boosts edge computing power in updated R700 RAIN RFID reader - SiliconANGLE Certinia powers professional services with AI - SiliconANGLE Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M - SiliconANGLE Developer tooling startup Expo nabs $45M investment - SiliconANGLE Solidroad lands $25M to bring AI to customer support interactions - SiliconANGLE DuploCloud lands compliance and AI governance certifications as enterprise buyers tighten scrutiny - SiliconANGLE Lua lands $5.8M to help businesses build and manage AI agent workforces - SiliconANGLE Best of frenemies: Oracle's and AWS' clouds unite with dedicated, private connectivity - SiliconANGLE NIST shifts National Vulnerability Database to risk-based triage as CVE submissions hit record levels - SiliconANGLE Cisco goes to the races with new Churchill Downs multiyear partnership - SiliconANGLE Susecon 2026 will tackle the future of open-source platforms - SiliconANGLE
Hybrid AI architecture for agentic workloads at scale - S...
Ryan Stevens · 2026-05-21 · via SiliconANGLE

GPUs changed the equation of enterprise compute. AMD and Dell say agentic AI is flipping the math once again

As enterprises graduate from AI experimentation to production-scale agentic deployments, the infrastructure assumptions of the chatbot era are rapidly giving way to a more distributed and cost-conscious hybrid AI architecture.

The shift is playing out in real time, with the AI factory emerging as the central organizing principle for rearchitecting enterprise compute. Token economics, data gravity and constrained data center power density are forcing a hard look at which workloads belong on-premises, which belong at the edge and which warrant frontier-model API calls, according to Suresh Andani (pictured, left), corporate vice president for compute and enterprise AI at Advanced Micro Devices Inc. That’s where AMD’s MI350P — a GPU card designed to slot into existing servers — comes in.

“About 70% of enterprise data centers are 30-kilowatt rack power density or lower — and about 50% of them are lower than about 15 kilowatts,” Andani said. “If you are a traditional data center and you have ambitions to be an AI-sophisticated enterprise, what do you do? That’s where [the MI350P] comes in, where you can take your existing servers … and plug in these 350P cards and still get 150, 170 billion parameter inference models running very efficiently.”

Andani and Melissa Crichton (right), vice president of server and AI solutions at Dell Technologies Inc., spoke with theCUBE’s John Furrier and Dave Vellante at Dell Technologies World 2026, during an exclusive broadcast on theCUBE, SiliconANGLE Media’s livestreaming studio. They discussed hybrid AI architecture, the AMD MI350P launch and evolving CPU-to-GPU ratios in agentic deployments. (* Disclosure below.)

Hybrid AI architecture and the agentic compute shift

The hybrid AI architecture discussion is inseparable from the economics of scale. Dell and AMD recently announced support for the MI350P in Dell PowerEdge servers, with the emphasis on enabling enterprises to run meaningful inference workloads within their existing power envelopes — without costly infrastructure rebuilds. Roughly 80% of data is created at the edge, meaning enterprises cannot run AI in a single location, Crichton noted.

“We believe we have to build out a hybrid platform for our customers — whether it runs at the edge, the core or in a hyperscaler,” Crichton said. “That’s how we think about the AI factory; creating that enterprise scalable model for our customers to start small, scale to the largest models … and creating that orchestration and abstraction layer to move workloads where it’s best suited.”

The compute ratio question is equally fundamental. Agentic AI has driven the GPU-to-CPU ratio from 8:1 toward 1:1 — and AMD believes it could soon invert entirely — because planning, orchestration and tool-calling in multi-step agent workflows are serial tasks that favor CPU architecture over massively parallel GPU compute, Andani explained.

“In the agentic flow, where you’re running multi-system agents, the first step you do when an agent request comes in [is] you need to start planning … that’s a combination of CPU and GPU,” he said. “Then you’ve got to go execute that plan, which involves a lot of orchestration, which is a serial job — it’s not a parallel job that GPUs do well. All of that tool execution is optimized on a serial architecture like a CPU versus a massively parallel architecture like a GPU. If you don’t do that, your very expensive GPUs are sitting idle, and that is a waste of money.”

Here’s the complete video interview, part of SiliconANGLE’s and theCUBE’s coverage of Dell Technologies World 2026:

(* Disclosure: AMD sponsored this segment of theCUBE. Neither AMD nor other sponsors have editorial control over content on theCUBE or SiliconANGLE.)

Photo: SiliconANGLE

A message from John Furrier, co-founder of SiliconANGLE:

Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities.

  • 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more
  • 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network.

About SiliconANGLE Media

SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement. As the parent company of SiliconANGLE, theCUBE Network, theCUBE Research, CUBE365, theCUBE AI and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.

Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.