惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 三生石上(FineUI控件)
G
Google Developers Blog
Vercel News
Vercel News
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
N
Netflix TechBlog - Medium
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Engineering at Meta
Engineering at Meta
B
Blog
博客园_首页
量子位
博客园 - 叶小钗
L
LangChain Blog
T
The Blog of Author Tim Ferriss
云风的 BLOG
云风的 BLOG
Blog — PlanetScale
Blog — PlanetScale
F
Fortinet All Blogs
S
SegmentFault 最新的问题
宝玉的分享
宝玉的分享
D
DataBreaches.Net
雷峰网
雷峰网
The Cloudflare Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
Last Week in AI
Last Week in AI
P
Proofpoint News Feed
TaoSecurity Blog
TaoSecurity Blog
罗磊的独立博客
MongoDB | Blog
MongoDB | Blog
The GitHub Blog
The GitHub Blog
I
Intezer
H
Help Net Security
The Hacker News
The Hacker News
The Register - Security
The Register - Security
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
AWS News Blog
AWS News Blog
V
V2EX
Microsoft Security Blog
Microsoft Security Blog
T
Tenable Blog
Spread Privacy
Spread Privacy
A
Arctic Wolf
P
Proofpoint News Feed
T
Threat Research - Cisco Blogs
Schneier on Security
Schneier on Security
C
CERT Recently Published Vulnerability Notes
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
The Last Watchdog
The Last Watchdog
Latest news
Latest news
T
Troy Hunt's Blog
L
LINUX DO - 热门话题

The Next Platform: In-depth coverage of high end computing

Uncle Sam Awards $2 Billion-Plus To Quantum Companies, But Wants A Cut Oak Ridge Starts Weaving Together A Quantum, Classical HPC, And AI System Stack Dell Bulks Up Hardware As AI Infrastructure Shifts To On-Premises Cisco Wins Over AI Customers With Merchant Silicon And Optics With Its IPO Done, Cerebras Can Get Back To Pushing The AI Envelope HPE Throws VM Users A Lifeline, Unifying Containers And VM Management In Cloud Stack OpenAI, Microsoft And Friends Build A Better, More Scalable Ethernet Compute And Memory Price Hikes Drive IT Spending Way Higher Sometimes, Air Is The Only Way For AI Systems To Keep Their Cool Arista Rides AI Scale Out Networks, Moves Into Scale Across, And Awaits Scale Up Cleveland Clinic Simulates Large Proteins With Quantum-Centric Supercomputing Broadcom Helps CPU And XPU Makers Go Vertical With Compute Microsoft Committed To Doubling AI Infrastructure In Two Years Google Is A Full Stack AI Player, And Is Playing Well AWS Will Be An OEM, Just Like Google And Maybe Microsoft New Google Networks Tuned Up For GenAI Inference And Training Microsoft And OpenAI Remain Friends, Are Looking To Hook Up With Others AI-Driven CPU Shortage Saves Intel’s Financial Cookies The GenAI Battle Shifts From Frontier Models To Agentic Platforms With TPU 8, Google Makes GenAI Systems Much Better, Not Just Bigger Cisco Scales Out Quantum Systems With A Quantum Network Switch The Second Time Will Be The IPO Charm For Cerebras Imagine An Army Of AI Minions Handling Incident Response AI Will Soon Drive A Third Of TSMC’s Business Bechtolsheim & Friends Breathe Life Into Pluggable Optics One Last Time How HPC And AI Digital Twins Accelerate Quantum Error Correction The Embrace Of AI In Design Transforms Cadence And Its Customers Nvidia Brings The Power Of Open Source AI Models To Quantum Computing Building The Imperfect Beast For Enterprises, GPUs Need Virtualization As Much As CPUs Ever Did CoreWeave Takes As Much Financial Engineering As It Does Datacenter Design Contemplating Meta’s Homegrown MTIA Compute Engine Roadmap Most Neoclouds, Sovereigns, And Enterprises Will Buy, Not Build, Their AI Stacks Broadcom And Google Benefit Mightily From Anthropic’s Meteoric Growth Rebellions AI Rings Up The Money To Rack Up AI Inference Systems Nvidia Software Pushes MLPerf Inference Benchmarks To New Highs Broadcom Makes Its Pitch To Run Kubernetes On VMware VCF The $2 Billion Nvidia Deal With Marvell Is About A Lot More Than NVLink Fusion Classiq Says Quantum Is On Its Way, But Patience Is Needed Demonstrating The Scientific Usefulness Of Quantum Systems We Need Servers – Lots Of Servers. . . . Arm Comes Full Circle With Homegrown, AI-Tuned Server CPU Riding The Memory Boom And Trying To Avoid The Bust Data Analytics Helps Make The Mighty Lionesses Roar Driving Down The AI System Roadmap With Nvidia The Open Agentic AI World According To Nvidia Nvidia Finally Admits Why It Shelled Out $20 Billion For Groq Nvidia Says OpenClaw Is To Agentic AI What GPT Was To Chattybots IBM Unrolls Blueprint For Quantum-Classical HPC Computing Women Get Data-Driven Health Boost As The FA Tackles Sports Science Four Months Into Its Comeback, Zapata Stakes Its Claim In Quantum Software Eridu Cuts To The AI Networking Chase With High Radix Switch System HPE Works Harder And Smarter To Chase Datacenter Profits We Need A Proper AI Inference Benchmark Test How AI Is Boosting Gender Equality In High Performance Racing Custom Compute Engine Biz Growing More Than Marvell Ever Hoped Broadcom May Become The Biggest Counterbalance To Nvidia Ayar Labs Gets $500 Million To Ramp Photonics Into 2028 AI Systems With Cisco Outshift, Agentic AI Is Teed Up For the Internet Of Cognition Nvidia Sees The Light On Silicon Photonics And Maybe Optical Switching AI Servers Finally Dominate Dell’s Systems Business VAST Data: What Controls The Data Is More Important Than What Stores It So Far, Nobody Turns Tokens Into Money Like Nvidia SambaNova Pits Its Engineering Against Nvidia For Agentic AI Some More Game Theory, This Time On The AMD-Meta Platforms Deal AMD Says “Helios” Racks And MI400 Series GPUs On Track For 2H 2026 CPU-Only Compute Still Matters To A Lot Of HPC Centers Taalas Etches AI Models Onto Transistors To Rocket Boost Inference Some Game Theory On That Nvidia-Meta Platforms Partnership AI Eats The World, And Most Of Its Flash Storage The Current AI Networking Wave Will Be A Tsunami Of Money By 2027 The Memory Crunch Pinches Cisco’s Profits Only A Few AI Platforms Can Survive The Greatest AI Show On Earth Cisco Doubles Up The Switch Bandwidth To Take On AI Scale Out And Eventually Scale Up Datacenter Spending Forecast Revised Upwards – Yet Again The Twin Engine Strategy That Propels AWS Is Working Well With GenAI Turbochargers, Google Is Shifting Its Cloud Into A Higher Gear AMD Finally Makes More Money On GPUs Than CPUs In A Quarter Dassault And Nvidia Bring Industrial World Models To Physical AI TACC Explores Mixed Precision And FP64 Emulation For HPC With Horizon Robotics Will Break AI infrastructure: Here's What Comes Next Oracle’s Financing Primes The OpenAI Pump Gartner Takes Another Stab At Forecasting AI Spending Microsoft Is More Dependent On OpenAI Than The Converse Big Blue Poised To Peddle Lots Of On Premises GenAI Microsoft Takes On Other Clouds With “Braga” Maia 200 AI Compute Engines Nvidia’s $2 Billion Investment In CoreWeave Is A Drop In A $250 Billion Bucket Intel Is Still Struggling In The Datacenter, But It Could Get Better Is Nvidia Assembling The Parts For Its Next Inference Platform? TSMC Has No Choice But To Trust The Sunny AI Forecasts Of Its Customers Cerebras Inks Transformative $10 Billion Inference Deal With OpenAI By Decade’s End, AI Will Drive More Than Half Of All Chip Sales Startup Quantum Elements Brings AI, Digital Twins To Quantum Computing D-Wave Makes Gate-Model Power Move With Quantum Circuits Buy Building The Future Of Software In The AI-Native Era Arista Modular Switches Aim At Scale Across Networks, Hit Scale Out, Too NextSilicon Takes Aim At CPUs And GPUs With “Maverick-2” Dataflow Engine How HPC Is Igniting Discoveries In Dinosaur Locomotion – And Beyond Oracle First In Line For AMD “Altair” MI450 GPUs, “Helios” Racks
If You Can Make A Compute Engine, You Can Sell A Compute Engine
Timothy Prickett Morgan · 2026-05-07 · via The Next Platform: In-depth coverage of high end computing

The title says it all, really. But make no mistake: There is still plenty of competition in the markets for CPUs, GPUs, and XPUs and the DRAM, HBM, and flash memory that they all depend upon despite the demand exceeding supply. The competition is now focusing on access to devices as much as it ever did on feeds, speeds, and price.

And, there is a new market emerging for CPUs, driven by the specific computing needs of agentic AI, that will, within the next few years, come to dominate the installed base of servers even if it does not represent the lion’s shares of revenues as do AI training systems by virtue of those very expensive GPUs they are laden with.

That is the new gospel according to AMD, whose top brass echoed many of the comments made last week by Intel as there is not just a resurgence on spending on CPUs – which there most certainly is – but a tectonic shift underway thanks to agentic AI.

This is the beginning of a trend, so its outlines are a bit vague. It was only back in early November 2025 at its Financial Analyst Day that Lisa Su, chief executive officer at AMD, put a stake in the ground and predicted that the total addressable market, or TAM, for datacenter CPUs was going to grow at a compound annual growth rate of 18 percent, reaching $60 billion by 2030. That is a lot of server CPUs in the datacenter. This compares to a roughly $13 billion to $14 billion server CPU silicon TAM back in 2015, when Intel accounted for close to 99 percent of shipments and when most of the RISC/Unix and Itanium alternatives were either dead or dying and Arm server CPUs had been at it in earnest for four years but had damned near zero traction.

If $60 billion sounds like a lot, how do you like $120 billion? Because that is the new server CPU TAM that Su & Co laid down in the conference call with Wall Street analysts yesterday going over the company’s first quarter of 2026 financials. Which we will get to in a second. So now, the CAGR between last year and 2030 for server CPUs is 35 percent, and most of the incremental money that AMD, Intel, Nvidia, Arm, and the homegrown CPU crowd will chase will be allocated to agentic AI systems, which are distinct from traditional system of record and system of engagement systems that have been the foundation of the server market for decades. Su called this general purpose compute, and she referred to the relatively thin slice of GenAI training and inference where the CPU is in the head node as a second part of the server CPU TAM. The third slice of the pie – one that appears to be growing exponentially – “are CPUs just for all of the agentic AI work,” as she put it.

Here is how Su explained it further:

“So you should think about we need all of the accelerators to run these foundational models, and then as these agents do work, they spawn more CPU tasks. So I would say it is largely incremental. The key is to make sure – what we are seeing is in these deployments – the ratio of CPUs to GPUs are the right ratio. So if you are installing a gigawatt of compute, the percentage of CPU as part of that gigawatt will increase.”

“Some of the conversation in the industry has been about CPU to GPU ratios. And it's very hard to call exactly, but we certainly see the movement towards where in the past, the CPU to GPU ratio was primarily just as a host node in like a 1:4 or 1:8 configuration node, now changing and getting closer to a 1:1 configuration – or you can even imagine if you get lots and lots of agents that you could have more CPUs than GPUs. But all in all, I think it is largely additive to the TAM. And the key is that everyone is now planning and thinking about CPUs at the same time that they are thinking about their accelerator deployments, which is a good thing.”

Of course, when you are counting sockets, you have to count what you are putting into the sockets and how big those sockets are. Clearly, the 72-core “Grace” CPU was a bit much for the “Hopper” CPU that Nvidia paired it with because with the next generation “Blackwell” CPUs, there were two GPUs attached to every CPU. And in fact, there were two Blackwell GPU chiplets per socket, so the ratio was really 1:4. And as far as we know, Nvidia will have one “Vera” CPU for every pair of “Rubin” GPUs – which is actually four GPU chiplets in total for a 1:4 ratio – later this year and will double up the GPU chiplets in 2027 in the Rubin Ultra socket to effectively get a 1:8 ratio. If there are eight compute tiles on the Vera CPU, then maybe we can say the ratio is really one to one.

It is tricky to count this, as Su has said. And it will get trickier when banks of CPUs start doing some inference for smaller models in a mixture of experts GenAI platform. The rule will be simple: Always use the smallest model that will get the right answer with acceptable accuracy. And that will mean that something we have contended for many years will come true: A lot of AI inference will be done by CPUs.

And that means there will be special cost-optimized CPUs from Intel, AMD, and Arm that are good at this agentic workflow and are different even from the host CPU job inside of GPU and XPU nodes. That is precisely what the “Verano” variation on the “Venice” Epyc 9006 theme seems to be. Verano is apparently coming out in 2027, very likely well ahead of the “Florence” Epyc 9007 processors expected perhaps in late 2027 or early 2028 and very likely paired with the next generation MI500 series GPUs and the kicker to the Helios racks that will sport Venice CPUs and MI400 GPUs. The Venice CPUs are slated for later this year, of course.

“We are clearly feeling like we are seeing significant share gain as we are going into our Turin portfolio that has ramped very nicely. Venice is extremely well positioned, and we are working with customers right now on – beyond Venice and what we are doing in those architectures. So we feel really good about the market as well as our opportunity to grow to greater than 50 percent share of that market.”

So there it is, gauntlet laid down at the feet of Lip-Bu Tan, chief executive officer of Intel.

In any event, this agentic AI server CPU boom is just starting to materialize and it is not what drove the impressive Epyc CPU revenues in Q1 2026. What drove that AMD boom is a shortage of supply and an excess of demand. As we have contended for some time now, companies need to upgrade their general purpose machines and consolidate down generations of older gear to save power, space, and money so they can afford to invest in AI. Which companies are doing, too.

There is more demand than meets supply, but both Intel and AMD have been increasing supply and therefore revenues are rising. It doesn’t hurt pricing that demand is still bigger than supply, and neither company has any incentive to make as many chips as they can sell. With substrates and other components also being in short supply, it is better to hang back, charge a premium, and bring more dough to the bottom line. Which is exactly what CPU and GPU makers are doing. They are selling lots of current generation products, and because of the high demand, they are also able to sell back generation products that are cheaper to make and more profitable at this point, too.

In the March quarter, AMD brought in a total of $10.25 billion, up 37.8 percent year on year and down a tad sequentially. Operating income, thanks to the pricing power that AMD is enjoying, rose by 83.1 percent to $1.48 billion, and net income rose by 95.1 percent to $1.38 billion. That’s only 13.5 percent of revenue, mind you, at the bottom line, so AMD is not gouging anyone so much as being able to command the some of the premium it has long deserved.

The company is also making massive investments in future server and PC architectures and that is also weighing on profits. But better that than get into a technical debt hole as both AMD and Intel have done in the past.

AMD exited the quarter with $12.35 billion in cash and equivalents, up 68.9 percent from a year ago. That is not quite enough money to buy Cerebras Systems before it goes public next week on May 14. . . .

AMD’s Data Center group, which sells server CPUs, DPUs, high-end FPGAs, and soon rack-scale components that its partners will use to create Helios double-wide racks for AI systems (and potentially for HPC applications), saw revenues grow by 57.2 percent to $5.78 billion. Operating income rose by 71.6 percent to $1.6 billion, and that income represented 27.7 percent of sales. That’s about average for AMD these days.

The Client group, which sells chips for PCs, had $2.9 billion in sales, up 25.8 percent, and the Gaming group, which sells high-end graphics cards for gamers and sometimes for those trying to do HPC and AI on the cheap, had $720 million in sales. AMD has combined operating income for the Client and Gaming groups because it is trying to mask something, and that something is that the PC chip business is not that profitable, with what we estimate is only $403 million, representing 14 percent of sales. The Gaming group does much better, at $173 million in our estimates, representing 24 percent of revenues for its operating income.

That leaves the Embedded group, which sells FPGAs and which had $873 million in sales, up 6.1 percent, with $338 million in operating profit, which is 38.7 percent of sales. Yup. FPGAs continue to be the most profitable part of AMD – and partially thanks to the bad management of FPGA rival Altera by Intel. (Altera has been spun out and we need to circle back and take a look at how it is doing with its newfound freedom.)

The Data Center group continues to grow in importance for AMD, and represented 56.3 percent of revenues in Q1 2026. Only three years ago, Data Center group only accounted for a quarter of AMD’s overall revenues, and a decade ago the datacenter represented a paltry 2.5 percent of sales.

Lisa Su said she would fix AMD and give it credibility in the datacenter, and she has absolutely done that.

Now comes the spreadsheet witchcraft that we so enjoyed. We were so pleased to report last quarter that the Instinct GPU business was larger than the Epyc CPU business for the first time in AMD’s history. Well, thanks to the CPU boom and the dearth of HBM memory out there in the world, and a product transition from the MI300 series to the MI400 series, the Instinct line was not able to keep pace with the Epyc line in the first quarter.

Our model shows the Instinct GPU business growing by 64 percent to a smidgen above $1.9 billion in the March quarter, which is great, but it also is a 28.2 percent sequential decline from the record $2.65 billion AMD racked up in datacenter GPU sales in Q4 2025.

The Epyc CPUs, by contrast, grew by 53 percent to $3.65 billion in our model, and based on hints that Su & Co gave out on the call, we think sales to hyperscalers and clouds represented 76.2 percent of that CPU pie, which is $2.78 billion, while enterprises, telcos, service providers, academia, and governments accounted for $867 million in Epyc sales, up 51.4 percent.

The rest of the Data Center group dough, which is barely material, went to datacenter network interface cards, DPUs, and datacenter-class Xilinx FPGAs.

A final note: When AI settles down – that was meant to be funny – all of this agentic management CPU function could end up being hard-coded in an agentic router. This is, in fact, what happened to network routers. They were originally just software running on proprietary minicomputers.