惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
WordPress大学
WordPress大学
aimingoo的专栏
aimingoo的专栏
博客园 - 聂微东
M
MIT News - Artificial intelligence
Microsoft Security Blog
Microsoft Security Blog
云风的 BLOG
云风的 BLOG
Google DeepMind News
Google DeepMind News
Apple Machine Learning Research
Apple Machine Learning Research
H
Hackread – Cybersecurity News, Data Breaches, AI and More
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
O
OpenAI News
N
News and Events Feed by Topic
TaoSecurity Blog
TaoSecurity Blog
有赞技术团队
有赞技术团队
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
MyScale Blog
MyScale Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
Hacker News: Ask HN
Hacker News: Ask HN
Help Net Security
Help Net Security
爱范儿
爱范儿
W
WeLiveSecurity
V2EX - 技术
V2EX - 技术
Forbes - Security
Forbes - Security
Hacker News - Newest:
Hacker News - Newest: "LLM"
SecWiki News
SecWiki News
A
About on SuperTechFans
Cloudbric
Cloudbric
N
Netflix TechBlog - Medium
博客园 - 司徒正美
S
Security @ Cisco Blogs
Martin Fowler
Martin Fowler
Schneier on Security
Schneier on Security
I
InfoQ
Engineering at Meta
Engineering at Meta
Google Online Security Blog
Google Online Security Blog
T
Troy Hunt's Blog
Latest news
Latest news
N
News and Events Feed by Topic
Security Archives - TechRepublic
Security Archives - TechRepublic
Spread Privacy
Spread Privacy
NISL@THU
NISL@THU
B
Blog
L
LINUX DO - 最新话题
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
V
V2EX
J
Java Code Geeks
N
News | PayPal Newsroom
T
Tor Project blog
The GitHub Blog
The GitHub Blog

The Next Platform: In-depth coverage of high end computing

Oak Ridge Starts Weaving Together A Quantum, Classical HPC, And AI System Stack Dell Bulks Up Hardware As AI Infrastructure Shifts To On-Premises Cisco Wins Over AI Customers With Merchant Silicon And Optics With Its IPO Done, Cerebras Can Get Back To Pushing The AI Envelope HPE Throws VM Users A Lifeline, Unifying Containers And VM Management In Cloud Stack OpenAI, Microsoft And Friends Build A Better, More Scalable Ethernet Compute And Memory Price Hikes Drive IT Spending Way Higher Sometimes, Air Is The Only Way For AI Systems To Keep Their Cool Arista Rides AI Scale Out Networks, Moves Into Scale Across, And Awaits Scale Up If You Can Make A Compute Engine, You Can Sell A Compute Engine Cleveland Clinic Simulates Large Proteins With Quantum-Centric Supercomputing Broadcom Helps CPU And XPU Makers Go Vertical With Compute Microsoft Committed To Doubling AI Infrastructure In Two Years Google Is A Full Stack AI Player, And Is Playing Well AWS Will Be An OEM, Just Like Google And Maybe Microsoft New Google Networks Tuned Up For GenAI Inference And Training Microsoft And OpenAI Remain Friends, Are Looking To Hook Up With Others AI-Driven CPU Shortage Saves Intel’s Financial Cookies The GenAI Battle Shifts From Frontier Models To Agentic Platforms With TPU 8, Google Makes GenAI Systems Much Better, Not Just Bigger Cisco Scales Out Quantum Systems With A Quantum Network Switch The Second Time Will Be The IPO Charm For Cerebras Imagine An Army Of AI Minions Handling Incident Response AI Will Soon Drive A Third Of TSMC’s Business Bechtolsheim & Friends Breathe Life Into Pluggable Optics One Last Time How HPC And AI Digital Twins Accelerate Quantum Error Correction The Embrace Of AI In Design Transforms Cadence And Its Customers Nvidia Brings The Power Of Open Source AI Models To Quantum Computing Building The Imperfect Beast For Enterprises, GPUs Need Virtualization As Much As CPUs Ever Did CoreWeave Takes As Much Financial Engineering As It Does Datacenter Design Contemplating Meta’s Homegrown MTIA Compute Engine Roadmap Most Neoclouds, Sovereigns, And Enterprises Will Buy, Not Build, Their AI Stacks Broadcom And Google Benefit Mightily From Anthropic’s Meteoric Growth Rebellions AI Rings Up The Money To Rack Up AI Inference Systems Nvidia Software Pushes MLPerf Inference Benchmarks To New Highs Broadcom Makes Its Pitch To Run Kubernetes On VMware VCF The $2 Billion Nvidia Deal With Marvell Is About A Lot More Than NVLink Fusion Classiq Says Quantum Is On Its Way, But Patience Is Needed Demonstrating The Scientific Usefulness Of Quantum Systems We Need Servers – Lots Of Servers. . . . Arm Comes Full Circle With Homegrown, AI-Tuned Server CPU Riding The Memory Boom And Trying To Avoid The Bust Data Analytics Helps Make The Mighty Lionesses Roar Driving Down The AI System Roadmap With Nvidia The Open Agentic AI World According To Nvidia Nvidia Finally Admits Why It Shelled Out $20 Billion For Groq Nvidia Says OpenClaw Is To Agentic AI What GPT Was To Chattybots IBM Unrolls Blueprint For Quantum-Classical HPC Computing Women Get Data-Driven Health Boost As The FA Tackles Sports Science Four Months Into Its Comeback, Zapata Stakes Its Claim In Quantum Software Eridu Cuts To The AI Networking Chase With High Radix Switch System HPE Works Harder And Smarter To Chase Datacenter Profits We Need A Proper AI Inference Benchmark Test How AI Is Boosting Gender Equality In High Performance Racing Custom Compute Engine Biz Growing More Than Marvell Ever Hoped Broadcom May Become The Biggest Counterbalance To Nvidia Ayar Labs Gets $500 Million To Ramp Photonics Into 2028 AI Systems With Cisco Outshift, Agentic AI Is Teed Up For the Internet Of Cognition Nvidia Sees The Light On Silicon Photonics And Maybe Optical Switching AI Servers Finally Dominate Dell’s Systems Business VAST Data: What Controls The Data Is More Important Than What Stores It So Far, Nobody Turns Tokens Into Money Like Nvidia SambaNova Pits Its Engineering Against Nvidia For Agentic AI Some More Game Theory, This Time On The AMD-Meta Platforms Deal AMD Says “Helios” Racks And MI400 Series GPUs On Track For 2H 2026 CPU-Only Compute Still Matters To A Lot Of HPC Centers Taalas Etches AI Models Onto Transistors To Rocket Boost Inference Some Game Theory On That Nvidia-Meta Platforms Partnership AI Eats The World, And Most Of Its Flash Storage The Current AI Networking Wave Will Be A Tsunami Of Money By 2027 The Memory Crunch Pinches Cisco’s Profits Only A Few AI Platforms Can Survive The Greatest AI Show On Earth Cisco Doubles Up The Switch Bandwidth To Take On AI Scale Out And Eventually Scale Up Datacenter Spending Forecast Revised Upwards – Yet Again The Twin Engine Strategy That Propels AWS Is Working Well With GenAI Turbochargers, Google Is Shifting Its Cloud Into A Higher Gear AMD Finally Makes More Money On GPUs Than CPUs In A Quarter Dassault And Nvidia Bring Industrial World Models To Physical AI TACC Explores Mixed Precision And FP64 Emulation For HPC With Horizon Robotics Will Break AI infrastructure: Here's What Comes Next Oracle’s Financing Primes The OpenAI Pump Gartner Takes Another Stab At Forecasting AI Spending Microsoft Is More Dependent On OpenAI Than The Converse Big Blue Poised To Peddle Lots Of On Premises GenAI Microsoft Takes On Other Clouds With “Braga” Maia 200 AI Compute Engines Nvidia’s $2 Billion Investment In CoreWeave Is A Drop In A $250 Billion Bucket Intel Is Still Struggling In The Datacenter, But It Could Get Better Is Nvidia Assembling The Parts For Its Next Inference Platform? TSMC Has No Choice But To Trust The Sunny AI Forecasts Of Its Customers Cerebras Inks Transformative $10 Billion Inference Deal With OpenAI By Decade’s End, AI Will Drive More Than Half Of All Chip Sales Startup Quantum Elements Brings AI, Digital Twins To Quantum Computing D-Wave Makes Gate-Model Power Move With Quantum Circuits Buy Building The Future Of Software In The AI-Native Era Arista Modular Switches Aim At Scale Across Networks, Hit Scale Out, Too NextSilicon Takes Aim At CPUs And GPUs With “Maverick-2” Dataflow Engine How HPC Is Igniting Discoveries In Dinosaur Locomotion – And Beyond Oracle First In Line For AMD “Altair” MI450 GPUs, “Helios” Racks
The Server Boom Balances Price Increases Against Chip Shortages
Timothy Prickett Morgan · 2026-06-18 · via The Next Platform: In-depth coverage of high end computing

The rise in prices for CPUs, GPUs, DRAM memory, and flash storage have counterbalanced the imbalance between supply and demand such that the overall server market had only a slight sequential decline in the first quarter of 2026. This is not only amazing, but perhaps also intentional on the part of the OEMs and ODMs that bend metal around components to make servers.

If you were in the place of the OEMs and ODMs, you would be passing on the increasing costs of these components, and then given that supply exceeds demand by maybe 25 percent to 30 percent (based on hints from the big OEMs), you might even engage in a little opportunistic price increases where contracts with customers have not locked in the cost of machines. Considering how hard these server manufacturers work, and for the most demanding and penny-pinching customers in the world, they deserve a little margin and we hope that for once they are getting a little vig.

In the first quarter of 2026 according to the market researchers at IDC, the world consumed $122.62 billion in servers, up 30.4 percent year on year but down 2.1 percent sequentially. This is nothing compared to the historical sequential decline from the fourth quarter to the following first quarter that the server market has had in the decades before the GenAI boom. Everybody in the world knew the fourth quarter was when IT budgets were being set for the following year, and if IT shops didn’t spend all of their budget, it would not be held pat or raised in the next year. So Q4 was always stronger that Q3 and Q2 was always stronger than Q1 and usually bigger than Q3. There is a sawtooth pattern there, which is affected by the server roadmaps of CPU and system makers.

These days, at least for the past several quarters, there seems to be a new normal level of server spending despite all of these puts and takes, as they say on Wall Street.

Here is what the quarterly server revenue situation looks since 1999, with a gap in the IDC data shown in the blue line:

If you inflation adjusted this data, which I need to do, it would make the spending way in the past look bigger than it is. But not enough to diminish the exponential explosion caused by the GenAI wave, which is now being met by a massive, global refresh cycle for non-AI systems at exactly the time that all of those components are seeing prices rise dramatically because there is too much demand and not enough supply.

According to IDC, the world spent $68.9 billion on GPU-accelerated servers in Q1 2026, up 24.8 percent but down 2.5 percent sequentially. Again, I think price increases nearly compensated for shipment declines due to component shortages. No one thinks this is a demand problem at all, including me. GPU systems represented 56.2 percent of overall server revenues, which is astounding until you consider that a datacenter GPU accelerator with its HBM memory costs like $50,000 these days, and will soon be close to $100,000 if not more with next generation of multi-chip devices.

Interestingly, IDC broke out revenues for other XPU-accelerated machines for the first time in its publicly available data, which is a juicy bit of news. In Q1, there were $17.1 billion in XPU systems sold – I think mostly based on Google TPUs and Amazon Web Services Trainiums with a smattering of others from Microsoft, Cerebras Systems, Groq, and a handful of others. These XPU systems accounted for 13.9 percent of server revenues, up from 8.2 percent a year ago and nearly nothing a few years before that.

Add them up, and GPU systems plus XPU systems accounted for $86 billion in revenues, comprising 70.2 percent of overall server revenues. Yup, that’s not a typo. That’s up from 62.2 percent of overall server revenues in the year ago period.

The age of X86 dominance is almost over if you count all of the money in the systems that use them. X86 hosts drove $63.9 billion in sales, down 2.9 percent year on year and down 8.5 percent sequentially. If components were not so expensive and in such short supply, it is likely that the X86 server business would have grown dramatically in the first quarter.

Non-X86 machines, dominated by the homegrown Arm processors created by the hyperscalers and cloud builders as well as Nvidia systems based on its “Grace” Arm server CPU, accounted for $58.7 billion in gear, up by a factor of 2.1X year on year and up 5.8 percent sequentially.

X86 systems had 52.1 percent revenue share in Q1 2026, while non-X86 machines had 47.1 percent share. IBM’s Power Systems RISC machines and System z mainframes accounted for a slice of this non-X86 business, so don’t think that it is all Arm gear. It won’t be too long before we have RISC-V machines on the rise, too. . . .  

Here is how the vendor breakdown looked in Q1 2026:

Dell, having been anointed the chosen OEM by Nvidia, is benefitting mightily from its status, and not just because of its enterprise customers but because of some big sovereign, neocloud, and model builder deals it has taken down. As we showed in our recent financial coverage of Dell and HPE, Dell’s AI server business is an order of magnitude larger than that of HPE, which is why Dell’s overall server business is 5.5X bigger than HPE’s is.

The ODMs that serve the hyperscalers and cloud builders are humming along, constrained only by component supplies. Their business would be much larger if all of that chippery was not so hard to get their hands on.