惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

量子位
博客园_首页
Google DeepMind News
Google DeepMind News
博客园 - Franky
The GitHub Blog
The GitHub Blog
GbyAI
GbyAI
有赞技术团队
有赞技术团队
Microsoft Azure Blog
Microsoft Azure Blog
G
Google Developers Blog
Recent Announcements
Recent Announcements
A
About on SuperTechFans
博客园 - 【当耐特】
博客园 - 三生石上(FineUI控件)
酷 壳 – CoolShell
酷 壳 – CoolShell
美团技术团队
罗磊的独立博客
IT之家
IT之家
博客园 - 聂微东
Stack Overflow Blog
Stack Overflow Blog
Jina AI
Jina AI
腾讯CDC
P
Proofpoint News Feed
Hugging Face - Blog
Hugging Face - Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com

Forbes - Business

Pickleball Slam 4 Preview — History Of The Event And Behind The Scenes Prep With The Players How To Get Masters 2027 Tickets Lottery Dates And Odds ‘Malcolm In The Middle: Life’s Still Unfair’ Is Likely A Wrap For Show Tony Gonzales, Eric Swalwell Will Resign Following Sexual Misconduct Allegations Suspect In Sam Altman Molotov Attack Charged With Attempted Murder Today’s Wordle #1760 Hints And Answer For Tuesday, April 14 Dan Orlovsky Compares Ty Simpson To Brock Purdy, Names Surprising NFC Contender As Fit For 2026 NFL Draft Prospect IndyCar’s Chip Ganassi Racing, OpenAI Hope For ‘Competitive Advantage’ Shingles Altered Achilles Rehab For Pacers Star Tyrese Haliburton, But He’s Back On The Court NYT Pips Today: Hints, Answers And Walkthrough For Tuesday, April 14 LVMH Founder Bernard Arnault’s Fortune Falls $50 Billion This Year Inter Miami CF Kicks Off New Era For South Florida Soccer In Nu Stadium IndyCar’s AJ Foyt Racing Hires Toby Sowery As Reserve Driver IndyCar’s Chip Ganassi Racing Goes Green With Green Sports Alliance Rory McIlroy Claims Second Straight Masters Title At Augusta Rockets Claim Fifth Seed In West Today’s Wordle #1759 Hints And Answer For Monday, April 13 NYT Pips Today: Hints, Answers And Walkthrough For Monday, April 13 Design Details In ‘The Drama’ Delve Deep Into Character AEW Dynasty 2026 Results, Winners And Live Updates On April 12 Former Dodgers Infielder, 3-Time MLB All-Star And Champion, Dies After Cancer Battle Townsend And Wild Secure Double Golds At Pro Pickleball Association Australia Moreton Bay Los Angeles Dodgers Prospect James Tibbs III Is Tearing Up Triple-A Hungary’s Authoritarian Orban—Boosted By Trump—Loses. European Leaders Celebrate. Review: Blackbraid Delivers Exteme Metal Masterclass To Dublin, Ireland Colorado Is Emerging As An Energy Innovation Hub U.S. Military Ships In Strait of Hormuz Violate Ceasefire, Iran Warns (Live Updates) Rosé’s All-Time Sales Chart Record Has Been Beaten IC3 Report Reveals Surge In Cryptocurrency Investment Scams The Top Contenders For The 2026 NCAA Gymnastics All-Around Title
The AI Trade Is Moving Beyond GPUs As Inference Demand Bu...
Andrew Graha · 2026-05-19 · via Forbes - Business
AI

(Photo Illustration by Omar Marques/SOPA Images/LightRocket via Getty Images)

SOPA Images/LightRocket via Getty Images

Over the past two years, the artificial intelligence trade has revolved around one central bet: companies would need far more computing power to train larger models.

That put GPUs, or graphics processing units, in the spotlight.

These chips can handle many calculations at once, making them essential for training large AI models. Since GPUs require vast physical infrastructure to operate at scale, the rush to secure chips quickly became a fight for data center space, power access and broader capacity.

Investors responded by pouring money into the businesses that powered this entire buildout. This trade has obviously done well and likely has room to run. As the market looks beyond the training-fed boom, though, the opportunity set may begin to widen.

The reason is that AI models do not create much value simply by existing. They only do so when people and businesses use them. That moves the discussion from training to inference, which is the process of running trained models to answer questions, complete tasks or power applications. For investors, the difference is not academic.

Training needs enormous computing power while models are being built. Inference, by contrast, depends on steady capacity as AI spreads through search, software, customer service, coding and other workflows. That brings CPUs, or central processing units, back into the discussion because they help coordinate activity across compute.

MORE FOR YOU

That would mark a notable turn. CPUs were long the workhorse of computing before GPUs seized the spotlight during the training boom. Now, CPUs may have a larger role again, not by replacing GPUs, but by helping to manage the steady flow of AI work running across servers, cloud platforms and data centers.

The cost of running AI models could make the inference phase even more compelling for investors. Tokens are the small pieces of text or data an AI model uses to generate a response. As hardware improves, companies appear to be producing each token at lower cost, allowing expensive chips to do more work.

At the same time, demand for tokens is likely to rise as AI agents become more common. Rather than answering a single question and stopping, agents can work through several steps before completing a task. That could drive far more usage across AI systems.

That combination matters for the hyperscalers. If token costs fall while usage grows and pricing holds, companies building AI infrastructure may earn a wider spread. In that case, spending on chips, data centers and power begins to look less like a speculative bet and more like the foundation for a larger operating business.

That broader demand is already showing up in how chip companies describe the inference market. Intel and Arm have both highlighted the growing role of CPUs as inference increases. Intel, for example, has said AI server configurations could shift from roughly eight GPUs for every CPU to about four GPUs for every CPU as inference demand grows. If that forecast proves accurate, it would support the broader point: inference could push AI spending beyond GPUs and deeper into CPUs, servers and the systems needed to run models at scale.

Servers may also become more important. The largest hyperscalers can design custom systems and work directly with global suppliers. Smaller cloud providers and neo-clouds built for inference often need equipment they can deploy quickly and support easily. That could help companies such as Dell and HPE, which sell the servers that carry AI workloads.

Notably, many companies are still preparing for broader AI use. They need to clean up data and connect systems before they can deploy agents across their businesses. That work takes time, but it also suggests inference demand could keep building as more companies move from preparation to real use.

Ultimately, this is not an argument against the GPU-driven trade. It is an argument that inference could spread the next phase of AI spending across a wider set of companies. If models are going to run constantly across real workflows, investors will need to look beyond the companies that trained them and toward the businesses that keep them running.