惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
Project Zero
Project Zero
Engineering at Meta
Engineering at Meta
Microsoft Azure Blog
Microsoft Azure Blog
小众软件
小众软件
T
The Blog of Author Tim Ferriss
S
SegmentFault 最新的问题
量子位
V
Visual Studio Blog
F
Full Disclosure
博客园 - 叶小钗
Recent Announcements
Recent Announcements
G
Google Developers Blog
博客园 - Franky
F
Fortinet All Blogs
有赞技术团队
有赞技术团队
B
Blog
aimingoo的专栏
aimingoo的专栏
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
N
Netflix TechBlog - Medium
L
LINUX DO - 最新话题
S
Security @ Cisco Blogs
Google DeepMind News
Google DeepMind News
B
Blog RSS Feed
Cloudbric
Cloudbric
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Last Week in AI
Last Week in AI
SecWiki News
SecWiki News
H
Hackread – Cybersecurity News, Data Breaches, AI and More
云风的 BLOG
云风的 BLOG
阮一峰的网络日志
阮一峰的网络日志
N
News and Events Feed by Topic
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Microsoft Security Blog
Microsoft Security Blog
D
DataBreaches.Net
人人都是产品经理
人人都是产品经理
Schneier on Security
Schneier on Security
Webroot Blog
Webroot Blog
T
Troy Hunt's Blog
P
Proofpoint News Feed
Application and Cybersecurity Blog
Application and Cybersecurity Blog
Jina AI
Jina AI
PCI Perspectives
PCI Perspectives
Martin Fowler
Martin Fowler
O
OpenAI News
Hacker News: Ask HN
Hacker News: Ask HN
S
Secure Thoughts
P
Privacy International News Feed
I
InfoQ
C
Cyber Attacks, Cyber Crime and Cyber Security

The Register - Special Features: The Future of the Datacenter

AI is rewriting how power flows through the datacenter All aglow about DCs, investors launch $300M at microreactor startup Why do bit barns keep bumping up our bills, Senators ask DC operators Delays? What delays? Oracle insists its $300B cloud contract with OpenAI is on track Galactic Brain space datacenter coming in 2027, pledges startup Aetherflux Activist groups urge Congress to pause datacenter buildouts Bezos-backed Unconventional AI addresses datacenter power Meta and Google tap NextEra to feed their hungry datacenters Datacenters accused of hoarding grid capacity Amazon’s Trainium3 is the latest to conform to Nvidia’s mold Palantir aims to help energy companies meet AI power crunch Datacenters planned for Scotland could drain a loch of power Datacenters must generate their own power or fail HPE to ship rack-scale AI system using AMD's Helios in 2026 London grid crunch delays new housing amid datacenter boom Britain plots atomic reboot as datacenter demand surges OCP learning how to get quantum computers into existing DCs AMD taking AI fight to Nvidia with Helios rack-scale system Nvidia's AI factory dream gets the Omniverse treatment
Qualcomm announces AI accelerators and racks they'll run in
Simon Sharwood Simon Sharwood · 2025-10-28 · via The Register - Special Features: The Future of the Datacenter

The Future of the Datacenter

House of the Snapdragon promises – without much detail – this kit will enable coolly efficient inferencing

Qualcomm has announced some details of its tilt at the AI datacenter market by revealing a pair of accelerators and rack scale systems to house them, all focused on inferencing workloads.

The company offered scant technical details about its new AI200 and AI250 “chip-based accelerator cards”, saying only that the AI 200 supports 768 GB of LPDDR memory per card, and the AI250 will offer “innovative memory architecture based on near-memory computing” and represent “a generational leap in efficiency and performance for AI inference workloads by delivering greater than 10x higher effective memory bandwidth and much lower power consumption.”

Qualcomm will ship the cards in pre-configured racks that will use “direct liquid cooling for thermal efficiency, PCIe for scale up, Ethernet for scale out, confidential computing for secure AI workloads, and a rack-level power consumption of 160 kW.”

In May, Qualcomm CEO Cristiano Amon offered somewhat cryptic statements that the company would only enter the AI datacenter market with “something unique and disruptive” and would use its expertise building CPUs to “think about clusters of inference that is about high performance at very low power.”

However, the house of the Snapdragon’s announcement makes no mention of CPUs. It does say its accelerators build on Qualcomm’s “NPU technology leadership” – surely a nod to the Hexagon-branded neural processing units it builds into processors for laptops and mobile devices.

Qualcomm’s most recent Hexagon NPU, which it baked into the Snapdragon 8 Elite SoC, includes 12 scalar accelerators and eight vector accelerators, and supports INT2, INT4, INT8, INT16, FP8, FP16 precisions.

Perhaps the most eloquent clue in Qualcomm’s announcement is that its new AI products “offer rack-scale performance and superior memory capacity for fast generative AI inference at high performance per dollar per watt” and has “low total cost of ownership.”

That verbiage addresses three pain points for AI operators.

One is the cost of energy to power AI applications. Another is that high energy consumption produces a lot of heat, meaning datacenters need more cooling infrastructure – which also consumes energy and impacts cost.

The third is the quantity of memory available to accelerators, a factor that determines what models they can run – or how many models can run in a single accelerator.

The 768 GB of memory Qualcomm says it’s packed into the AI 200 is comfortably mode than Nvidia or AMD offer in their flagship accelerators.

Qualcomm therefore appears to be suggesting its AI products can do more inferencing with fewer resources, a combination that will appeal to plenty of operators as (or if) adoption of AI workloads expands.

The house of Snapdragon also announced a customer for its new kit, namely Saudi AI outfit Humain, which “is targeting 200 megawatts starting in 2026 of Qualcomm AI200 and AI250 rack solutions to deliver high-performance AI inference services in the Kingdom of Saudi Arabia and globally.”

But Qualcomm says it expects the AI250 won’t be available until 2027. Humain’s announcement, like the rest of this news, is therefore hard to assess because it omits important details about exactly what Qualcomm has created and if it will be truly competitive with other accelerators.

Also absent from Qualcomm’s announcement is whether major hyperscalers have expressed any interest in its kit, or if it will be viable to run on-prem.

The announcement does, however, mark Qualcomm's return to the datacenter after past forays focused on CPUs flopped. Investors clearly like this new move as the company's share price popped 11 percent on Monday. ®