惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

美团技术团队
阮一峰的网络日志
阮一峰的网络日志
T
The Blog of Author Tim Ferriss
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
宝玉的分享
宝玉的分享
L
LangChain Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Last Week in AI
Last Week in AI
博客园 - 司徒正美
M
MIT News - Artificial intelligence
人人都是产品经理
人人都是产品经理
WordPress大学
WordPress大学
B
Blog RSS Feed
H
Hackread – Cybersecurity News, Data Breaches, AI and More
博客园 - Franky
B
Blog
V
V2EX
J
Java Code Geeks
D
Docker
博客园 - 叶小钗
The Cloudflare Blog
量子位
博客园_首页
MongoDB | Blog
MongoDB | Blog

AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

Cisco: Latest news and insights 2026 network outage report and internet health check Selector targets the network visibility gap in multi-cloud infrastructure AI reshapes cybersecurity workforce priorities as IT teams brace for new risks Top network and data center events of 2026 How AI is transforming network incident response (and where it still falls short) Google opens TPUs to enterprises beyond its own cloud via Blackstone JV AI, cybersecurity skills top IT pay premiums Startup Bolt Graphics promises 5x performance over Nvidia’s best GPU Wireless security is a battle of AI vs. AI NetOps teams look to AI to automate Day 2 operations Digital twins reshape network and data center management Network outages, power failures strain data center resiliency Five takeaways from Cisco's blowout quarter and what it means to customers Cisco to cut nearly 4,000 jobs despite strong growth in AI, enterprise networking Startup SPAN teams with Nvidia to put data center nodes in your backyard Hard drive shortage affecting enterprise storage needs Wi-Fi 8 is closer than you think. Here’s what you need to know Cisco open-sources agentic AI security spec HPE revamps private cloud stack for enterprises rethinking VMware Versa takes aim at fragmented enterprise security with CSPM, orchestration update, and AI agent controls Red Hat opens Ansible to AI agents, within limits Red Hat offers endless Linux support — for a fee Red Hat: Sovereignty is more than just compliance Tech job postings hit three-year high as AI demand fuels hiring rebound HPE memory server targets compute-heavy and agentic AI workloads PCI group begins work on new spec to support bandwidth-hungry apps like AI, HPC Q&A: Quantum physicist Sonia Fernández-Vidal on why classical computing isn't going anywhere OpenAI-led consortium seeks to address AI processing bottlenecks AWS hit by US-East-1 outage after data center thermal event
Nvidia Rubin GPUs may be delayed, slowing the next phase ...
2026-04-09 · via AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

Nvidia’s latest generation of AI chips, the Nvidia Rubin GPUs, expected to ship later this year, may face supply delays amid ongoing geopolitical pressures and supply chain constraints.

The share of Rubin GPUs in Nvidia’s overall shipments was earlier expected to be significantly higher at 29%, but is now projected to remain limited at 22% for 2026, according to TrendForce. The research firm calls HBM4 validation, along with transitioning network interconnects from CX8 to CX9, managing significantly higher power consumption, and optimizing performance under more advanced liquid cooling solutions, as some of the key challenges for this delay.

[ RelatedMore Nvidia news and insights ]

The delay could potentially impact the next wave of AI infrastructure upgrades.

Nvidia did not immediately respond to a request for comment.

Rubin’s role in next-gen AI infrastructure

Current-generation platforms such as Nvidia’s Blackwell and Hopper architectures remain sufficient for training and inference workloads today. But Rubin represents more than a routine GPU upgrade. It is designed to improve the economics of AI at scale by increasing compute density, memory bandwidth, and overall efficiency.

“Rubin is meant to improve cost per token, reduce the number of GPUs required for large workloads, and make large-scale inference economically sustainable. That matters, especially as AI shifts toward agentic workloads that multiply compute demand,” said Sanchit Vir Gogia, chief analyst at Greyhound Research.

The Rubin platform was expected to see early adoption among hyperscalers and AI-native companies, which have the infrastructure to support high-density systems, advanced cooling, and tightly integrated architectures.

Hyperscalers to absorb shock

Typically, hyperscalers lead early adoption of advanced GPUs, deploying them internally and through cloud platforms, with enterprises gaining access later via APIs and services over the next 6-12 months.

“Hyperscalers (will) absorb the initial shock by extending Blackwell lifecycles and prioritizing high-ROI workloads, reducing external capacity. This tightens cloud availability, increases pricing volatility, and elevates the importance of reserved capacity,” said Manish Rawat, semiconductor analyst at TechInsights.

He added that enterprises are likely to face a second-order impact, including constrained access to cloud-based AI infrastructure and delays in the availability of next-generation instances.

Enterprise impact: delays, cost pressure

If Rubin’s rollout is delayed, it is unlikely to halt enterprise AI adoption. But it will affect deployment timelines and cost expectations.

Many enterprise AI strategies are quietly built on the expectation that future hardware will fix today’s inefficiencies. Better performance per dollar, higher density, improved energy efficiency, Gogia said.

This will not result in AI activities being halted, but more phased rollouts, more hybrid consumption, and more aggressive financial scrutiny. Enterprises will prioritise inference-led deployments, smaller clusters, and hybrid architectures that allow them to scale without committing too early to a specific hardware curve.

Rawat noted that this may accelerate diversification toward alternatives like AMD and custom silicon, while increasing focus on software portability beyond CUDA.

AI factory ambitions might recalibrate

The impact will be visible in large-scale initiatives such as private AI clusters and AI factory environments.

“Rubin represents a shift to system-level AI infrastructure optimized for always-on AI factories with superior cost efficiency and throughput. If delayed, enterprises continue deploying on Nvidia Blackwell and Nvidia Hopper, preserving architectural direction but at weaker economics, lower utilization, higher power costs, and greater hardware intensity,” Rawat said.

Given that in the absence of Rubin, enterprises will continue to rely on existing platforms, particularly Nvidia’s Blackwell architecture, TrendForce expects the Blackwell platform to dominate shipments with over 70% share, led by the GB300 or B300 series in 2026.

Rawat stated deployment cycles may stretch quarters, creating a temporary deferral pocket rather than demand destruction. As AI adoption remains workload-driven, this would imply near-term friction, with a likely surge once Rubin platforms become widely available.

SUBSCRIBE TO OUR NEWSLETTER

From our editors straight to your inbox

Get started by entering your email address below.