惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Last Week in AI
Last Week in AI
C
CERT Recently Published Vulnerability Notes
博客园 - 叶小钗
大猫的无限游戏
大猫的无限游戏
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
月光博客
月光博客
T
Tailwind CSS Blog
博客园 - 三生石上(FineUI控件)
Jina AI
Jina AI
S
SegmentFault 最新的问题
人人都是产品经理
人人都是产品经理
Hugging Face - Blog
Hugging Face - Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
WordPress大学
WordPress大学
J
Java Code Geeks
V
Visual Studio Blog
腾讯CDC
博客园 - 【当耐特】
博客园 - 司徒正美
小众软件
小众软件
宝玉的分享
宝玉的分享
博客园 - Franky
量子位
有赞技术团队
有赞技术团队
The Cloudflare Blog
Apple Machine Learning Research
Apple Machine Learning Research
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
雷峰网
雷峰网
美团技术团队
阮一峰的网络日志
阮一峰的网络日志
酷 壳 – CoolShell
酷 壳 – CoolShell
N
News | PayPal Newsroom
D
Docker
Google Online Security Blog
Google Online Security Blog
博客园 - 聂微东
A
About on SuperTechFans
S
Security Affairs
N
News and Events Feed by Topic
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
S
Securelist
T
The Exploit Database - CXSecurity.com
爱范儿
爱范儿
MyScale Blog
MyScale Blog
V
Vulnerabilities – Threatpost
S
Security @ Cisco Blogs
T
Threatpost
Scott Helme
Scott Helme
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Application and Cybersecurity Blog
Application and Cybersecurity Blog
P
Palo Alto Networks Blog

AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

Cisco: Latest news and insights 2026 network outage report and internet health check Selector targets the network visibility gap in multi-cloud infrastructure AI reshapes cybersecurity workforce priorities as IT teams brace for new risks Top network and data center events of 2026 How AI is transforming network incident response (and where it still falls short) Google opens TPUs to enterprises beyond its own cloud via Blackstone JV AI, cybersecurity skills top IT pay premiums Startup Bolt Graphics promises 5x performance over Nvidia’s best GPU Wireless security is a battle of AI vs. AI NetOps teams look to AI to automate Day 2 operations Digital twins reshape network and data center management Network outages, power failures strain data center resiliency Five takeaways from Cisco's blowout quarter and what it means to customers Cisco to cut nearly 4,000 jobs despite strong growth in AI, enterprise networking Startup SPAN teams with Nvidia to put data center nodes in your backyard Hard drive shortage affecting enterprise storage needs Wi-Fi 8 is closer than you think. Here’s what you need to know Cisco open-sources agentic AI security spec HPE revamps private cloud stack for enterprises rethinking VMware Versa takes aim at fragmented enterprise security with CSPM, orchestration update, and AI agent controls Red Hat opens Ansible to AI agents, within limits Red Hat offers endless Linux support — for a fee Red Hat: Sovereignty is more than just compliance Tech job postings hit three-year high as AI demand fuels hiring rebound HPE memory server targets compute-heavy and agentic AI workloads PCI group begins work on new spec to support bandwidth-hungry apps like AI, HPC Q&A: Quantum physicist Sonia Fernández-Vidal on why classical computing isn't going anywhere OpenAI-led consortium seeks to address AI processing bottlenecks AWS hit by US-East-1 outage after data center thermal event Gluware's Titan rises to meet Mythos network vulnerability challenge AMD launches AI-targeted PCIe cards for current servers Supply constraints, optical advances dominate Arista's Q1 Lumen advances cloud networking vision with $475M Alkira buy HPE bolsters autonomous network operations for Mist, Aruba Central Netskope launches AI agents for SOC and NOC automation Intel, behind in AI chips, bets on quantum and neuromorphic processors Switch storm coming: Gartner forecasts price hikes, long lead times for enterprise data center switches Extreme moves toward autonomous networking with advanced AI agent, management tools Broadcom bets big on VMware Cloud Foundation 9.1 IBM unveils its blueprint to help enterprises run AI at the core of their business Ruckus Networks on the move again, this time acquired by Belden for $1.85 billion AMD and Intel partner to deliver AI performance advancement Cisco grabs Astrix to secure AI agents Beyond the pitch: A look at Atlético Madrid's connected stadium StarlingX 12.0 is right on time for mixed-hardware edge deployments Cisco nerds out: May the Fourth be with your AI assistant Memory shortage and cost surge push enterprises toward the cloud Extreme Networks: Memory advantage, Wi-Fi 7 and competitive flux drive momentum Scenes from the great data center revolt Enterprise Spotlight: Transforming software development with AI When 170,000 people show up: Network refresh readies Churchill Downs for Kentucky Derby IT certification pay surges as noncertified skills slump QuEra claims quantum error correction breakthrough with 2-to-1 qubit ratio HPE expands ProLiant line with rugged edge servers Deconstructing the data center: A massive (and massively liberating) project Cisco bolsters security, AI support in latest SD-WAN release The era of chatbot AIOps is fading as agentic AI gains traction Auvik bets agentic AI can fill the networking skills gap AI data flows force rethink of data center networking at Backblaze Nvidia's 'AI insurance policy' balances immediate and future AI approaches Cirrascale to offer on-prem Google Gemini models Space data-center news: Roundup of extraterrestrial AI endeavors Network jobs watch: Hiring, skills and certification trends Cisco switch aimed at building practical quantum networks How AI is changing copper, fiber networking Almost 40% of data center projects will be late this year, 2027 looks no better It’s the end of set-and-forget security Google bets on workload-specific TPUs with 8t and 8i launch SUSE bets automated migration can break VMware's grip on virtualization How Zero Networks is closing the network enforcement gap for AI agents Cloudflare wants to rebuild the network for the age of AI agents AI fuels wireless talent shortage Broadcom's Facebook friend will help train it to accelerate AI workloads Data centers are costing local governments billions Equinix offering targets automated AI-centric network operations AI shifts IT roles from operator to orchestrator IBM unveils security services for thwarting agentic attacks, automating threat assessment Maine to put brakes on big data centers as AI expansion collides with power limits Satellite backhaul service Globalstar has a new, rich owner amid challenging market conditions DNS security is often inadequate, and network engineers should get more involved Curious about quantum? Check out training options from ISC2, IBM, AWS and more Cisco just made moves to own the AI infrastructure stack Data centers are moving inland, away from some traditional locations Fixing encryption isn't enough. Quantum developments put focus on authentication Intel: Latest news and insights Linux 7.0 debuts with some big changes for networking Intel secures Google cloud and AI infrastructure deal OpenAI puts part of Stargate project on hold over runaway power costs Broadcom strikes chip deals with Google, Anthropic Cisco to acquire Galileo for AI observability Neoclouds gain momentum in a supply-constrained world Lumen: Upstream network visibility is enterprise security's new front line Yael Nardi joins Minimus as Chief Business Officer to head growth strategy What is AI networking? How it adds intelligence to your infrastructure Google owns the most AI compute, and it built it its way Aria Networks raises $125M and debuts its approach for AI-optimized networks Intel bets on Terafab to help it reassert itself in the AI chip race New v2 UALink specification aims to catch up to NVLink Cisco joins Anthropic’s multivendor effort to secure AI software
Nvidia Rubin GPUs may be delayed, slowing the next phase of AI infrastructure
2026-04-09 · via AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

Nvidia’s latest generation of AI chips, the Nvidia Rubin GPUs, expected to ship later this year, may face supply delays amid ongoing geopolitical pressures and supply chain constraints.

The share of Rubin GPUs in Nvidia’s overall shipments was earlier expected to be significantly higher at 29%, but is now projected to remain limited at 22% for 2026, according to TrendForce. The research firm calls HBM4 validation, along with transitioning network interconnects from CX8 to CX9, managing significantly higher power consumption, and optimizing performance under more advanced liquid cooling solutions, as some of the key challenges for this delay.

[ RelatedMore Nvidia news and insights ]

The delay could potentially impact the next wave of AI infrastructure upgrades.

Nvidia did not immediately respond to a request for comment.

Rubin’s role in next-gen AI infrastructure

Current-generation platforms such as Nvidia’s Blackwell and Hopper architectures remain sufficient for training and inference workloads today. But Rubin represents more than a routine GPU upgrade. It is designed to improve the economics of AI at scale by increasing compute density, memory bandwidth, and overall efficiency.

“Rubin is meant to improve cost per token, reduce the number of GPUs required for large workloads, and make large-scale inference economically sustainable. That matters, especially as AI shifts toward agentic workloads that multiply compute demand,” said Sanchit Vir Gogia, chief analyst at Greyhound Research.

The Rubin platform was expected to see early adoption among hyperscalers and AI-native companies, which have the infrastructure to support high-density systems, advanced cooling, and tightly integrated architectures.

Hyperscalers to absorb shock

Typically, hyperscalers lead early adoption of advanced GPUs, deploying them internally and through cloud platforms, with enterprises gaining access later via APIs and services over the next 6-12 months.

“Hyperscalers (will) absorb the initial shock by extending Blackwell lifecycles and prioritizing high-ROI workloads, reducing external capacity. This tightens cloud availability, increases pricing volatility, and elevates the importance of reserved capacity,” said Manish Rawat, semiconductor analyst at TechInsights.

He added that enterprises are likely to face a second-order impact, including constrained access to cloud-based AI infrastructure and delays in the availability of next-generation instances.

Enterprise impact: delays, cost pressure

If Rubin’s rollout is delayed, it is unlikely to halt enterprise AI adoption. But it will affect deployment timelines and cost expectations.

Many enterprise AI strategies are quietly built on the expectation that future hardware will fix today’s inefficiencies. Better performance per dollar, higher density, improved energy efficiency, Gogia said.

This will not result in AI activities being halted, but more phased rollouts, more hybrid consumption, and more aggressive financial scrutiny. Enterprises will prioritise inference-led deployments, smaller clusters, and hybrid architectures that allow them to scale without committing too early to a specific hardware curve.

Rawat noted that this may accelerate diversification toward alternatives like AMD and custom silicon, while increasing focus on software portability beyond CUDA.

AI factory ambitions might recalibrate

The impact will be visible in large-scale initiatives such as private AI clusters and AI factory environments.

“Rubin represents a shift to system-level AI infrastructure optimized for always-on AI factories with superior cost efficiency and throughput. If delayed, enterprises continue deploying on Nvidia Blackwell and Nvidia Hopper, preserving architectural direction but at weaker economics, lower utilization, higher power costs, and greater hardware intensity,” Rawat said.

Given that in the absence of Rubin, enterprises will continue to rely on existing platforms, particularly Nvidia’s Blackwell architecture, TrendForce expects the Blackwell platform to dominate shipments with over 70% share, led by the GB300 or B300 series in 2026.

Rawat stated deployment cycles may stretch quarters, creating a temporary deferral pocket rather than demand destruction. As AI adoption remains workload-driven, this would imply near-term friction, with a likely surge once Rubin platforms become widely available.

SUBSCRIBE TO OUR NEWSLETTER

From our editors straight to your inbox

Get started by entering your email address below.