惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 叶小钗
D
Darknet – Hacking Tools, Hacker News & Cyber Security
S
SegmentFault 最新的问题
博客园 - 三生石上(FineUI控件)
雷峰网
雷峰网
WordPress大学
WordPress大学
有赞技术团队
有赞技术团队
博客园 - 【当耐特】
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
V
V2EX
V
Visual Studio Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园 - 聂微东
P
Proofpoint News Feed
Last Week in AI
Last Week in AI
U
Unit 42
W
WeLiveSecurity
博客园 - Franky
Recent Announcements
Recent Announcements
Hacker News - Newest:
Hacker News - Newest: "LLM"
Attack and Defense Labs
Attack and Defense Labs
月光博客
月光博客
The Cloudflare Blog
Spread Privacy
Spread Privacy
腾讯CDC
P
Privacy International News Feed
N
News and Events Feed by Topic
AWS News Blog
AWS News Blog
NISL@THU
NISL@THU
T
Troy Hunt's Blog
小众软件
小众软件
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
Microsoft Security Blog
Microsoft Security Blog
L
Lohrmann on Cybersecurity
Webroot Blog
Webroot Blog
Y
Y Combinator Blog
量子位
P
Palo Alto Networks Blog
N
News and Events Feed by Topic
V
Vulnerabilities – Threatpost
K
Kaspersky official blog
IT之家
IT之家
T
Threat Research - Cisco Blogs
Cloudbric
Cloudbric
云风的 BLOG
云风的 BLOG
C
Check Point Blog
Blog — PlanetScale
Blog — PlanetScale
爱范儿
爱范儿
G
Google Developers Blog
S
Secure Thoughts

TechWire Asia

Microsoft ERP Implementation Partners Roundup: Which Is Best For Your Business? Semiconductor talent retention now comes with a price tag, and it keeps climbing AI token usage is not productivity. Here's what to measure instead Oracle brings governed AI agents to Fusion Applications Nvidia expands Japan AI infrastructure and robotics push AI Appreciation Day 2026 puts trust and governance in focus NVIDIA pours its full stack into Japan. The flip side of its China lockout? Malaysia's digital regulations are becoming a real cost for its startups Malaysia's AI data center vision: How EdgeConneX is building for the future Southeast Asia tech funding doubled to $7.4 billion. One company took most of it SK Hynix's Nasdaq listing raises $26.5 billion to fund Korea's AI memory expansion OpenAI launches GPT-5.6 for coding, cyber and science Meta rolls out Muse Image AI model for Instagram, WhatsApp, and advertisers Malaysia businesses face AI and password cybersecurity risks How AI workloads will test APAC mobile networks Enterprise AI costs don't have to spiral, argues ManageEngine Microsoft launches $2.5B Frontier Company for enterprise AI FIFA World Cup: How To Win Fans in APAC With Technology Kanga enters a new phase of global growth and launches Kanga Global Vertiv ramps up manufacturing in Johor's tightening data centre market U Mobile completes migration to own ULTRA5G network after DNB exit Anthropic Claude models launch in Microsoft Foundry on Azure Asia built the AI infrastructure boom. The BIS just flagged who's exposed if it stalls. Why Apple is lobbying Washington to buy China’s memory chips Nvidia-backed Firmus plans 170,000-GPU Batam AI data centre Taiwan robot makers march into humanoid systems IBM claims world’s first sub-1 nm chip technology using nanostack design Can Alibaba bridge Malaysia’s SME talent gap via agentic AI for business? Huawei’s new tech explains why mobile AI network tech is no longer optional Apple-Intel chip deal faces years-long production timeline China beats US in TOP500 ranking with world’s fastest supercomputer The global memory squeeze hits the Mainland China PC market, leading to a decline IBM joins OpenAI cyber program for vulnerability detection Is the Shopee ChatGPT integration the blueprint for the future of Southeast Asian e-commerce? How the global AI boom dropped a record RM1.127 trillion trade windfall on Malaysia Philippines expands Google Cloud public sector AI partnership South Korea takes a positive spin on AI Apple's price hikes trace the memory chip shortage straight back to Asia Why enterprises need clearer accountability for AI agents Google sues Chinese network over AI text phishing scams AI Won't Fix Broken Personalisation: Braze Report Reveals How Media and Entertainment Can Drive Real Success Across APAC Anthropic builds out Claude as OpenAI and Google stay ahead How APAC firms are handling software supply chain security Meta Business Agent turns WhatsApp into a salesperson, and Southeast Asia will decide if it works CrowdStrike: Chinese hackers lead tech sector espionage threats NVIDIA deals in South Korea cover AI memory, cloud and robotics Alibaba Cloud's Johor region launch comes packaged with an agentic AI push in Malaysia Digital Realty Malaysia is open and already looking beyond Cyberjaya AI’s invisible metal: Why tin demand is surging, and supplies are running thin WeChat is opening up to AI agents, and Southeast Asia’s super apps should be nervous TNG eWallet is eyeing agentic payments and its CEO sees Malaysia’s regulatory climate as encouraging AI data centres could double power and water use by 2030 TNG eWallet is no longer just a payment app, and the numbers prove it Nvidia GTC Taipei recap: RTX Spark, Vera, data centres and more Alipay wants AI agents to handle your payments. But who’s really in control? Huawei’s Her’s Law eyes AI chips as China reduces Nvidia reliance Kong Konnect now available in Singapore AWS is quietly building one of Southeast Asia’s most ambitious green data centre footprints China launches offshore wind-powered underwater AI data centre Has Huawei just rewritten the rules of chip design? OpenAI Daybreak and the patching cycle AirTrunk to invest MYR12 billion in Johor data centres China orders Meta to unwind Manus AI acquisition Kong reveals ‘agent-to-agent communication’ critical for Asian enterprises Huawei picked Malaysia for its biggest AI move outside China. Anwar told you exactly why. DeepSeek launches V4 model adapted for Huawei AI chips MATCH Act passes first hurdle–targeting semiconductor tools, not just chips The real cost of AI in APAC isn’t the software licence–it’s the mess underneath Cisco shows Universal Quantum Switch prototype to connect quantum systems The global smartphone market just had its worst quarter in two years, and memory is to blame Google Cloud introduces AI agent platform and new TPU chips at Next 2026 Tesla plans to use Intel 14A chips for Terafab project Meta deploys tracking tool to train AI on employee workflows Tuned Global’s service manipulation detector for streaming clients and rights holders Malaysia is rushing into AI faster than anyone. Its governance gap is the price Apple’s CEO transition puts a hardware engineer in charge–at exactly the right moment Memory shortage to persist through 2027 as supply lags demand xAI provides GPU infrastructure to Cursor for AI model training Amazon Leo just gave Southeast Asia’s satellite internet market a second player Meta extends Broadcom deal to develop AI chips Can Malaysia Build a USD1 Trillion Economy on the Strength of Its Geography? How will MyDigital ID progress in Malaysia? Southeast Asia leads the world in AI optimism. Its governance frameworks are nowhere near ready. A chatbot is not an AI strategy Japan is building physical AI it controls–and its biggest companies are all in India is leading Asia’s agentic AI adoption race. The rest of the region is still catching up. Ericsson frames 6G as an intelligent fabric Mandatory AI literacy: China joins the UAE and India. Where is Southeast Asia? AWS AI revenue hits US$15 billion. Andy Jassy says the hard part is keeping up with demand Minor Hotels builds data and AI platform with Google Cloud The MATCH Act would cut off China’s last chipmaking lifeline–Asia is already feeling it Amperity expands to Australian AWS Regions and invests in local talent Chinese memory giants are scaling fast, and the AI boom is giving them cover Intel joins Musk’s Terafab AI chip project with Tesla and SpaceX TikTok’s second data centre in Finland a European push Custom AI chips, 3.5 gigawatts, and a quiet SEC clause: the Broadcom deal explained Kong names Bruce Felt as chief financial officer DeepSeek V4 points to growing use of Huawei chips in AI models Microsoft to invest $10 billion in Japan for AI and cybersecurity Which CRMs offer the most powerful reporting tools?
Microsoft to deploy AMD Helios AI systems on Azure
Muhammad Zulhusni · 2026-07-21 · via TechWire Asia
  • Microsoft will deploy AMD Helios on Azure.
  • Azure will add AMD EPYC VMs and Pensando networking.

Microsoft plans to deploy AMD’s Helios rack-scale computing platform on Azure under an expanded infrastructure partnership covering processors, graphics chips, networking hardware, and software. The systems will support AI inference for Microsoft, Azure customers, and Azure AI services.

Microsoft also plans to offer the Helios infrastructure through its upcoming ND MI455X v7 virtual machines. The instances are designed for production-scale reasoning, search, and agentic inference workloads.

Helios targets large-scale AI inference

Helios combines AMD Instinct MI455X GPUs, sixth-generation EPYC “Venice” processors, Pensando networking technology, and the ROCm software platform. The reference design includes 72 MI455X GPUs, EPYC host processors, and Pensando “Vulcano” networking hardware in a double-wide rack.

AMD developed the system around Meta’s Open Rack Wide specification, which was submitted through the Open Compute Project. The design provides a common rack architecture for high-density computing, power delivery, and liquid cooling.

“AMD and Microsoft have spent years building high-performance infrastructure together, and today we’re extending that partnership across the full stack of AMD AI solutions on Azure,” AMD Chair and CEO Lisa Su said. “Microsoft’s new AMD deployments mark an important milestone as we deliver leadership compute solutions to Azure customers and scale the next generation of AI infrastructure together,” Su said.

Each MI455X GPU includes 432GB of HBM4 memory and provides memory bandwidth of up to 19.6TB per second, according to AMD. A full Helios rack provides up to 31TB of combined HBM4 memory across its 72 accelerators.

AMD rates the system at up to 1.4 exaflops of FP8 compute and 2.9 exaflops of FP4 compute. The figures represent peak manufacturer specifications rather than measured Azure application performance.

The MI455X GPUs handle the primary AI calculations, while the EPYC processors support host computing, workload coordination, and data movement. Microsoft will initially use Helios for frontier-model inference across its own services, Azure AI services, and customer applications.

Although Helios supports both model training and inference, Microsoft’s announced ND MI455X v7 deployment focuses on running trained models at scale. The company has not provided a timetable for customer access to the instances.

“Customers are looking for AI infrastructure that is optimised for a wide range of workloads, from training and inference to data preparation, search, and reinforcement learning,” Microsoft Chairman and CEO Satya Nadella said. “Through our collaboration with AMD, we are expanding the Azure infrastructure portfolio with AMD Helios to give customers the performance, scale, and choice they need to build and run the next generation of AI applications,” Nadella said.

Enterprise customers will also be able to access AMD-based infrastructure through Microsoft Foundry Managed Compute. The service hosts open-source models on dedicated GPU capacity, with Microsoft managing the GPU topology, runtime, container image, and security patching.

Customers select the model, accelerator family, deployment template, and scaling settings. Managed Compute remains in public preview, has no service-level agreement, and is not currently recommended by Microsoft for production workloads.

Azure expands CPU and networking infrastructure

The partnership also covers two Azure virtual machine series powered by AMD’s sixth-generation EPYC “Venice” processors. Azure HDv2 will include nearly 500 physical EPYC CPU cores, 4TB of RAM, 32TB of local NVMe storage, and 400Gb Azure Boost networking.

Microsoft is positioning HDv2 for CPU-intensive AI workloads, including data preparation, search, reinforcement learning, agent coordination, and data pipelines. The company said the VMs will handle processes that prepare data, coordinate workloads, and support GPU-based training and inference.

Azure HXv2 will include 176 sixth-generation EPYC cores running at more than 5GHz, with 50% more addressable cache per core than the previous HX generation. Microsoft plans to offer configurations with nearly 2TB or 4TB of RAM and 800Gb InfiniBand connectivity.

HXv2 will retain AMD’s 3D V-Cache technology, which is also used in the current Azure HX series. Microsoft and AMD introduced the first HX virtual machines in 2023 for electronic design automation workloads.

The HXv2 series will target electronic design automation, scientific simulation, engineering analysis, and distributed-memory computing. Microsoft also identified register-transfer level simulation as a target workload for the new series.

The 800Gb InfiniBand connection is intended to support large-scale Message Passing Interface simulations across distributed computing environments. AMD also uses Azure HX infrastructure for electronic design automation as it develops future EPYC processors and Instinct accelerators, according to AMD Executive Vice-President and CTO Mark Papermaster.

Microsoft has not announced pricing, launch dates, or the Azure regions where HDv2 and HXv2 will initially be available.

AMD and Microsoft are also expanding their work on cloud networking. Azure already uses AMD Pensando data processing units, or DPUs, to handle infrastructure functions separately from a server’s main processors.

Pensando DPUs offload networking, storage, security, and encryption services from host CPUs. Azure Boost also moves selected networking, storage, and virtualisation processing onto dedicated hardware and software.

The expanded agreement will extend AMD’s role in Azure connection processing and backend networking, although the companies have not disclosed how each component will be deployed. Within Helios, UALink-over-Ethernet provides scale-up connectivity among the rack’s GPUs, while Pensando Ethernet components support scale-out networking between racks and clusters.

ROCm provides the software environment used to develop and run workloads on AMD accelerators. Its inclusion gives Helios a common software layer across the rack’s GPU infrastructure.

AMD expects Helios-based systems to enter volume deployment in the second half of 2026. Microsoft has not announced when ND MI455X v7 instances will become available to Azure customers or which regions will receive them first.

Want to learn more about AI and big data from industry leaders? Check out AI & Big Data Expo taking place in Amsterdam, California, and London. The comprehensive event is part of TechEx and is co-located with other leading technology events, click here for more information.

Tech Wire Asia is powered by TechForge Media. Explore other upcoming enterprise technology events and webinars here.