惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
L
LangChain Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
罗磊的独立博客
J
Java Code Geeks
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
博客园 - 叶小钗
小众软件
小众软件
博客园 - Franky
D
Docker
Google DeepMind News
Google DeepMind News
Microsoft Azure Blog
Microsoft Azure Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
U
Unit 42
宝玉的分享
宝玉的分享
C
Check Point Blog
B
Blog
V
V2EX
博客园 - 三生石上(FineUI控件)
MyScale Blog
MyScale Blog
The Cloudflare Blog
博客园 - 聂微东
博客园_首页
Engineering at Meta
Engineering at Meta

VMware Blogs

Diagnostics for VMware Cloud Foundation (VCF) 9.1 with Old Versions of VCF Components Mastering Infrastructure Policies in VMware Cloud Foundation Automation 9.1 Modernizing the Private Cloud: Why VCF 9.1 Lifecycle Management is a Game Changer Announcing the VMware Cloud Foundation 9.1 Upgrade Planning Tool VCF Breakroom Chats Episode 86 – Containers Made Easy: The New “Container-as-a-Service” in VCF 9.1 Securing Your VCF 9.1 Infrastructure with the Symantec Identity Security Platform Virtually Speaking: The AI Reality Check with Dave Linthicum Zero Touch Provisioning: Activating Edge Sites with VMware Cloud Foundation Edge 9.1 VCF Breakroom Chats Episode 85 – Cloning Success at Scale: Inside VCF 9.1’s App Stack Formation VMware Cloud on AWS の使用状況を確認できる API Unlocking the Full Potential of Programmable Infrastructure with VMware Cloud Foundation 9.1 – New Features and Capabilities Smarter Patching at Scale: Vulnerability Assessment and Remediation with VMware Tanzu Platform Encrypted vMotion Offload to Intel QAT in VMware Cloud Foundation 9.1 Deepen Your Expertise: Four Key Benefits of Attending Increase Deployment Flexibility with VCF Edge Automation 1.0.3 Avi Advantage: Automating Certificate Management of VCF Workloads More Memory, Less Effort: Configuring Memory Tiering in VCF 9.1 VCF 9.1 Licensing: Programmatic, Centralized, and Built to Scale Why APJ Networking Professionals Need Private Cloud Expertise VCF 9.1 Networking: Simpler VPC Connectivity Control VCF 9.1 Networking: Exploring Network Services for Virtual Private Clouds VCF Networking 9.1: Seamless DDI Integration with Infoblox The Open Source Advantage: Building from Source for Ultimate Security Expand Shared VMDKs with Clustered Applications in VMware vSAN for VCF 9.1 Monetizing Zero-Trust Security with VCF 9.1 and VMware vDefend VMware vSAN Protection and Recovery Enhancements for VCF 9.1 Deliver Production SQL Server DBaaS with VMware Data Services Manager 9.1 Maximizing Profitability: VCF 9.1 Cost-Focused Approach for VMware Cloud Service Providers Modernizing Your Infrastructure: Introducing VMware Cloud Foundation 9.1 to VCSPs VCF 9.1 is Available: Explore the New Features in Hands-on Labs
AI with VCF 9.1 on AMD GPUs: Build with open frameworks a...
shobhit bhut · 2026-05-05 · via VMware Blogs

Artificial Intelligence (AI) is rapidly emerging as one of the most transformational technologies of our time, with massive potential to redefine how organizations operate, compete, and innovate. With the ability to learn from vast amounts of data, identify complex patterns, and make intelligent decisions, AI is moving enterprises towards smarter, faster, and highly adaptive operations at-scale. However, enterprises face tremendous challenges in implementing AI – privacy, security, compliance, cost, and governance when deploying AI workloads.

At Explore Vegas 2025, Broadcom and AMD announced the expansion of our collaboration to advance AI for enterprises. With VCF on AMD Instinct GPUs, our partnership is addressing these enterprise challenges.

Announcement 

With the release of VMware Cloud Foundation 9.1, Broadcom and AMD are announcing the next step in furthering our mission of helping enterprises run and manage their AI workloads. Broadcom and AMD will add support for VCF on AMD Instinct MI350 Series GPUs. This new release will help enterprises run and manage AI workloads, unlocking powerful virtualization capabilities and reducing TCO.

Details of the Release

Let’s get into the details.

1. Exceptional Performance at Lower TCO (Total Cost of Ownership):

We are adding several capabilities to accelerate performance with low TCO for enterprises. Here are the capabilities in this category:

  • High-performance infrastructure powered by the AMD Instinct MI350 GPU series: The MI350 series delivers a 4x generational increase in AI compute over previous AMD GPUs. With up to 10 PetaFLOPS of FP4 and FP6 operations, these GPUs are optimized for highly efficient training and inferencing.
  • Larger-sized models on a single GPU: Equipped with 288 GB of HBM3E memory, the MI350 series GPUs can run models with up to 520 billion parameters on a single GPU. This reduces the number of GPUs required for deployment, significantly reducing infrastructure costs.
  • Hot-add and remove virtual hardware for maximizing model performance: Enterprises can scale AI workloads dynamically by adding or removing vCPUs, memory, storage, and network adapters to running VMs without downtime. Removing virtual hardware helps right-size the infrastructure, optimizing TCO seamlessly.

2. Simplified Infrastructure Management:

These capabilities streamline operations and ensure high availability:

  • vSphere High Availability for automated AI workload recovery: By pooling VMs and hosts into a cluster, AI apps running on VMs can be restarted on a different host if an issue occurs, reducing downtime.
  • ESX Live Patch for non-disruptive security updates for AI hosts: Administrators can perform ESX updates by applying critical security patches to running hosts in partial maintenance mode. This helps continuous operations of AI workloads by ensuring uninterrupted access to compute resources. 
  • Storage vMotion and Snapshots for ensuring continuous operations for AI workloads. Storage vMotion allows for the seamless migration of running VMs’ data to a different datastore, while storage snapshots preserve the exact state of virtual disks at specific points in time.

3. Choice of Open Standards and Frameworks

AMD and Broadcom have collaborated to build  an infrastructure on which AI workloads can be created using open standards and frameworks. This enhances flexibility while continuing to deliver the consistent operations experience that VCF provides. Let’s understand the details on this:

  • Industry Standard Frameworks, PyTorch and vLLM: We provide native support for PyTorch and vLLM, allowing enterprises to move workloads between accelerators with minimal code changes.
  • Standardized Architecture through OPEA (Open Platform for Enterprise AI): By leveraging OPEA, enterprises can build on vetted, industry-standard blueprints for RAG and generative AI.
  • Massive Quantity of Available Open-source models: Through the partnership with Hugging Face, over 1.8 million open-source models are available out-of-the-box, ready to be deployed on AMD GPUs.

Want to know more? 


Discover more from VMware Cloud Foundation (VCF) Blog

Subscribe to get the latest posts sent to your email.