惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Engineering at Meta
Engineering at Meta
G
Google Developers Blog
WordPress大学
WordPress大学
M
MIT News - Artificial intelligence
D
DataBreaches.Net
云风的 BLOG
云风的 BLOG
爱范儿
爱范儿
Microsoft Security Blog
Microsoft Security Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Blog — PlanetScale
Blog — PlanetScale
T
Tailwind CSS Blog
S
SegmentFault 最新的问题
阮一峰的网络日志
阮一峰的网络日志
博客园 - 三生石上(FineUI控件)
酷 壳 – CoolShell
酷 壳 – CoolShell
Recent Announcements
Recent Announcements
T
The Blog of Author Tim Ferriss
I
InfoQ
MyScale Blog
MyScale Blog
V
V2EX
B
Blog
罗磊的独立博客
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More

Supermicro Data Center Stories

Supermicro Offers Edge AI Powerhouse in Compact, Fanless SYS-E103-14P Bringing Rack-Scale AI to the Enterprise Together with Cisco AI Inference Solutions for Financial Trading Powering Enterprise Agentic AI: NVIDIA Nemotron 3.5 Lightning and Supermicro Join Supermicro at FMS 2026 – Here’s What to Expect Inside Supermicro Unlocking Flexible, High-Availability Edge Infrastructure with Supermicro Servers Supermicro’s 100% Liquid-Cooled NVIDIA Vera CPU Rack: Built for Agentic AI and HPC at Scale Rethinking Retail Edge Infrastructure: Why Efficiency and Scalability Matter Supermicro NVIDIA Blackwell Systems Demonstrate Linear Scalability for MLPerf Training v6.0 Right-Sizing Edge AI: Choosing the Right Processor Type for Inferencing SPEC CPU 2026 Benchmark Suites Released: See Supermicro's Strong Results Supermicro Announces General Availability of the NVIDIA DGX GB300-Powered Super AI Station at COMPUTEX 2026 Building More Efficient, Reliable AI Infrastructure with NVIDIA's Photonics Switches and NVIDIA Vera Rubin A Closer Look: Building a Modern Data Center with Supermicro Networking and Switching Solutions Supermicro 5U PCIe GPU Servers Using AMD Instinct™ MI350P GPUs Provides Ready-to-Deploy Enterprise AI for Your Existing Infrastructure From Platforms to Production: How Supermicro Is Powering the Rise of AI Factories Powering the Next Wave of AI Infrastructure with Cloud-Native MegaDC Systems Secure AI: Supermicro’s HGX B300 & GB300 NVL72 with NVIDIA Confidential Computing Experience the AI Factory SuperCloud Director: Operationalizing NeoCloud Infrastructure with NVIDIA NCX Infra Controller Built to Accelerate: Supermicro Delivers Powerful AI Factory Clusters and Intelligent Data Platforms for Enterprises Supermicro Announces General Availability of NVIDIA GB300-Powered Super AI Station at GTC 2026
Supermicro Leads Whisper Benchmark in MLPerf v6.0 with NV...
Supermicro Experts · 2026-04-01 · via Supermicro Data Center Stories

In the latest MLPerf v6.0 Inference Datacenter, closed division results, Supermicro posted the top result on the Whisper benchmark, achieving over 50,562 samples per second. This result reflects our ongoing collaborations with NVIDIA and our shared goal of ensuring that our mutual customers can benefit from the performance demonstrated in the MLPerf results.

The submission used the Supermicro AS-8126GS-NB3RT, incorporating the 8-GPU NVIDIA HGX B300 system, featuring Blackwell Ultra GPUs with 5th Generation NVIDIA NVLink 1.8 TB/s, 2.3 TB of HBM3e GPU memory per system, and dual AMD EPYC™ 9575F CPUs. NVIDIA-powered solutions support a wide range of inference workloads—from generative recommenders to language, vision, and speech AI models—making them a practical platform for varied deployment requirements.

HGX-systems-portfolio-shot (2)

About the Whisper Benchmark

Whisper is an open-source model from OpenAI and serves as the industry’s “gold standard” baseline for speech AI. Whenever a new speech-to-text model is released, developers typically measure its performance against Whisper, which serves as a useful reference point for comparing systems across speech-processing workloads.

The benchmark covers a range of real-world speech workloads:

  • Transcribing speech in its native language or translating it directly into English.

  • Creating highly accurate subtitles for videos or podcasts without needing a human editor.

  • Processing sensitive audio—such as medical dictation or legal meetings—locally on private servers, without sending data to the cloud.

  • Converting voice notes into text that is then fed into large language models such as Gemini or GPT-4 for summarization.

The benchmark also tests a system’s ability to handle challenging speech conditions:

  • Heavy Accents: Regional dialects that typically trip up AI systems.

  • Background Noise: Technical chatter, ambient sound, or wind interference.

  • Technical Jargon: Specialized terminology across 99 different languages.

Software Efficiency and Infrastructure ROI

Our systems are optimized for use with NVIDIA inference software, including NVIDIA Dynamo, which can deliver efficiency gains on existing AI infrastructure—helping lower cost per token and improve return on investment for operators running inference workloads.

Submissions Across Multiple Systems

For MLPerf Inference v6.0, Supermicro submitted benchmarks on a variety of systems. Red Hat partnered with Supermicro on one of the submissions.

Supermicro and NVIDIA continue to work together across speech, language, and vision AI workloads, with the goal of giving customers access to well-tested, efficient systems for inference.

For more information about NVIDIA and MLPerf results, please visit: https://developer.nvidia.com/blog/nvidia-extreme-co-design-delivers-new-mlperf-inference-records/

Subscribe to Data Center Stories

By clicking subscribe, you consent to allow Supermicro to store and process the personal information submitted above to provide you the content requested.

You can unsubscribe from these communications at any time. For more information on how to unsubscribe, our privacy practices, and how we are committed to protecting and respecting your privacy, please review our Privacy Policy.